iFANN
    iFANN durchsuchen...
    Anmelden
    Startseite
    Nachrichten
    Videos
    Fotos
    GIFs
    Entdecken
    Umfragen
    Auszeichnungen
    iFAMOUS
    Wiki
    Anime
    Räume
    Benachrichtigungen
    Nachrichten
    Lesezeichen
    Profil
    WikiAuszeichnungeniFAMOUSRanglistenBranchenCreator-BelohnungenBenutzerbelohnungenNutzungsbedingungenDatenschutzCommunity-RichtlinienEntfernung / DMCAHilfeEntwickler

    © 2026 iFANN

    Startseite
    Suche
    Nachrichten
    Alarme
    Profil

    Beitrag

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 Gefällt mir0 Gefällt mir nicht1 Reposts0 Kommentare
    ?

    Kommentare

    Noch keine Kommentare. Sei der Erste!

    Beitrag

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 Gefällt mir0 Gefällt mir nicht1 Reposts0 Kommentare
    ?

    Kommentare

    Noch keine Kommentare. Sei der Erste!