iFANN
    Cerca su iFANN...
    Accedi
    Home
    Notizie
    Video
    Foto
    GIF
    Esplora
    Sondaggi
    Premi
    iFAMOUS
    Wiki
    Anime
    Stanze
    Notifiche
    Messaggi
    Segnalibri
    Profilo
    WikiPremiiFAMOUSClassificheSettoriRicompense CreatorRicompense UtenteTerminiPrivacyLinee guida della communityRimozione / DMCAAiutoSviluppatori

    © 2026 iFANN

    Home
    Cerca
    Messaggi
    Avvisi
    Profilo

    Pubblica

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 Mi piace0 Non mi piace1 Repliche0 Commenti
    ?

    Commenti

    Ancora nessun commento. Sii il primo!

    Pubblica

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 Mi piace0 Non mi piace1 Repliche0 Commenti
    ?

    Commenti

    Ancora nessun commento. Sii il primo!