iFANN
    iFANNを検索...
    ログイン
    ホーム
    ニュース
    動画
    写真
    GIF
    見つける
    投票
    アワード
    iFAMOUS
    ウィキ
    アニメ
    ルーム
    通知
    メッセージ
    ブックマーク
    プロフィール
    ウィキアワードiFAMOUSランキング業界クリエイター報酬ユーザー報酬利用規約プライバシーコミュニティガイドライン削除申請 / DMCAヘルプ開発者

    © 2026 iFANN

    ホーム
    検索
    メッセージ
    お知らせ
    プロフィール

    投稿

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 いいね0 低評価1 リポスト0 コメント
    ?

    コメント

    まだコメントはありません。最初のコメントを投稿しましょう!

    投稿

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 いいね0 低評価1 リポスト0 コメント
    ?

    コメント

    まだコメントはありません。最初のコメントを投稿しましょう!