iFANN
    iFANN 검색...
    로그인
    홈
    뉴스
    동영상
    사진
    GIF
    탐색
    투표
    어워드
    iFAMOUS
    위키
    애니
    룸
    알림
    메시지
    북마크
    프로필
    위키어워드iFAMOUS랭킹산업크리에이터 리워드사용자 리워드약관개인정보커뮤니티 가이드라인게시 중단 / DMCA도움말개발자

    © 2026 iFANN

    홈
    검색
    메시지
    알림
    프로필

    게시물

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 좋아요0 싫어요1 리포스트0 댓글
    ?

    댓글

    아직 댓글이 없습니다. 첫 댓글을 남겨보세요!

    게시물

    Nate
    Nate@nate_512
    🏢DeepSeek🏢Nvidia🏢Hugging Face

    DeepSeek V4.1 Flash NVFP4 locally

    LOCAL INFERENCE JUST GOT A MASSIVE UPGRADE atomic_chat_hq did it again 💥 They’ve made DeepSeek V4.1 Flash much more practical to run on Blackwell by squeezing the 552B MoE model into NVFP4. The result? lower inference costs without giving up the model’s core capabilities. You still get a 1M context window and native visual understanding, while retaining 98% agreement on agentic dialogue. That makes this a super strong fit for local agentic workflows. Grab the weights on huggingface 🤗 (link in 🧵↓)

    2d

    5 좋아요0 싫어요1 리포스트0 댓글
    ?

    댓글

    아직 댓글이 없습니다. 첫 댓글을 남겨보세요!