Post

Log inSign up

Post

atomic.chat on X: "@AntLingAGI Explore Atomic's Ling 3.0 Flash quants: https://t.co/ksDZBmJMrm Run AI Models locally: https://t.co/8dXwcTiJbo"

  • user avatar
    atomic.chat
    @atomic_chat_hq
    Aug 5
    Run Ling 3.0 Flash locally 🌀 We released GGUF quants on Hugging Face, from lossless BF16 to 1-bit, plus NVFP4! AD-Q5_K_M is the best fit for 128GB hardware (tested on DGX Spark). It matches the original's token choice 97.5% of the time and drifts 31% less than the llama.cpp
    user avatar
    Ant Ling
    @AntLingAGI
    Aug 5
    🚀 Today, we’re releasing INT4 and FP4 (MXFP4) variants of Ling-3.0-flash. Both run end to end on a single NVIDIA DGX Spark via our Spark-adapted SGLang path. For FP4, W4A16 is the stable default, while W4A8 is tuned for higher throughput. The efficiency and accuracy of the
  • user avatar
    atomic.chat
    @atomic_chat_hq
    Explore Atomic's Ling 3.0 Flash quants: huggingface.co/collections/At… Run AI Models locally: atomic.chat
    Ling 3.0 Flash - a AtomicChat Collection
    From huggingface.co
    4:21 PM · Aug 5, 20262.5KViews

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email

Relevant people

Avatar
atomic.chat@atomic_chat_hqFollow
Local AI chat and Inference Engine. Enhanced by TurboQuant. Team: @gladkos @skinbagwbones @AlexFromAtomic @danyurkin @worthant_

Trending now

Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.