Post

Log inSign up

Post

Cursor on X: "We're open-sourcing Mixture-of-Kittens (MoK), our MoE training megakernel for NVL72s. It fuses all Mixture-of-Experts communication and computation into a single, fully deterministic kernel, and runs up to 2.37x faster than the strongest public baselines. https://t.co/yHu5E6RXp9"

  • user avatar
    Cursor
    @cursor_ai
    We're open-sourcing Mixture-of-Kittens (MoK), our MoE training megakernel for NVL72s. It fuses all Mixture-of-Experts communication and computation into a single, fully deterministic kernel, and runs up to 2.37x faster than the strongest public baselines.
    Bar chart: Mixture-of-Kittens reaches up to 2.37× higher MXFP8 forward throughput than the fastest public baseline on GB300 NVL72s, across Kimi K2.7, GLM 5.2, Qwen 3.5-397B-A17B, and DeepSeek V4 Pro.
    4:00 PM · Aug 4, 2026541.3KViews
  • user avatar
    Cursor
    @cursor_ai
    Aug 4
    MoK now powers training across tens of thousands of GPUs at Cursor. In production, it raised end-to-end training throughput by 1.41x over our previous DeepEP-based stack.
    user avatar
    Cursor
    @cursor_ai
    Aug 4
    Our hope is that this lowers the barrier to AI research, so more labs can train models efficiently. Here's how we built it:
    Mixture-of-Kittens: our open-source MoE megakernel for NVL72s · Cursor
    From cursor.com
  • user avatar
    Lee Robinson
    Cursor
    @leerob
    Aug 4
    The real mixture of kittens

Log in or sign up for X

See what’s happening and join the conversation

Continue with phone
or
Log in with username or email

Relevant people

Avatar
Cursor@cursor_aiFollow
Coding agent for building ambitious software

Trending now

Terms·Privacy·Cookies·Accessibility·Ads Info·© 2026 X Corp.