TLDRocket
Sign in

Mixture-of-Experts

43 summarised stories about Mixture-of-Experts, each linking back to the original source. Browse all topics →

+ Follow this topic

Tuesday, 4 August 2026

Cursor Open-Sources Mixture-of-Kittens (MoK): A Deterministic MoE Training Megakernel for GB300 NVL72 Racks

MarkTechPost 3 weeks ago 44 2 sources

Cursor Research open-sourced Mixture-of-Kittens, a mixture-of-experts training kernel that fuses all MoE communication and computation into a single deterministic megakernel for large GPU clusters.The kernel achieves up to 2.37x higher throughput than existing baselines and requires NVIDIA Blackwell GPUs in GB300 NVL72 racks with Python 3.12+, PyTorch 2.10+, and CUDA 13.0+.Organizations with access to large-scale GPU infrastructure can now use MoK under Apache-2.0 to accelerate training of mixture-of-experts models like DeepSeek-V3-style architectures.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.