TLDRocket
Sign in

Tools & Coding

975 summarised stories in Tools & Coding, each linking back to the original source. Browse all topics →

Saturday, 3 February 2024

SegMoE: Segmind Mixture of Diffusion Experts

Hugging Face 2 years ago 13

Segmind released SegMoE, a framework that creates Mixture-of-Experts diffusion models by combining multiple expert models through a router network that selectively activates them during inference. Three pre-built models are available on Hugging Face (SegMoE-2x1, SegMoE-4x2, and SegMoE-SD-4x2), with SegMoE-4x2 requiring 24GB of VRAM in half-precision. Users can now create custom MoE diffusion models by configuring a YAML file specifying base and expert models, though inference speed decreases when multiple experts process tokens simultaneously.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.