TLDRocket
Sign in

Together AI

54 summarised stories about Together AI, each linking back to the original source. Browse all topics →

+ Follow this topic

Wednesday, 15 July 2026

New in Together GPU Clusters: Reliability and control for production GPU clusters

Together AI 1 month ago 15

Together AI released updates to its GPU Clusters platform including passive health checks that monitor running workloads for failures like GPU bus drops and thermal throttling, auto node repair with human approval, and a rebuilt Slurm-on-Kubernetes stack addressing daemon crashes and process cleanup. The platform added operational features including a redesigned cluster overview dashboard showing health and utilization, external OIDC authentication for per-user Kubernetes access, and startup scripts for self-serve node customization. These changes reduce incident resolution time from hours to minutes and enable teams to manage clusters at scale without sharing admin credentials or performing manual node setup.

Together AI brings Thinking Machines Lab’s new model Inkling on day 0

Together AI 1 month ago 50 2 sources

Thinking Machines Lab released Inkling, a 975-billion-parameter mixture-of-experts model with 40B active parameters that accepts text, image, and audio inputs for multimodal reasoning tasks. Inkling is available on Together AI's inference platform starting today with a 1M token context window and adjustable inference effort settings. Developers can now access a unified multimodal model through a single API endpoint that supports reasoning, coding, forecasting, and agentic workflows without managing their own infrastructure.

The daily briefing

Every AI story that matters, in your inbox by 8am.

TLDRocket reads all relevant sources, removes duplicate coverage, and summarises the day in two minutes. Follow companies and topics for alerts, or get the briefing in Slack. Free, no spam, unsubscribe anytime.