Key research and product announcements at the AI Native Conf
Together AI
Together Research announced seven research and product releases at AI Native Conf, including FlashAttention-4, a Reinforcement Learning API, ThunderAgent for agentic workflows, and optimization techniques like ATLAS-2. FlashAttention-4 achieves 2.7x faster performance than Triton on NVIDIA Blackwell GPUs, while Together Megakernel reduced latency from 281ms to 77ms for voice agents and together.compile improved image generation speed by 41%. These advances integrate research directly into production infrastructure to optimize AI model inference and training workloads at scale.
Why it matters
At AI Native Conf, Together AI announced breakthroughs across kernels, RL, and inference optimization — including FlashAttention-4, ThunderAgent, and together.compile. Research that ships to production. That's the AI Native Cloud.