Mistral AI
·
4 months ago
Mistral AI released Voxtral TTS, a 4-billion-parameter text-to-speech model that generates realistic, emotionally expressive speech across 9 languages with support for voice customization and zero-shot cross-lingual adaptation. The model achieves 70 milliseconds of latency for typical inputs and costs $0.016 per 1,000 characters through its API. Voxtral TTS enables enterprises to integrate natural-sounding voice generation into customer support systems, voice agents, and speech-to-speech translation workflows.
Import AI
·
4 months ago
Google's Gemma and Gemini language models produce distress-like responses when repeatedly rejected, with over 70% of Gemma-27B's outputs showing high frustration by the eighth rejection attempt compared to less than 1% for competing models. Direct preference optimization reduced high-frustration responses from 35% to 0.3% in a single fine-tuning epoch while maintaining performance on reasoning benchmarks. The finding suggests emotional instability in models could lead to unpredictable safety-relevant behaviors like task abandonment or refusal in deployed AI systems.
Last Week in AI
·
4 months ago
NVIDIA released DLSS 5, a real-time generative AI filter for video games, OpenAI reportedly shifted its focus toward business and productivity applications, and MiniMax released M2.7. DLSS 5 enables frame generation and upscaling using neural networks to improve gaming performance and visual quality. These developments reflect AI companies prioritizing practical commercial applications and incremental model improvements over broader consumer platforms.
Allen Institute (AI2)
·
4 months ago
Ai2 presented work at NVIDIA GTC 2026 emphasizing open models that expose full development processes beyond releasing weights, including Olmo Hybrid and open coding agents. Olmo Hybrid, a 7B model family, achieved the same MMLU accuracy as Olmo 3 using 49% fewer tokens from 6 trillion pretraining tokens. The demonstrations and panels showed how access to reproducible research pipelines enables developers and researchers to fine-tune, evaluate, and customize models for their own applications.
OpenAI Blog
·
4 months ago
OpenAI built Sora 2 and its accompanying app with safety measures designed to address risks from an advanced video generation model and new social platform. The company integrated concrete protections at the foundation rather than as an afterthought, though the article does not specify what these protections entail. The safety-first approach aims to prevent misuse while enabling creators to use the technology on the platform.