AWS deploys vLLM-Omni Deep Learning Containers on Amazon SageMaker AI to enable real-time image/video generation and streaming text-to-speech
Deployment Provisional 78% confidence first seen
AWS made the vLLM-Omni Deep Learning Container available on Amazon SageMaker AI, adding support for generating still images and animated videos from text prompts using SageMaker endpoints. The coverage also describes deploying a text-to-speech model (Qwen3-TTS) with vLLM-Omni to stream voice output in real time while the model generates it.