AWS provided deployment guidance and managed endpoints on Amazon SageMaker AI for WhisperX speaker-labeled transcription and Qwen3-TTS real-time personalized voice cloning
Feature update Provisional 66% confidence first seen
AWS published how to deploy the WhisperX speech-to-text container on Amazon SageMaker AI to generate transcripts with word-level timestamps and speaker (diarization) labels, including guidance on using synchronous versus asynchronous endpoints. Separately, Amazon SageMaker JumpStart added a publicly available Qwen3-TTS-12Hz-1.7B model deployment that supports real-time voice cloning through a managed inference endpoint.