Home icon

Three new models for speech recognition and text-to-speech are now available in Amazon SageMaker JumpStart

News



This article announces the availability of three new Qwen3 speech models in Amazon SageMaker JumpStart for speech recognition and text-to-speech capabilities.

  • Qwen3-TTS-12Hz-1.7B-CustomVoice: Multilingual TTS with customizable voice styles across 10 languages
  • Qwen3-TTS-12Hz-1.7B-Base: Multilingual TTS with 3-second rapid voice cloning from audio input
  • Qwen3-ASR-1.7B: Automatic speech recognition supporting 52 languages with state-of-the-art accuracy
  • Models enable real-time interactive voice applications, virtual assistants, and transcription services
  • Deploy models with few clicks using SageMaker Studio or Python SDK

These models expand SageMaker JumpStart's foundation model portfolio, enabling customers to build intelligent voice-powered applications on AWS infrastructure.



Go to article

The AWS News Feed is currently looking for gold sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.

Related articles

May 14
2026
New models for image generation and text embeddings are now available in Amazon SageMaker JumpStart
May 6
2026
4 new Qwen models for multimodal reasoning, agentic coding, and multilingual applications are now available in Amazon SageMaker JumpStart
May 14
2026
Two new models for agentic coding and efficient AI are now available in Amazon SageMaker JumpStart
Jul 13
2026
Voxtral-Mini-4B-Realtime for real-time speech transcription is now available in Amazon SageMaker JumpStart

The AWS News Feed is currently looking for silver sponsors. If you want to support the AWS community and reach a large audience of AWS professionals, consider sponsoring the AWS News Feed.