Amazon Nova Sonic Audio Generation AI Model Released, Can Process Speech in Real-Time

Amazon Nova Sonic AI model comes with a context window of 32,000 tokens.

Advertisement
Written by Akash Dutta, Edited by Siddharth Suvarna | Updated: 9 April 2025 11:59 IST
Highlights
  • The model offers up to eight minutes of speech generation per session
  • Amazon says Nova Sonic can understand the context behind the input speech
  • Currently, it only supports the English language in multiple accents

Amazon Nova Sonic can be accessed via the company’s Bedrock console via an API

Photo Credit: Amazon

Amazon introduced a new artificial intelligence (AI) model in its flagship Nova family of models on Tuesday. Dubbed Amazon Nova Sonic, it is a voice generation model capable of generating human-like speech. However, it is not a text-to-speech (TTS) tool; instead, it can process voice input in real time and respond to it. The Seattle-based tech giant says developers can use the model to build conversational AI chatbots and similar tools. Notably, the Amazon Nova Sonic AI model also supports functional calling and tool use, making it compatible with agentic application developments as well.

Amazon Nova Sonic Is Available As an API

In a blog post, the tech giant announced the release of the Amazon Nova Sonic. The company said traditional approaches to voice-enabled applications use a complex with multiple models such as text recognition, speech-to-text conversion, data processing, and TTS models. This often leads to an increase in latency, and failure in preserving linguistic context, the post added.

Advertisement

Amazon said its approach with the Nova Sonic model was to unify speech understanding and speech generation components. The AI model is said to be able to process data and generate speech in real time, giving it a conversation-like experience. This unified system also allows the model to better understand the pace and timbre of input speech to contextualise the intent of the user.

Additionally, the AI model can understand different speaking styles as well as separate masculine and feminine-sounding voices in different accents. It can also understand when a user misspeaks, mumbles, or pauses while speaking. Amazon says the model can pick up speech even in a noisy setting.

Advertisement

In response generation, the company claims the model can be more expressive and human-like, and can adjust its response style to match the context of the conversation. Currently, the AI model only supports the English language. Amazon said support for more languages will be added soon. The model supports a context window of 32,000 tokens for audio, with an additional window to handle longer conversations. It has a default session limit of eight minutes.

To use the Nova Sonic model, developers can head to Amazon Bedrock and find it under the model access option. It can also be accessed via a bidirectional streaming application programming interface (API) that can both process audio input and generate output.

 

Get your daily dose of tech news, reviews, and insights, in under 80 characters on Gadgets 360 Turbo. Connect with fellow tech lovers on our Forum. Follow us on X, Facebook, WhatsApp, Threads and Google News for instant updates. Catch all the action on our YouTube channel.

Advertisement

Related Stories

Popular Mobile Brands
  1. Android 17 Beta Rolls Out to Motorola Edge 60 Pro With These Features
  2. WhatsApp Testing Home Screen Voice Message Widget for Android Users: Report
  3. Honor X6e Launched With 7,500mAh Battery
  4. JBL Bar 1000MK2 Review: More Than Just an Audio System
  5. Here's How Much the Vivo S2 Could Cost in India: See Expected Features
  1. Redmi Turbo 6 Max Render Leaks Ahead of Launch: Expected Specifications
  2. Gemini Can Now Read, Summarise, and Draft Replies to Comments in Google Docs 
  3. Paytm Introduces Split Bills for Shared Group Expenses on Its App
  4. India's Telecom User Base Rises in June, Airtel Tops Subscriber Additions: TRAI
  5. Crypto Hacks Top $1 Billion in H1 2026 as Ethereum and Solana Lead Losses: Report
  6. 1inch Introduces Aqua to Unite DeFi Liquidity on 13 Chains
  7. Wuchang: Fallen Feathers Sequel Announced by 505 Games With Original Creator's New Studio Involved
  8. WhatsApp Reportedly Testing Home Screen Voice Message Widget for Android Users
  9. Samsung Galaxy S26 FE Charging Speed Upgrade Seems Unlikely as Phone Reportedly Appears on a Certification Site
  10. Google Launches Gemini-Powered Ask Google Pay in India With a Dedicated In-App AI Chatbot
Download Our Apps
Available in Hindi
© Copyright Red Pixels Ventures Limited 2026. All rights reserved.