Search

Amazon Nova Sonic Audio Generation AI Model Released, Can Process Speech in Real-Time

Amazon Nova Sonic AI model comes with a context window of 32,000 tokens.

Advertisement
Highlights
  • The model offers up to eight minutes of speech generation per session
  • Amazon says Nova Sonic can understand the context behind the input speech
  • Currently, it only supports the English language in multiple accents
Amazon Nova Sonic Audio Generation AI Model Released, Can Process Speech in Real-Time

Amazon Nova Sonic can be accessed via the company’s Bedrock console via an API

Photo Credit: Amazon

Amazon introduced a new artificial intelligence (AI) model in its flagship Nova family of models on Tuesday. Dubbed Amazon Nova Sonic, it is a voice generation model capable of generating human-like speech. However, it is not a text-to-speech (TTS) tool; instead, it can process voice input in real time and respond to it. The Seattle-based tech giant says developers can use the model to build conversational AI chatbots and similar tools. Notably, the Amazon Nova Sonic AI model also supports functional calling and tool use, making it compatible with agentic application developments as well.

Amazon Nova Sonic Is Available As an API

In a blog post, the tech giant announced the release of the Amazon Nova Sonic. The company said traditional approaches to voice-enabled applications use a complex with multiple models such as text recognition, speech-to-text conversion, data processing, and TTS models. This often leads to an increase in latency, and failure in preserving linguistic context, the post added.

Amazon said its approach with the Nova Sonic model was to unify speech understanding and speech generation components. The AI model is said to be able to process data and generate speech in real time, giving it a conversation-like experience. This unified system also allows the model to better understand the pace and timbre of input speech to contextualise the intent of the user.

Additionally, the AI model can understand different speaking styles as well as separate masculine and feminine-sounding voices in different accents. It can also understand when a user misspeaks, mumbles, or pauses while speaking. Amazon says the model can pick up speech even in a noisy setting.

In response generation, the company claims the model can be more expressive and human-like, and can adjust its response style to match the context of the conversation. Currently, the AI model only supports the English language. Amazon said support for more languages will be added soon. The model supports a context window of 32,000 tokens for audio, with an additional window to handle longer conversations. It has a default session limit of eight minutes.

To use the Nova Sonic model, developers can head to Amazon Bedrock and find it under the model access option. It can also be accessed via a bidirectional streaming application programming interface (API) that can both process audio input and generate output.

For the latest tech news and reviews, follow Gadgets 360 on X, Facebook, WhatsApp, Threads and Google News. For the latest videos on gadgets and tech, subscribe to our YouTube channel. If you want to know everything about top influencers, follow our in-house Who'sThat360 on Instagram and YouTube.

 
Show Full Article
Please wait...
Advertisement

Related Stories

Popular Mobile Brands
  1. Poco F7 5G With 7,550mAh Battery Launched in India: See Price
  2. OnePlus Nord 5 Camera Details Revealed Ahead of India Launch
  3. JBL Tune Beam 2 Review: Punchy Sound Meets Powerful ANC
  4. Amazon Prime Day 2025 Sale Dates Announced: Check Upcoming Discounts
  5. Realme GT 8 Pro Said to Get an Anti-Glare 2K Resolution Display
  6. Vivo T4 Lite 5G With 6,000mAh Battery Launched in India: See Price
  7. Samsung Galaxy Z Fold 7 and Galaxy Z Flip 7 Prices Leaked Ahead of Launch
  8. Tecno Spark Go 2 With Free Link App Support Launched in India: See Price
  1. Lenovo Chromebook Plus With MediaTek Kompanio Ultra 910, Google AI Features and Dolby Atmos Launched
  2. Google Earth Gets Upgraded With Historical Street View Imagery, AI-Driven Insights to Arrive Soon
  3. Xiaomi Mix Flip 2 Design Teased; to Feature Snapdragon 8 Elite SoC, 5,165mAh Battery
  4. CD Projekt Red Delays Cyberpunk 2077 Update 2.3, Says Patch Will Similar in Scope to Previous One
  5. Oppo Pad SE India Launch Timeline Tipped; Could Launch Alongside Reno 14 Series
  6. UK May Compel Google to Change Search Rankings, Offer Alternatives
  7. Samsung Opens Pre-Reservations for Upcoming Galaxy Z Foldables in India
  8. Asus ROG Strix G16, TUF Gaming F16 Refreshed with Nvidia GeForce RTX 5050 GPUs: Price, Specifications
  9. Poco F7 5G With Snapdragon 8s Gen 4 SoC, 7,550mAh Battery Launched in India: Price, Specifications
  10. Google Expands AI Mode in Search to India, Adds Support for Voice and Image Inputs
Gadgets 360 is available in
Download Our Apps
App Store App Store
Available in Hindi
App Store
© Copyright Red Pixels Ventures Limited 2025. All rights reserved.
Trending Products »
Latest Tech News »