Amazon Nova 2 Sonic
Amazon Web Services
- Released
- December 2, 2025
- Data date
- October 3, 2026
- Model class
- Access
Amazon Nova 2 Sonic is Bedrock’s speech-to-speech model for bidirectional real-time conversations. It accepts audio and text and returns both. Turn taking and expressive polyglot voices belong to the conversation model, while asynchronous tools run alongside the audio exchange.
AWS cites a one-million-token context and availability in US regions and Tokyo. Audio and text tokens are billed separately at the applicable Bedrock rate. It is a hosted Bedrock service rather than a local weights checkpoint.
Specifications and access
| Specification | Value and source |
|---|---|
| Model class | Speech-to-speech conversationSource |
| Access | Amazon Bedrock APISource |
| Input | Audio and textSource |
| Output | Audio and textSource |
| API model ID | amazon.nova-2-sonicSource |
| Transport | Bidirectional streaming in Amazon BedrockSource |
| Capabilities | Expressive polyglot voices, turn-taking control, and asynchronous toolsSource |
| Regions | US East, US West, and TokyoSource |
Page 1 of 2
Measurements without matching peer values
These measurements have no matching peer values under the same test conditions. Their original values and sources remain available here.
Big Bench Audio accuracy (Show measurement, test conditions, and source)
- Source value
- 87
- Score
- 87
- Metric
- accuracy
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 12; Big Bench Audio version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. Artificial Analysis evaluation of speech-to-speech models
- Context
- Artificial Analysis is named in Amazon's report; this record preserves Amazon's report rather than an independently retrieved leaderboard run.
BFCL accuracy (Show measurement, test conditions, and source)
- Source value
- 74.5
- Score
- 74.5
- Metric
- accuracy
- Unit
- %
- Benchmark version
- BFCL-v3 filtered snapshot, repository commit c67d246
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. single-turn, single-function, python-only subset; text prompts transformed with Google TTS; report explicitly cites BFCL state as of 2025-03-02
- Context
- The cited 2025-03-02 BFCL snapshot is a dataset state, not a Nova 2 Sonic run date.
ComplexFunction accuracy (Show measurement, test conditions, and source)
- Source value
- 65.2
- Score
- 65.2
- Metric
- accuracy
- Unit
- %
- Benchmark version
- ComplexFunction; version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. text prompts transformed with Google TTS
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
IFBench prompt-level accuracy (Show measurement, test conditions, and source)
- Source value
- 33.3
- Score
- 33.3
- Metric
- accuracy
- Unit
- %
- Benchmark version
- IFBench; version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. text prompts transformed with Google TTS
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
IFBench instruction-level accuracy (Show measurement, test conditions, and source)
- Source value
- 37.5
- Score
- 37.5
- Metric
- accuracy
- Unit
- %
- Benchmark version
- IFBench; version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. text prompts transformed with Google TTS
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, American English masculine, versus GPT-Realtime (Show measurement, test conditions, and source)
- Source value
- 53.9
- Score
- 53.9
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, American English masculine, versus Gemini 2.5 Flash Live (Show measurement, test conditions, and source)
- Source value
- 60
- Score
- 60
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, Indian English masculine, versus GPT-Realtime (Show measurement, test conditions, and source)
- Source value
- 55.5
- Score
- 55.5
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, Indian English masculine, versus Gemini 2.5 Flash Live (Show measurement, test conditions, and source)
- Source value
- 64.2
- Score
- 64.2
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, Spanish masculine, versus GPT-Realtime (Show measurement, test conditions, and source)
- Source value
- 68.4
- Score
- 68.4
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, Spanish masculine, versus Gemini 2.5 Flash Live (Show measurement, test conditions, and source)
- Source value
- 70.3
- Score
- 70.3
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Human preference, French masculine, versus GPT-Realtime (Show measurement, test conditions, and source)
- Source value
- 54.7
- Score
- 54.7
- Metric
- win rate
- Unit
- %
- Benchmark version
- Nova 2 technical report Table 11; preference test version not stated
- Category
- audio
- Direction
- Higher is better
- Source type
- vendor-reported
- Evaluator
- Amazon Nova 2 technical report
- Status
- active
- Retrieved at
- 2026-10-04
- Methodology
- Vendor-reported result. human evaluation; rows are winning rates for Nova 2 Sonic against GPT-Realtime / Gemini 2.5 Flash Live
- Context
- Source-specific observation; it does not establish a compatible cross-model cohort.
Page 1 of 2
Published prices
Prices apply to the stated unit. Resolution, output length, and provider can change the cost.
- Billing
- Separate token rates for audio and text input and audio and text output; text rates also apply to transcription, tool calls, grounding, and history. Standard, Priority, and Flex tiers must be checked separately.Source
Sources and data date
Every statement links to its underlying documentation or leaderboard.
| Type | Evidence and data date |
|---|---|
| Research status | Research date October 4, 2026. 1 source URLs checked. This documents the inspected sources, not an exhaustive inventory of every publication. |
| Additional source | AWS Amazon Nova 2 Sonic announcement (retrieved October 3, 2026) · AWS Amazon Nova 2 Sonic announcement · Editorial description reviewed October 3, 2026 |
| Additional source | AWS Nova pricing (retrieved October 3, 2026) · AWS Nova pricing · Editorial description reviewed October 3, 2026 |
| Additional source | Amazon Nova 2 technical report (retrieved October 4, 2026) |