Enterprise image generation and editing, product visualization, advertising and marketing assets, image variations, background removal, virtual try-on, and brand or subject-consistent visual content.
Type
Image Generation
Speed
7/10
Multimodal
Image input
Media output
Status
Legacy; currently accessible as of 2026-09-25; scheduled for end of life on 2026-09-30
View model
→
Low-cost multimodal document analysis, image and video understanding, visual question answering, summarization, RAG, and tool-enabled agents.
Type
Multimodal
Context
300K
Reasoning
5/10
Speed
9/10
Multimodal
Image input
Video input
Input
$0.06 per 1 million input tokens
Output
$0.24 per 1 million output tokens
View model
→
High-volume, low-latency text classification, summarization, translation, extraction, routing, FAQs and narrowly defined fine-tuned tasks
Type
Lightweight
Context
128K
Reasoning
4/10
Speed
10/10
Tool use
Fine-tuning
Streaming
Input
$0.035 per 1 million input tokens
Output
$0.14 per 1 million output tokens
View model
→
Long-context multimodal analysis, enterprise document workflows, complex tool calling, agentic orchestration, codebase analysis, and teacher-model distillation before retirement.
Type
Multimodal
Context
1M
Reasoning
8/10
Speed
7/10
Multimodal
Image input
Video input
View model
→
Enterprise multimodal applications, document analysis, visual question answering, video understanding, long-context summarization, RAG, and tool-using assistants
Type
Multimodal
Context
300K
Reasoning
7/10
Speed
8/10
Multimodal
Image input
Video input
Input
$0.80 per 1 million input tokens
Output
$3.20 per 1 million output tokens
View model
→
Short-form advertising, marketing concepts, product visualization, storyboards, social video drafts, and image-guided cinematic clips.
Multimodal
Image input
Media output
Status
Active through the current Nova Reel 1.1 workflow; the original Nova Reel v1.0 model is legacy and scheduled for end of life on 2026-09-30.
Input
Not token-priced; video generation is priced per generated video second.
Output
$0.08 per generated video second, subject to AWS region, pricing-tier, and current Bedrock pricing conditions.
View model
→
Real-time voice assistants, customer-service automation, interactive education, language learning, and speech-enabled enterprise workflows
Type
Multimodal
Context
300K
Reasoning
5/10
Speed
9/10
Multimodal
Audio input
Media output
Status
Retired; legacy model with official end-of-life date of September 14, 2026
View model
→
High-volume multimodal applications, document and video analysis, customer service, business automation, software engineering, long-context workflows, and cost-sensitive AI agents.
Type
Multimodal
Context
1M
Reasoning
7/10
Speed
8/10
Multimodal
Image input
Video input
Status
Active; generally available through Amazon Bedrock. AWS states EOL is no sooner than 2026-12-02.
Input
$0.30 per 1 million input tokens on the global standard rate; regional and service-tier prices may vary.
Output
$2.50 per 1 million output tokens on the global standard rate; regional and service-tier prices may vary.
View model
→
Real-time voice assistants, customer-service automation, telephony, interactive learning, multilingual conversations, and tool-enabled speech agents.
Type
Multimodal
Context
1M
Reasoning
6/10
Speed
9/10
Multimodal
Audio input
Media output
Status
Active; EOL no sooner than 2026-12-02
Input
$0.003 per 1,000 speech-input units; text-input pricing may apply separately and should be verified on the current Amazon Bedrock pricing page.
Output
$0.012 per 1,000 speech-output units; text-output pricing may apply separately and should be verified on the current Amazon Bedrock pricing page.
View model
→
Browser automation, visual UI navigation, repetitive web workflows, agentic QA, tool-oriented tasks, and human-supervised enterprise processes
Type
Other
Reasoning
7/10
Speed
7/10
Multimodal
Image input
Tool use
Status
Generally available
Input
$4.75 per agent hour; token-level input pricing is not published for this model
Output
$4.75 per agent hour; token-level output pricing is not published for this model
View model
→
Amazon Nova Multimodal Embeddings
Cross-modal semantic search, multimodal RAG, digital asset discovery, recommendations, classification, and clustering
Type
Other
Context
8K
Reasoning
1/10
Speed
8/10
Multimodal
Image input
Audio input
Status
Active; generally available through Amazon Bedrock
Input
Modality-dependent pricing; AWS charges based on processed input and processing mode. Current pricing should be checked on the Amazon Bedrock pricing page.
Output
Not applicable as output-token pricing; the model returns embeddings and pricing is based primarily on input modality and processing mode.
View model
→
Multimodal search, text-to-image retrieval, image similarity, visual recommendations, personalization, and image-text matching.
Type
Multimodal
Context
256
Reasoning
1/10
Speed
8/10
Multimodal
Image input
Fine-tuning
Input
Text: $0.0008 per 1,000 input tokens; images: $0.00006 per input image. Pricing may vary by AWS Region and current Bedrock pricing terms.
Output
No separate output-token price; the model returns embedding vectors.
View model
→
Text-to-image generation, image editing, reference-guided composition, background removal, color-controlled visuals, image variations, and subject-consistent branded content.
Type
Other
Reasoning
1/10
Speed
7/10
Multimodal
Image input
Media output
Status
Retired; AWS lists the model as legacy with an end-of-life date of 2026-06-30.
Input
Not token-priced; image-generation pricing applies. AWS pricing examples list $0.01 per 1,024×1,024 standard-quality image.
Output
$0.01 per 1,024×1,024 standard-quality image in the AWS pricing example; smaller images and premium quality use different pricing tiers.
View model
→
Semantic search, vector indexing, retrieval-augmented generation, personalization, clustering, classification, and recommendation pipelines.
Type
Other
Context
8K
Reasoning
1/10
Speed
8/10
Status
Available; older G1/V1 text-embedding generation
Input
$0.10 per 1 million input tokens
View model
→
Cost-efficient text embeddings for semantic search, RAG, document retrieval, classification, clustering, reranking, and recommendations
Type
Embedding
Context
8K
Reasoning
1/10
Speed
8/10
Input
$0.02 per 1 million input tokens
View model
→