Titan Text Embeddings

Amazon Titan Embeddings G1 - Text

by Amazon · Available; older G1/V1 text-embedding generation

Amazon Titan Embeddings G1 - Text is a text-only Amazon Bedrock embedding model that converts up to 8,192 input tokens into a fixed 1,536-dimensional vector. It is designed for semantic search, RAG indexing, retrieval, personalization, clustering, classification, and recommendation systems, with pricing listed at $0.10 per 1 million input tokens.

Embeddings Reasoning Coding
Amazon Titan Embeddings G1 - Text, also called Titan Text Embeddings V1, turns text into numerical representations that applications can compare and search. It accepts up to 8,192 tokens and returns a 1,536-dimensional embedding through Amazon Bedrock. Its main value is efficient semantic matching: finding passages with related meaning even when they do not share the same keywords.
Outputs

What Amazon Titan Embeddings G1 - Text can produce

Embeddings
Inputs

What it can understand

Text
Model profile

Performance characteristics

1/10 Reasoning
1/10 Coding
8/10 Speed
8/10 Cost efficiency
Specifications

Technical details

Model family Titan Text Embeddings
Model type Other
Context window 8K tokens
Release date 2023-09-28
Status Available; older G1/V1 text-embedding generation
Knowledge cutoff notes

AWS does not publish a direct knowledge-cutoff date for this embedding model. Embedding models are used to transform supplied input text and should not be treated as general-purpose language models with a conventional public knowledge cutoff.

Model notes

Canonical Amazon Bedrock model ID is amazon.titan-embed-text-v1. AWS also refers to this model as Titan Text Embeddings V1. It accepts a non-empty text string of up to 8,192 tokens and returns one 1,536-dimensional floating-point embedding plus the input token count. The G1 request supports only inputText and does not support generative inference parameters such as maxTokenCount or topP. AWS documentation lists the model as available in selected Bedrock regions, including ap-northeast-1, eu-central-1, us-east-1, and us-west-2 for Bedrock Knowledge Bases. The official AWS launch material lists pricing of $0.10 per million input tokens; regional, service-tier, and account-specific pricing conditions may apply. Titan Text Embeddings V2 is the newer successor with configurable output dimensions, but it is a distinct model and is not an alias for G1.

Cost

Model pricing

Input $0.10 per 1 million input tokens
Model guide

Amazon Titan Embeddings G1 - Text for Semantic Search and RAG

Amazon Titan Embeddings G1 - Text is an Amazon Bedrock embedding model that converts text into fixed 1,536-dimensional floating-point vectors. It is designed for semantic search, retrieval-augmented generation, recommendation, personalization, clustering, classification, and vector indexing rather than conversational text generation.

What Amazon Titan Embeddings G1 - Text does

Amazon Titan Embeddings G1 - Text is an embedding-only model provided by Amazon through Amazon Bedrock. Instead of producing paragraphs, answers, code, images, audio, or video, it transforms supplied text into a numerical vector. That vector is a compact representation of the text's meaning and can be stored in a vector database or another search index.

For example, a knowledge base might contain a passage about resetting an account password. A user could search for “I cannot log in,” even if those exact words do not appear in the passage. The application can embed both the stored passage and the query, compare their vectors, and retrieve the text with the closest semantic relationship.

The model is therefore a building block for search and machine-learning systems, not a complete chatbot. Retrieval, similarity calculations, ranking, filtering, and response generation must be handled by the surrounding application or by other services.

Position in Amazon's model lineup

The canonical Bedrock model ID is amazon.titan-embed-text-v1. AWS also refers to it as Titan Text Embeddings V1. AWS documentation lists its launch date as September 28, 2023, and the model remains available in selected Amazon Bedrock regions and account configurations.

G1 is an earlier generation in Amazon's Titan text-embedding family. The newer Titan Text Embeddings V2 model is a separate successor, not an alias for G1. V2 supports configurable output dimensions and additional embedding options, while G1 returns a fixed-size vector. That distinction matters when choosing a model for a new index or when integrating with an existing vector database whose dimensions are already fixed.

Verified technical specifications

SpecificationAmazon Titan Embeddings G1 - Text
ProviderAmazon
Bedrock model IDamazon.titan-embed-text-v1
Model familyTitan Text Embeddings
Input modalityText only
Maximum input8,192 tokens
OutputOne 1,536-dimensional floating-point embedding and the input token count
Text generationNot supported
Inference parametersNot supported
AccessAmazon Bedrock Runtime InvokeModel API
Published price$0.10 per 1 million input tokens

The model accepts a non-empty text string and returns a JSON response containing an embedding field and an inputTextTokenCount field. It does not have a maximum output-token setting because its output is a fixed embedding rather than generated text. The supplied documentation does not identify a conventional knowledge-cutoff date; the model processes the text provided by the application.

How the embedding output is used

An embedding is useful because numerical vectors can be compared mathematically. A typical workflow first divides documents into logical passages, sends each passage to Titan Embeddings G1 - Text, and stores the returned vectors alongside the original text and metadata. When a user submits a search query, the application embeds that query using the same model and searches for nearby vectors.

Common downstream operations include:

  • Semantic search: retrieve content by meaning rather than exact keyword matches.
  • Retrieval-augmented generation: find relevant passages before passing them to a separate text-generation model.
  • Document retrieval: locate related paragraphs, policies, support articles, or records.
  • Personalization and recommendation: compare user interests, items, or content representations.
  • Clustering: group related documents or messages without manually assigning every category.
  • Classification: use vector representations as input to a downstream classifier.
  • Indexing: prepare content for vector databases and knowledge-base systems.

Long documents should generally be split into meaningful sections before embedding. Sending an entire large document as one vector can blur distinct topics and make retrieval less precise, even when the document fits within the 8,192-token input limit.

Strengths and practical trade-offs

The model's clearest strength is specialization. It is narrowly focused on converting text into embeddings, so applications do not pay for conversational generation when they only need indexing or semantic comparison. The documented price of $0.10 per 1 million input tokens is suited to processing collections of text, although actual costs depend on token volume, region, and the surrounding Bedrock or database services.

Its fixed 1,536-dimensional output also creates predictability. Teams can design an index around one known vector size and use the same dimensionality for documents and queries. The trade-off is reduced flexibility compared with Titan Text Embeddings V2, whose configurable dimensions may help applications balance storage, search performance, and compatibility requirements.

The supplied evaluation rates the model highly for speed and cost relative to the evaluated alternatives. Those are editorial assessments, not AWS-published benchmark results. The practical reason for the favorable cost and speed profile is that G1 performs a focused embedding task rather than extended reasoning or text generation. It should not, however, be described as a reasoning model.

Capabilities it does not provide

Amazon Titan Embeddings G1 - Text supports text input and embedding output only. It does not accept images, audio, or video according to the supplied specifications, and it does not generate text, images, speech, music, or video. It also does not provide native tool or function calling, web search, streaming generation, batch API support, or fine-tuning in the researched specification.

There are no generative inference controls such as maxTokenCount or topP. These parameters would be relevant to a text-generation model, but they do not apply to G1's fixed embedding response. Similarly, it has no conversational reasoning or coding capability. It can help a software system retrieve code documentation or classify source-code text, but it does not write, execute, debug, or reason through code as a coding model would.

The model also does not perform retrieval by itself. A Bedrock call returns the vector, while the application must store it, compare it, apply filters, and decide what to show to a user. A separate generation model is needed if the final experience must answer questions in natural language.

Pricing and availability

AWS launch material and the supplied research list pricing of $0.10 per 1 million input tokens. This is token-based inference pricing rather than a monthly subscription. No separate output-token price is listed because the model produces an embedding rather than generated text. Regional, service-tier, and account-specific pricing conditions may apply, so the current AWS pricing documentation should be checked before deployment.

Access is through Amazon Bedrock, specifically the Bedrock Runtime InvokeModel API. Availability is regional. The researched AWS materials identify selected regions, including ap-northeast-1, eu-central-1, us-east-1, and us-west-2 for Bedrock Knowledge Bases, but availability can vary by region and account configuration.

When to choose this model

Choose Amazon Titan Embeddings G1 - Text when the central requirement is text-to-vector conversion and the surrounding system can manage storage and retrieval. It is a reasonable fit for:

  • a semantic search index for support documentation or internal policies;
  • a retrieval layer for a question-answering or RAG application;
  • large-scale document ingestion where predictable output dimensions are useful;
  • recommendation, personalization, clustering, or classification pipelines based on text similarity;
  • an existing Bedrock or vector-database architecture built around 1,536-dimensional embeddings.

Another option may be more appropriate when the application needs configurable embedding dimensions or newer embedding features; Titan Text Embeddings V2 is the relevant named alternative in the supplied research. A text-generation model is more appropriate when the system must answer questions, summarize retrieved passages, write content, or generate code. A multimodal embedding model is needed when images, audio, or video must be represented alongside text.

Bottom line

Amazon Titan Embeddings G1 - Text is best understood as a focused infrastructure component for semantic text matching. Its defining characteristics are the amazon.titan-embed-text-v1 model ID, an 8,192-token input limit, a fixed 1,536-dimensional floating-point output, and token-based Bedrock pricing. It is useful when an application needs consistent, relatively low-cost text embeddings, but it is not a general-purpose AI assistant and cannot replace the search, database, retrieval, or generation layers around it.


Answers to Frequently Asked Questions

How does Titan Embeddings G1 compare with Titan Text Embeddings V2?
Titan Text Embeddings V2 is a separate, newer successor rather than an alias for G1. G1 produces a fixed 1,536-dimensional vector, while V2 supports configurable output dimensions and additional embedding options.
Does Amazon Titan Embeddings G1 - Text generate answers or perform searches by itself?
No. It only creates text embeddings through the Amazon Bedrock Runtime InvokeModel API. The surrounding application must store and compare vectors, retrieve relevant content, apply filters, and use a separate generation model if natural-language answers are required.
What is the Amazon Titan Embeddings G1 - Text model ID and vector dimension?
The Amazon Bedrock model ID is amazon.titan-embed-text-v1. It returns one fixed-size 1,536-dimensional floating-point embedding for each input text.
What are the input limits and pricing for Amazon Titan Embeddings G1 - Text?
The model accepts up to 8,192 tokens of text and is listed at $0.10 per 1 million input tokens. Actual costs may vary by AWS region, account, service tier, and related Bedrock or database services.
What is Amazon Titan Embeddings G1 - Text used for?
Amazon Titan Embeddings G1 - Text converts text into numerical vectors that represent meaning. These embeddings can power semantic search, retrieval-augmented generation, document retrieval, recommendations, clustering, classification, and vector database indexing.


Sources 5
Provider

About Amazon