Amazon Titan Text Embeddings V2

GA

Text embedding model (2nd-gen Titan). Output is an embedding vector, not tokens, so there is no output-token price. Input limit 8,192 tokens (~50,000 chars); output dimensions flexible 256 / 512 / 1,024 (default 1,024). Price is the us-east-1 on-demand rate from the AWS Price List API ($0.00002/1K input; batch $0.01/M). Model card lifecycle: Active; 'Model EOL date: No sooner than 4/30/2024' is a floor, not a fixed retirement, so lifecycle fields stay null.

Amazon Titan Text Embeddings V2 by Amazon is tracked live in our LLM catalog, with a 8K (8192-token) context window. It is generally available (GA).

Last verified: 13 Aug 2026 · sourced from official provider documentation

Provider
Amazon
Status
GA
Input price
$0.02 / 1M tokens
Output price
Cached input
Blended price
Context window
8192 tokens (8K)
Max output
Embedding dimensions
1024-dim output vectors
Modality
text
Open weights
No — API only
Knowledge cutoff
Released
30 Apr 2024
API string
amazon.titan-embed-text-v2:0

Source: Amazon official documentation ↗

FAQ

Amazon Titan Text Embeddings V2 — questions & answers

What is the context window of Amazon Titan Text Embeddings V2?

Amazon Titan Text Embeddings V2 has a 8192-token context window (8K).

Is Amazon Titan Text Embeddings V2 deprecated?

No — Amazon Titan Text Embeddings V2 is generally available (GA) and not currently scheduled for deprecation or retirement in our tracker.

Does Amazon Titan Text Embeddings V2 support image or vision input?

Our catalog lists Amazon Titan Text Embeddings V2's modalities as text; image/vision input is not among them.

Can I self-host Amazon Titan Text Embeddings V2?

No — Amazon Titan Text Embeddings V2 is proprietary. Its weights are not publicly released, so it is available only through Amazon's API (or hosted partners), not for self-hosting.

What is the API model string for Amazon Titan Text Embeddings V2?

The API model identifier for Amazon Titan Text Embeddings V2 is "amazon.titan-embed-text-v2:0" when calling Amazon's API.