Home  /  Models  /  Text Embedding 3 (Small)
⚡ Quick Pricing Summary

Text Embedding 3 (Small) by OpenAI costs $0.020 per 1 Million input tokens and N/A (Vectors) per 1 Million output tokens. It features a 8k context window and has a blended 3:1 production rate of $0.015/1M tokens.

OpenAI Embedding

Text Embedding 3 (Small)

High-efficiency 1536-dimension embedding model with native dimension reduction.

Calculate Spend in App →
Input Token Price
$0.020
Per 1,000,000 tokens
Output Token Price
N/A (Vectors)
Per 1,000,000 tokens
Prompt Caching Rate
Not Supported
Save up to 80% on cached inputs
Max Context Window
8k
Dense Vector Output

Verified Specifications & Benchmark Data

Developer / Provider OpenAI
Model Family Embeddings
Blended 3:1 Rate (Production Benchmark) $0.015 / 1M tokens
Batch API Discount (24hr SLA) 50% off standard rate
Multimodal Vision ❌ Text Only
Function Calling / Structured Outputs ❌ Not Available
MMLU Benchmark Score N/A (Specialized)
Latency & Throughput Tier Ultra-Fast
Optimal Architecture & Use Cases Vector search, semantic clustering, RAG index creation

Frequently Asked Questions about Text Embedding 3 (Small)

How much does Text Embedding 3 (Small) cost per 1M tokens?

Text Embedding 3 (Small) pricing is set at $0.020 per 1 million input tokens and N/A (Vectors) per 1 million output tokens. For high-volume batch workloads, 24-hour batch queues provide a 50% off standard rate.

How much money does Text Embedding 3 (Small) prompt caching save?

Prompt caching is currently not natively offered for Text Embedding 3 (Small) on direct serverless endpoints.

What is the context window limit of Text Embedding 3 (Small)?

Text Embedding 3 (Small) has a maximum context window of 8k (8,191 tokens), supporting up to 4,096 completion tokens per response.

Direct Matchups with Text Embedding 3 (Small)

See how Text Embedding 3 (Small) compares against other leading frontier and open-weights models in cost and latency.