OpenAI's small, low-cost embedding model. Returns 1536-dim vectors by default and supports shortening output dimensions via the dimensions parameter without retraining. Replaced text-embedding-ada-002 with better retrieval quality at a fraction of the price, and is the default choice for general-purpose semantic search and RAG.
Models in this article
Introduction to OpenAI Text Embedding 3 Small (TE3 Small)
text-embedding-ada-002, this model is specifically engineered to balance high performance with extreme cost efficiency. In the world of Large Language Models (LLMs), embeddings are the backbone of semantic understanding, converting text into numerical vectors that capture the underlying meaning. Whether you are building a recommendation engine or a sophisticated Retrieval-Augmented Generation (RAG) pipeline, understanding the nuances of text-embedding-3-small is critical for optimizing both latency and accuracy.Technical Specifications and Dimensions
text-embedding-3-small model utilizes a default output of 1,536 dimensions. However, one of its most innovative features is the support for Matryoshka Representation Learning (MRL). This allows developers to truncate the vector dimensions (e.g., down to 512 or 256) while retaining a surprising amount of semantic information. This flexibility is vital for reducing storage costs in vector databases like Pinecone or Weaviate without necessitating a complete re-indexing of your data. The model supports a context window of 8,191 tokens, making it suitable for embedding everything from short queries to lengthy technical documentation.The Power of Matryoshka Embeddings
text-embedding-3-small, OpenAI has trained the model such that the most important information is 'packed' into the earlier dimensions of the vector. You can simply slice the 1,536-dimensional vector at 512 dimensions and still maintain roughly 98% of the performance on standard benchmarks. This allows for tiered search architectures where you perform a fast initial search on small vectors and a reranking step on the full dimensions.Benchmark Performance: MTEB and MIRACL
text-embedding-3-small consistently outperforms its predecessor, ada-002. On the Massive Text Embedding Benchmark (MTEB), which evaluates models across tasks like clustering, classification, and retrieval, TE3 Small achieved an average score of 62.3%, compared to 61.0% for the older model. While the jump might seem incremental, the real strength lies in its multilingual performance. On the MIRACL benchmark, which focuses on cross-language information retrieval, the model shows a double-digit percentage improvement in accuracy for non-English languages.- MTEB Retrieval Score: 51.6% (Significant improvement over Ada-002)
- MIRACL Multilingual Score: 44.0% average across 18 languages
- Maintains high accuracy even when truncated to 512 dimensions
- Optimized for zero-shot classification tasks
- Stronger performance in identifying technical jargon and code snippets
Pricing and Cost Efficiency Analysis
text-embedding-3-small as their most accessible model yet. The pricing is set at $0.02 per 1 million tokens. To put this into perspective, this is a 5x reduction in cost compared to text-embedding-ada-002, which was already considered affordable. For startups and enterprises processing billions of documents, these savings are transformative. It allows for more frequent re-indexing of content and more granular chunking strategies without breaking the budget. Combined with the storage savings from dimensionality reduction, the Total Cost of Ownership (TCO) for a RAG system using TE3 Small is remarkably low.- OpenAI text-embedding-3-smallUS$0.024 / 1M input tokens1K in + 500 out tokens: ≈ US$0.00010
- OpenAI text-embedding-3-largeUS$0.156 / 1M input tokens1K in + 500 out tokens: ≈ US$0.00020
Current prices from Railwail's rules, 7 October 2026. Billed in USD from a prepaid balance.
Ideal Use Cases for TE3 Small
Semantic Search and Document Retrieval
text-embedding-3-small is semantic search. Unlike keyword-based search (BM25), embeddings allow a system to understand that 'how to fix a flat' and 'tire repair instructions' are semantically identical. Because TE3 Small is so fast and cheap, it is the perfect 'first-pass' retriever in a complex search pipeline. You can embed your entire knowledge base and perform a cosine similarity search in milliseconds. If you're ready to build, sign up today to get your API key and start indexing.Clustering and Topic Modeling
Strengths and Limitations
text-embedding-3-small is a powerhouse, it is important to be data-driven about its limitations. Its greatest strength is its efficiency-to-performance ratio. It provides 'good enough' performance for 90% of business applications at a fraction of the cost of 'Large' models. However, for extremely high-stakes legal or medical document retrieval where every percentage point of accuracy matters, its bigger brother, text-embedding-3-large, or a domain-specific model might be preferable. Additionally, like all dense embeddings, it can occasionally struggle with very short, specific strings like serial numbers or highly unique acronyms that weren't prevalent in its training set.- Strength: Unmatched price-to-performance ratio.
- Strength: Flexible dimensionality (Matryoshka).
- Strength: Robust multilingual support.
- Limitation: Lower absolute accuracy than TE3 Large on complex reasoning tasks.
- Limitation: Fixed context window of 8k tokens (requires chunking for long books).
- Limitation: Closed-source nature means you rely on OpenAI's uptime.
Comparison with Competitors
embed-english-v3.0 and Voyage AI's specialized models. Cohere often performs slightly better on specific retrieval benchmarks but at a higher price point. Open-source models like BGE-Small-v1.5 are also popular for local deployment; however, they often require significant GPU resources to host at scale, whereas TE3 Small's API-based approach offers 'infinite' scalability with zero maintenance. For most developers, the integration ease of the OpenAI ecosystem makes TE3 Small the default choice.How to Implement Text Embedding 3 Small
dimensions. If you do not specify it, the model defaults to 1,536. If you want to use the Matryoshka feature to save space, you simply pass dimensions=512. It is recommended to use cosine similarity as your distance metric when performing searches, as the model was optimized for this during training. Remember to normalize your vectors if your vector database does not do it automatically, although OpenAI's API returns pre-normalized vectors by default.Conclusion: Is TE3 Small Right for You?
text-embedding-3-large exists for those who need the absolute ceiling of performance, TE3 Small is the 'workhorse' model that will power the next generation of AI-driven applications. Start building with it today and experience the efficiency of modern vector embeddings.Models in this article
Live prices from Railwail's current rules, 7 October 2026.
≈ billed by actual tokens or GPU time
Next step
Try OpenAI text-embedding-3-small on Railwail
Sign in with Google for 10 free credits (usable 24 hours after sign-up, runs up to 2 credits), or top up from US$5.00. Unused balance does not expire.