Dedicated embedding models are optimized to map text to vectors where semantic similarity = closeness — the retrieval engine behind search, RAG, clustering, and recommendations. Options range from API services (OpenAI, Cohere) to strong open models (BGE, E5). Key choices: dimension (storage/speed), domain fit, and matching the same model for indexing and querying. Good embeddings are half of good retrieval.