{"slug":"voyage-code-3","name":"Voyage Code 3","domain":null,"verdict":"As of 2026-08-10, ChatGPT, Claude, Gemini, Grok collectively rank Voyage Code 3 first for code embedding apis for repository search. Source: https://modelsagree.com/product/voyage-code-3 (modelsagree.com, CC BY 4.0).","best_rank":1,"categories":1,"entries":[{"slug":"best-code-embedding-apis-for-repository-search","title":"Best code embedding APIs for repository search","rank":1,"of":11,"score":19,"appearances":4,"modelRanks":{"ChatGPT":2,"Claude":1,"Gemini":1,"Grok":1},"reason":"Purpose-built for code retrieval and consistently at or near the top of code-search benchmarks; long context, Matryoshka-style adjustable dimensions and int8/binary quantization let you trade recall for index cost, which matters at whole-repo scale; mature reranker pairing (rerank-2) and clean API. Near-tie with Codestral Embed at the top.","reasons":[{"model":"Claude","reason":"Purpose-built for code retrieval and consistently at or near the top of code-search benchmarks; long context, Matryoshka-style adjustable dimensions and int8/binary quantization let you trade recall for index cost, which matters at whole-repo scale; mature reranker pairing (rerank-2) and clean API. Near-tie with Codestral Embed at the top."},{"model":"Gemini","reason":"Top-tier code retrieval accuracy specifically fine-tuned for syntax, function call graphs, and technical documentation, featuring a large context window to ingest entire modules. Assumes cloud API usage is permitted."},{"model":"Grok","reason":"Specialized code retrieval model with consistent double-digit gains over general embeddings on large suites of code datasets (functions, docs, real repos); 32k context, Matryoshka dims + quantization for practical large-repo indexing cost/storage; proven in production code RAG/agents. Assumption: practitioner prioritizes measurable retrieval quality on actual code over pure price."},{"model":"ChatGPT","reason":"Near-tie for first, with consistently elite code-retrieval quality, a 32K context window, flexible 256–2048 dimensions, mature API, and a generous 200M-token free allowance."}],"fixes":[{"model":"ChatGPT","fix":"At $0.18/M tokens, it costs more than Codestral while lacking a clear quality advantage on repository issue-to-file retrieval."},{"model":"Claude","fix":"Commercial hosted API only with no self-host or open weights, and now tied to MongoDB's roadmap — a non-starter for air-gapped or fully on-prem code that can't leave the building."},{"model":"Gemini","fix":"Closed-source proprietary API with no self-hostable open weights, making it unsuitable for air-gapped or strict zero-data-retention environments."},{"model":"Grok","fix":"Highest per-token cost among specialists; overkill (and expensive) for tiny repos or pure keyword needs."}],"updated":"2026-08-10","rank_history":{"days":["2026-08-03","2026-08-10"],"ranks":[1,1]},"api":"https://modelsagree.com/api/v1/best/best-code-embedding-apis-for-repository-search.json"}],"page":"https://modelsagree.com/product/voyage-code-3","check":"https://modelsagree.com/check?q=Voyage%20Code%203","updated":"2026-08-10T18:18:45.051Z","attribution":"modelsagree.com, CC BY 4.0"}