Close Menu
TechurzTechurz

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    What's Hot

    North Korean Hackers Use EtherHiding to Hide Malware Inside Blockchain Smart Contracts

    October 16, 2025

    Rent a Cyber Friend will pay you to talk to strangers online and will show off its platform at TechCrunch Disrupt 2025

    October 16, 2025

    One Republican Now Controls a Huge Chunk of US Election Infrastructure

    October 16, 2025
    Facebook X (Twitter) Instagram
    Trending
    • North Korean Hackers Use EtherHiding to Hide Malware Inside Blockchain Smart Contracts
    • Rent a Cyber Friend will pay you to talk to strangers online and will show off its platform at TechCrunch Disrupt 2025
    • One Republican Now Controls a Huge Chunk of US Election Infrastructure
    • Deel hits $17.3B valuation after raising $300M from big-name VCs
    • Everything Apple launched on Oct. 15: M5 chipset, MacBook Pro, iPad, Vision Pro, more
    • Final 2 days to claim your exhibit table at Disrupt 2025
    • How to Assess and Choose the Right AI-SOC Platform
    • General Intuition lands $134M seed to teach agents spatial reasoning using video game clips
    Facebook X (Twitter) Instagram Pinterest Vimeo
    TechurzTechurz
    • Home
    • AI
    • Apps
    • News
    • Guides
    • Opinion
    • Reviews
    • Security
    • Startups
    TechurzTechurz
    Home»AI»Mistral launches new code embedding model that outperforms OpenAI and Cohere in real-world retrieval tasks
    AI

    Mistral launches new code embedding model that outperforms OpenAI and Cohere in real-world retrieval tasks

    TechurzBy TechurzMay 29, 2025No Comments4 Mins Read
    Share Facebook Twitter Pinterest LinkedIn Tumblr Reddit Telegram Email
    Mistral launches new code embedding model that outperforms OpenAI and Cohere in real-world retrieval tasks
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Join our daily and weekly newsletters for the latest updates and exclusive content on industry-leading AI coverage. Learn More

    With demand for enterprise retrieval augmented generation (RAG) on the rise, the opportunity is ripe for model providers to offer their take on embedding models. 

    French AI company Mistral threw its hat into the ring with Codestral Embed, its first embedding model, which it said outperforms existing embedding models on benchmarks like SWE-Bench.

    The model specializes in code and “performs especially well for retrieval use cases on real-world code data.” The model is available to developers for $0.15 per million tokens. 

    The company said the Codestral Embed “significantly outperforms leading code embedders” like Voyage Code 3, Cohere Embed v4.0 and OpenAI’s embedding model, Text Embedding 3 Large. 

    Super excited to announce @MistralAI Codestral Embed, our first embedding model specialized for code.

    It performs especially well for retrieval use cases on real-world code data. pic.twitter.com/ET321cRNli

    — Sophia Yang, Ph.D. (@sophiamyang) May 28, 2025

    Codestral Embed, part of Mistral’s Codestral family of coding models, can make embeddings that transform code and data into numerical representations for RAG. 

    “Codestral Embed can output embeddings with different dimensions and precisions, and the figure below illustrates the trade-offs between retrieval quality and storage costs,” Mistral said in a blog post. “Codestral Embed with dimension 256 and int8 precision still performs better than any model from our competitors. The dimensions of our embeddings are ordered by relevance. For any integer target dimension n, you can choose to keep the first n dimensions for a smooth trade-off between quality and cost.”

    Mistral tested the model on several benchmarks, including SWE-Bench and Text2Code from GitHub. In both cases, the company said Codestral Embed outperformed leading embedding models. 

    SWE- Bench

    Text2Code

    Use cases

    Mistral said Codestral Embed is optimized for “high-performance code retrieval” and semantic understanding. The company said the code works best for at least four kinds of use cases: RAG, semantic code search, similarity search and code analytics. 

    Embedding models generally target RAG use cases, as they can facilitate faster information retrieval for tasks or agentic processes. Therefore, it’s not surprising that Codestral Embed would focus on that. 

    The model can also perform semantic code search, allowing developers to find code snippets using natural language. This use case works well for developer tool platforms, documentation systems and coding copilots. Codestral Embed can also help developers identify duplicated code segments or similar code strings, which can be helpful for enterprises with policies regarding reused code. 

    The model supports semantic clustering, which involves grouping code based on its functionality or structure. This use case would help analyze repositories, categorize and find patterns in code architecture. 

    Competition is increasing in the embedding space

    Mistral has been on a roll with releasing new models and agentic tools. It released Mistral Medium 3, a medium version of its flagship large language model (LLM), which currently powers its enterprise-focused platform Le Chat Enterprise. 

    It also announced the Agents API, which allows developers to access tools for creating agents that perform real-world tasks and orchestrate multiple agents. 

    Mistral’s moves to offer more model options to developers have not gone unnoticed in developer spaces. Some on X note that Mistral’s timing in releasing Codestral Embed is “coming on the heels of increased competition.”

    Mistral AI Just Dropped a Game-Changer: Codestral Embed Crushes OpenAI and Google in Code Search Race

    French AI startup Mistral AI has quietly unleashed what could be the most significant breakthrough in code intelligence this year. Their brand-new Codestral Embed model isn’t…

    — Rahul Khorwal (@rkrahulkhorwal) May 28, 2025

    Mistral on a delivery mission

    — Joel Basson (@joelbasson) May 28, 2025

    However, Mistral must prove that Codestral Embed performs well not just in benchmark testing. While it competes against more closed models, such as those from OpenAI and Cohere, Codestral Embed also faces open-source options from Qodo, including Qodo-Embed-1-1.5 B.

    VentureBeat reached out to Mistral about Codestral Embed’s licensing options. 

    Daily insights on business use cases with VB Daily

    If you want to impress your boss, VB Daily has you covered. We give you the inside scoop on what companies are doing with generative AI, from regulatory shifts to practical deployments, so you can share insights for maximum ROI.

    Read our Privacy Policy

    Thanks for subscribing. Check out more VB newsletters here.

    An error occured.

    code Cohere embedding launches Mistral model OpenAI outperforms Realworld retrieval tasks
    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email
    Previous ArticleFederal Court Blocks Trump’s Tariffs, Finding the President Overstepped His Authority
    Next Article Shark FlexBreeze HydroGo review: an incredible portable fan to make summer easier indoors and outdoors
    Techurz
    • Website

    Related Posts

    Security

    MCPTotal Launches to Power Secure Enterprise MCP Workflows

    October 16, 2025
    Security

    Over 100 VS Code Extensions Exposed Developers to Hidden Supply Chain Risks

    October 16, 2025
    Security

    Source code and vulnerability info stolen from F5 Networks

    October 16, 2025
    Add A Comment
    Leave A Reply Cancel Reply

    Top Posts

    The Reason Murderbot’s Tone Feels Off

    May 14, 20259 Views

    Start Saving Now: An iPhone 17 Pro Price Hike Is Likely, Says New Report

    August 17, 20258 Views

    CNET’s Daily Tariff Price Tracker: I’m Keeping Tabs on Changes as Trump’s Trade Policies Shift

    May 27, 20258 Views
    Stay In Touch
    • Facebook
    • YouTube
    • TikTok
    • WhatsApp
    • Twitter
    • Instagram
    Latest Reviews

    Subscribe to Updates

    Get the latest tech news from FooBar about tech, design and biz.

    Most Popular

    The Reason Murderbot’s Tone Feels Off

    May 14, 20259 Views

    Start Saving Now: An iPhone 17 Pro Price Hike Is Likely, Says New Report

    August 17, 20258 Views

    CNET’s Daily Tariff Price Tracker: I’m Keeping Tabs on Changes as Trump’s Trade Policies Shift

    May 27, 20258 Views
    Our Picks

    North Korean Hackers Use EtherHiding to Hide Malware Inside Blockchain Smart Contracts

    October 16, 2025

    Rent a Cyber Friend will pay you to talk to strangers online and will show off its platform at TechCrunch Disrupt 2025

    October 16, 2025

    One Republican Now Controls a Huge Chunk of US Election Infrastructure

    October 16, 2025

    Subscribe to Updates

    Get the latest creative news from FooBar about art, design and business.

    Facebook X (Twitter) Instagram Pinterest
    • About Us
    • Contact Us
    • Privacy Policy
    • Terms and Conditions
    • Disclaimer
    © 2025 techurz. Designed by Pro.

    Type above and press Enter to search. Press Esc to cancel.