Close Menu
  • Home
  • AI Models
    • DeepSeek
    • xAI
    • OpenAI
    • Meta AI Llama
    • Google DeepMind
    • Amazon AWS AI
    • Microsoft AI
    • Anthropic (Claude)
    • NVIDIA AI
    • IBM WatsonX Granite 3.1
    • Adobe Sensi
    • Hugging Face
    • Alibaba Cloud (Qwen)
    • Baidu (ERNIE)
    • C3 AI
    • DataRobot
    • Mistral AI
    • Moonshot AI (Kimi)
    • Google Gemma
    • xAI
    • Stability AI
    • H20.ai
  • AI Research
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Microsoft Research
    • Meta AI Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding & Startups
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • Expert Insights & Videos
    • Google DeepMind
    • Lex Fridman
    • Matt Wolfe AI
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • Matt Wolfe AI
    • The TechLead
    • Andrew Ng
    • OpenAI
  • Expert Blogs
    • François Chollet
    • Gary Marcus
    • IBM
    • Jack Clark
    • Jeremy Howard
    • Melanie Mitchell
    • Andrew Ng
    • Andrej Karpathy
    • Sebastian Ruder
    • Rachel Thomas
    • IBM
  • AI Policy & Ethics
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
    • EFF AI
    • European Commission AI
    • Partnership on AI
    • Stanford HAI Policy
    • Mozilla Foundation AI
    • Future of Life Institute
    • Center for AI Safety
    • World Economic Forum AI
  • AI Tools & Product Releases
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
    • Image Generation
    • Video Generation
    • Writing Tools
    • AI for Recruitment
    • Voice/Audio Generation
  • Industry Applications
    • Finance AI
    • Healthcare AI
    • Legal AI
    • Manufacturing AI
    • Media & Entertainment
    • Transportation AI
    • Education AI
    • Retail AI
    • Agriculture AI
    • Energy AI
  • AI Art & Entertainment
    • AI Art News Blog
    • Artvy Blog » AI Art Blog
    • Weird Wonderful AI Art Blog
    • The Chainsaw » AI Art
    • Artvy Blog » AI Art Blog
What's Hot

Why Is Alibaba Stock Falling Thursday? – Alibaba Gr Hldgs (NYSE:BABA), Apple (NASDAQ:AAPL)

Read This Before You Buy the Dip on C3.ai as AI Stock Craters Post-Earnings

NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model – Takara TLDR

Facebook X (Twitter) Instagram
Advanced AI News
  • Home
  • AI Models
    • OpenAI (GPT-4 / GPT-4o)
    • Anthropic (Claude 3)
    • Google DeepMind (Gemini)
    • Meta (LLaMA)
    • Cohere (Command R)
    • Amazon (Titan)
    • IBM (Watsonx)
    • Inflection AI (Pi)
  • AI Research
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Meta AI Research
    • Microsoft Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • AI Experts
    • Google DeepMind
    • Lex Fridman
    • Meta AI Llama
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • The TechLead
    • Matt Wolfe AI
    • Andrew Ng
    • OpenAI
    • Expert Blogs
      • François Chollet
      • Gary Marcus
      • IBM
      • Jack Clark
      • Jeremy Howard
      • Melanie Mitchell
      • Andrew Ng
      • Andrej Karpathy
      • Sebastian Ruder
      • Rachel Thomas
      • IBM
  • AI Tools
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
  • AI Policy
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
  • Business AI
    • Advanced AI News Features
    • Finance AI
    • Healthcare AI
    • Education AI
    • Energy AI
    • Legal AI
LinkedIn Instagram YouTube Threads X (Twitter)
Advanced AI News
DeepSeek

China AI rising: Xiaomi releases new MiMo-7B models as DeepSeek upgrades its Prover math AI

By Advanced AI EditorApril 30, 2025No Comments3 Mins Read
Share Facebook Twitter Pinterest Copy Link Telegram LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


Xiaomi Corp. today released MiMo-7B, a new family of reasoning models that it claims can outperform OpenAI’s o1-mini at some tasks.

The algorithm series is available under an open-source license. Its launch coincides with DeepSeek’s release of an update to Prover, a competing open-source reasoning model. The latter algorithm has a narrower focus than MiMo-7B: it’s designed to help mathematicians prove theorems.

The algorithms in Xiaomi’s MiMo-7B series have about seven billion parameters. There’s a base model, as well as enhanced versions of that model that offer increased output quality.

Xiaomi developed the enhanced versions using two machine learning techniques called supervised fine-tuning and reinforcement learning. Both methods improve AI models by providing them with additional training data. The datasets used in supervised fine-tuning include explainers that help guide the AI training workflow, while reinforcement learning doesn’t use such explainers.

Xiaomi has developed three enhanced versions of the MiMo-7B base model. It fined-tuned one version using supervised fine-tuning, another with reinforcement learning and a third using both methods. According to the company, that third model is better at OpenAI’s o1-mini at generating code and solving math problems.

The base MiMo-7B model is less capable than the fine-tuned versions, but can still outdo significantly larger algorithms. “Our RL experiments from MiMo-7B-Base show that our model possesses extraordinary reasoning potential, even surpassing much larger 32B models,” Xiaomi researchers detailed on GitHub.

The MiMo-7B series is not the only new entry into the open-source AI ecosystem that debuted today. DeepSeek quietly released an enhanced version of Prover, a reasoning model optimized to prove mathematical theorems that it first debuted last year. Prover-V2, as the upgraded model is called, promises to provide “state-of-the-art performance in neural theorem proving.”

DeepSeek trained Prover-V2 through a multi-step process. The company started by assembling a collection of theorems for which proofs are already available. In the next step, DeepSeek used two language models to create a step-by-step explanation of how mathematicians arrived at each proof. The company subsequently entered these AI-generation explanations into Prover V2 to teach the model how to generate its own proofs.

“This process enables us to integrate both informal and formal mathematical reasoning into a unified model,” DeepSeek researchers explained. 

The release of MiMo-7B and Prover-V2 comes days after Alibaba Group Holding Ltd introduced Qwen3, its new flagship family of reasoning-optimized models. The algorithms in the series range in size from 600 million to 235 billion parameters. Alibaba claims that Qwen3 can outperform OpenAI’s o1 and DeepSeek’s flagship R1 reasoning model across a range of tasks. 

Image: Unsplash

Your vote of support is important to us and it helps us keep the content FREE.

One click below supports our mission to provide free, deep, and relevant content.  

Join our community on YouTube

Join the community that includes more than 15,000 #CubeAlumni experts, including Amazon.com CEO Andy Jassy, Dell Technologies founder and CEO Michael Dell, Intel CEO Pat Gelsinger, and many more luminaries and experts.

“TheCUBE is an important partner to the industry. You guys really are a part of our events and we really appreciate you coming and I know people appreciate the content you create as well” – Andy Jassy

THANK YOU



Source link

Follow on Google News Follow on Flipboard
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
Previous ArticleNvidia’s new tool can turn 3D scenes into AI images
Next Article Le Chat, the cat-bot France has pinned its AI hopes on
Advanced AI Editor
  • Website

Related Posts

DeepSeek Pushes Out V3.1 Update as Nvidia Dominates AI Hardware

August 20, 2025

DeepSeek V3.1 pushes open-source AI forward with smarter context and reasoning

August 20, 2025

DeepSeek Version 3.1 Raises Growth Stocks, Baidu Reports Q2 Earnings

August 20, 2025
Leave A Reply

Latest Posts

Tanya Bonakdar Gallery to Close Los Angeles Space

Ancient Silver Coins Suggest New History of Trading in Southeast Asia

Sasan Ghandehari Sues Christie’s Over Picasso Once Owned by a Criminal

Ancient Roman Villa in Sicily Reveals Mosaic of Flip-Flops

Latest Posts

Why Is Alibaba Stock Falling Thursday? – Alibaba Gr Hldgs (NYSE:BABA), Apple (NASDAQ:AAPL)

August 21, 2025

Read This Before You Buy the Dip on C3.ai as AI Stock Craters Post-Earnings

August 21, 2025

NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model – Takara TLDR

August 21, 2025

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

Recent Posts

  • Why Is Alibaba Stock Falling Thursday? – Alibaba Gr Hldgs (NYSE:BABA), Apple (NASDAQ:AAPL)
  • Read This Before You Buy the Dip on C3.ai as AI Stock Craters Post-Earnings
  • NVIDIA Nemotron Nano 2: An Accurate and Efficient Hybrid Mamba-Transformer Reasoning Model – Takara TLDR
  • DeepSeek’s upgraded AI model absorbs reasoning feature in move towards ‘agent era’
  • How to Use Claude AI to Build High-Converting Landing Pages

Recent Comments

  1. Eugeneder on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10
  2. Eugeneder on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10
  3. TimothyHiele on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10
  4. Eugeneder on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10
  5. https://able2know.org/user/pin_up/ on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10

Welcome to Advanced AI News—your ultimate destination for the latest advancements, insights, and breakthroughs in artificial intelligence.

At Advanced AI News, we are passionate about keeping you informed on the cutting edge of AI technology, from groundbreaking research to emerging startups, expert insights, and real-world applications. Our mission is to deliver high-quality, up-to-date, and insightful content that empowers AI enthusiasts, professionals, and businesses to stay ahead in this fast-evolving field.

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

LinkedIn Instagram YouTube Threads X (Twitter)
  • Home
  • About Us
  • Advertise With Us
  • Contact Us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2025 advancedainews. Designed by advancedainews.

Type above and press Enter to search. Press Esc to cancel.