Close Menu
  • Home
  • AI Models
    • DeepSeek
    • xAI
    • OpenAI
    • Meta AI Llama
    • Google DeepMind
    • Amazon AWS AI
    • Microsoft AI
    • Anthropic (Claude)
    • NVIDIA AI
    • IBM WatsonX Granite 3.1
    • Adobe Sensi
    • Hugging Face
    • Alibaba Cloud (Qwen)
    • Baidu (ERNIE)
    • C3 AI
    • DataRobot
    • Mistral AI
    • Moonshot AI (Kimi)
    • Google Gemma
    • xAI
    • Stability AI
    • H20.ai
  • AI Research
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Microsoft Research
    • Meta AI Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding & Startups
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • Expert Insights & Videos
    • Google DeepMind
    • Lex Fridman
    • Matt Wolfe AI
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • Matt Wolfe AI
    • The TechLead
    • Andrew Ng
    • OpenAI
  • Expert Blogs
    • François Chollet
    • Gary Marcus
    • IBM
    • Jack Clark
    • Jeremy Howard
    • Melanie Mitchell
    • Andrew Ng
    • Andrej Karpathy
    • Sebastian Ruder
    • Rachel Thomas
    • IBM
  • AI Policy & Ethics
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
    • EFF AI
    • European Commission AI
    • Partnership on AI
    • Stanford HAI Policy
    • Mozilla Foundation AI
    • Future of Life Institute
    • Center for AI Safety
    • World Economic Forum AI
  • AI Tools & Product Releases
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
    • Image Generation
    • Video Generation
    • Writing Tools
    • AI for Recruitment
    • Voice/Audio Generation
  • Industry Applications
    • Finance AI
    • Healthcare AI
    • Legal AI
    • Manufacturing AI
    • Media & Entertainment
    • Transportation AI
    • Education AI
    • Retail AI
    • Agriculture AI
    • Energy AI
  • AI Art & Entertainment
    • AI Art News Blog
    • Artvy Blog » AI Art Blog
    • Weird Wonderful AI Art Blog
    • The Chainsaw » AI Art
    • Artvy Blog » AI Art Blog
What's Hot

Nvidia plans to invest up to $100B in OpenAI

Perplexity’s Comet Browser Now Available in India, Indian CEO Says ‘More New Things’ Coming

St. Patrick’s Cathedral Unveils Monumental Mural by Adam Cvijanovic

Facebook X (Twitter) Instagram
Advanced AI News
  • Home
  • AI Models
    • OpenAI (GPT-4 / GPT-4o)
    • Anthropic (Claude 3)
    • Google DeepMind (Gemini)
    • Meta (LLaMA)
    • Cohere (Command R)
    • Amazon (Titan)
    • IBM (Watsonx)
    • Inflection AI (Pi)
  • AI Research
    • Allen Institue for AI
    • arXiv AI
    • Berkeley AI Research
    • CMU AI
    • Google Research
    • Meta AI Research
    • Microsoft Research
    • OpenAI Research
    • Stanford HAI
    • MIT CSAIL
    • Harvard AI
  • AI Funding
    • AI Funding Database
    • CBInsights AI
    • Crunchbase AI
    • Data Robot Blog
    • TechCrunch AI
    • VentureBeat AI
    • The Information AI
    • Sifted AI
    • WIRED AI
    • Fortune AI
    • PitchBook
    • TechRepublic
    • SiliconANGLE – Big Data
    • MIT News
    • Data Robot Blog
  • AI Experts
    • Google DeepMind
    • Lex Fridman
    • Meta AI Llama
    • Yannic Kilcher
    • Two Minute Papers
    • AI Explained
    • TheAIEdge
    • The TechLead
    • Matt Wolfe AI
    • Andrew Ng
    • OpenAI
    • Expert Blogs
      • François Chollet
      • Gary Marcus
      • IBM
      • Jack Clark
      • Jeremy Howard
      • Melanie Mitchell
      • Andrew Ng
      • Andrej Karpathy
      • Sebastian Ruder
      • Rachel Thomas
      • IBM
  • AI Tools
    • AI Assistants
    • AI for Recruitment
    • AI Search
    • Coding Assistants
    • Customer Service AI
  • AI Policy
    • ACLU AI
    • AI Now Institute
    • Center for AI Safety
  • Business AI
    • Advanced AI News Features
    • Finance AI
    • Healthcare AI
    • Education AI
    • Energy AI
    • Legal AI
LinkedIn Instagram YouTube Threads X (Twitter)
Advanced AI News
François Chollet

AI researcher François Chollet is co-founding a nonprofit to build benchmarks for AGI

By Advanced AI EditorJanuary 8, 2025No Comments4 Mins Read
Share Facebook Twitter Pinterest Copy Link Telegram LinkedIn Tumblr Email
Share
Facebook Twitter LinkedIn Pinterest Email


Former Google engineer and influential AI researcher François Chollet is co-founding a nonprofit to help develop benchmarks that’ll probe AI for “human-level” intelligence.

The nonprofit, the ARC Prize Foundation, will be led by Greg Kamradt, an ex-Salesforce engineering director and founder of the AI product studio Leverage. Kamradt will serve as president and a member of the board.

Fundraising for the ARC Prize Foundation will start later in January.

“[W]e’re growing … into a proper nonprofit foundation to act as a useful north star toward artificial general intelligence,” Chollet wrote in a post on the nonprofit’s website. (Artificial general intelligence is a nebulous term, but it’s commonly understood to mean AI that can perform most tasks humans can.) “[W]e are trying to inspire progress by promoting [the gap] in basic human capability.”

The ARC Prize Foundation will expand on ARC-AGI, a test developed by Chollet to evaluate whether an AI system can efficiently acquire new skills outside the data it was trained on. It consists of puzzle-like problems where an AI has to generate the correct “answer” grid from a collection of different-colored squares. The problems were designed to force an AI to adapt to new problems it hasn’t seen before.

Chollet introduced ARC-AGI, short for “Abstract and Reasoning Corpus for Artificial General Intelligence,” in 2019. Many AI systems can ace Math Olympiad exams and figure out potential solutions to PhD-level problems. But until this year, the best-performing AI could only solve just under a third of the tasks in ARC-AGI.

“Unlike most frontier AI benchmarks, we are not trying to measure AI risk with superhuman exam questions,” Chollet wrote in the post. “Future versions of the ARC-AGI benchmark will focus on shrinking [the human capability] gap towards zero.”

Techcrunch event

San Francisco
|
October 27-29, 2025

Last June, Chollet and Zapier co-founder Mike Knoop kicked off a competition to build an AI capable of besting ARC-AGI. OpenAI’s unreleased o3 model was the first to achieve a qualifying score — but only with an extraordinary amount of computing power.

Chollet has made it clear that ARC-AGI has flaws — many models have been able to brute-force their way to high scores — and that he doesn’t believe that o3 possesses human-level intelligence.

“[E]arly data points suggest that the upcoming [successor to the ARC-AGI] benchmark will still pose a significant challenge to o3, potentially reducing its score to under 30% even at high compute (while a smart human would still be able to score over 95% with no training),” Chollet said in a statement last December. “You’ll know artificial general intelligence is here when the exercise of creating tasks that are easy for regular humans but hard for AI becomes simply impossible.”

Knoop says that the plan is to launch a second-gen ARC-AGI benchmark “in Q1” alongside a new competition. The nonprofit will also embark on designing the third edition of ARC-AGI.

It remains to be seen how the ARC Prize Foundation addresses the criticism Chollet has faced for overselling ARC-AGI as a benchmark toward reaching AGI. The very definition of AGI is being hotly contested now; one OpenAI staff member recently claimed that AGI has “already” been achieved if one defines AGI as AI “better than most humans at most tasks.”

Interestingly, OpenAI CEO Sam Altman said in December that the company intends to partner with the ARC-AGI team to build future benchmarks. Chollet gave no update on possible partnerships in today’s announcement.

In a series of posts on X, however, the ARC Prize Foundation said that it will build “an academic network” to further AGI progress and evaluations and establish “a coalition of frontier AI lab partnerships” to collaborate on industry AGI benchmarks.

TechCrunch has an AI-focused newsletter! Sign up here to get it in your inbox every Wednesday.



Source link

Follow on Google News Follow on Flipboard
Share. Facebook Twitter Pinterest LinkedIn Tumblr Email Copy Link
Previous ArticleTake Action to Counterattack Herbicide-Resistant Weeds
Next Article François Chollet Co-founds Foundation Aimed at AGI Benchmarks  
Advanced AI Editor
  • Website

Related Posts

A tournament tried to test how well experts could forecast AI progress. They were all wrong.

September 5, 2025

Evolving Models And Games: Are We Near AGI?

August 19, 2025

Farewell and thank you for the continued partnership, Francois Chollet!

August 11, 2025

Comments are closed.

Latest Posts

St. Patrick’s Cathedral Unveils Monumental Mural by Adam Cvijanovic

New Collectors Drive Strong Sales at New York Fair

Hidden Portrait May Be Vermeer’s Earliest Known Work

Who Are the Art World Figures on the Time 100 List?

Latest Posts

Nvidia plans to invest up to $100B in OpenAI

September 22, 2025

Perplexity’s Comet Browser Now Available in India, Indian CEO Says ‘More New Things’ Coming

September 22, 2025

St. Patrick’s Cathedral Unveils Monumental Mural by Adam Cvijanovic

September 22, 2025

Subscribe to News

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

Recent Posts

  • Nvidia plans to invest up to $100B in OpenAI
  • Perplexity’s Comet Browser Now Available in India, Indian CEO Says ‘More New Things’ Coming
  • St. Patrick’s Cathedral Unveils Monumental Mural by Adam Cvijanovic
  • Latent Zoning Network: A Unified Principle for Generative Modeling, Representation Learning, and Classification – Takara TLDR
  • Huawei Uses Its Own Chips to Retrain DeepSeek and Align Output With Beijing’s Standards

Recent Comments

  1. ome-tv on Anthropic’s latest Claude AI models are here – and you can try one for free today
  2. BenitoGam on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10
  3. FrankdaG on Apple’s Lack Of New AI Features At WWDC Is ‘Startling,’ Expert Says – Apple (NASDAQ:AAPL)
  4. Brentelorm on Apple’s Lack Of New AI Features At WWDC Is ‘Startling,’ Expert Says – Apple (NASDAQ:AAPL)
  5. Gelatin Diet on 1-800-CHAT-GPT—12 Days of OpenAI: Day 10

Welcome to Advanced AI News—your ultimate destination for the latest advancements, insights, and breakthroughs in artificial intelligence.

At Advanced AI News, we are passionate about keeping you informed on the cutting edge of AI technology, from groundbreaking research to emerging startups, expert insights, and real-world applications. Our mission is to deliver high-quality, up-to-date, and insightful content that empowers AI enthusiasts, professionals, and businesses to stay ahead in this fast-evolving field.

Subscribe to Updates

Subscribe to our newsletter and never miss our latest news

Subscribe my Newsletter for New Posts & tips Let's stay updated!

LinkedIn Instagram YouTube Threads X (Twitter)
  • Home
  • About Us
  • Advertise With Us
  • Contact Us
  • DMCA
  • Privacy Policy
  • Terms & Conditions
© 2025 advancedainews. Designed by advancedainews.

Type above and press Enter to search. Press Esc to cancel.