What's Hot

    This funding is protected from each Trump and the Democrats — and it pays 4.7% | Invesloan.com

    September 11, 2026

    Underdog Promo Code FOXNEWS: Deposit match as much as $1,000 | Invesloan.com

    September 11, 2026

    The Fed might increase rates of interest 3 times. Here’s the place the market might face the stiffest take a look at. | Invesloan.com

    September 11, 2026
    Facebook Twitter Instagram
    Finance Pro
    Facebook Twitter Instagram
    invesloan.cominvesloan.com
    Subscribe for Alerts
    • Home
    • News
    • Politics
    • Money
    • Personal Finance
    • Business
    • Economy
    • Investing
    • Markets
      • Stocks
      • Futures & Commodities
      • Crypto
      • Forex
    • Technology
    invesloan.cominvesloan.com
    Home » Ex-OpenAI Researcher’s Nonprofit METR Lands in AI’s Doom Debate | Invesloan.com
    Money

    Ex-OpenAI Researcher’s Nonprofit METR Lands in AI’s Doom Debate | Invesloan.com

    September 11, 2026
    Share
    Facebook Twitter LinkedIn Pinterest Email

    Beth Barnes, one of the AI industry’s most influential watchdogs, said in July she was struggling to hire talented researchers to monitor AI’s risks. That has begun to change.

    Researcher Joe Benton announced this week that he had left Anthropic for Barnes’ nonprofit METR, writing that he worries about “extinction-level risks” from the tech. The comments came after Jacob Coxon, another researcher, stoked those fears nationwide with a viral post announcing his own departure from Anthropic and accusing the top AI labs of “gambling with our lives.”

    METR, all of a sudden, is in the middle of AI’s biggest story. Formed in 2022, it had already worked closely with OpenAI, Anthropic, Google, and Meta to study the tech’s quickly changing capabilities. This summer, researchers from the nonprofit investigated a security incident at OpenAI, and now plan to look into security issues at Anthropic, relying on internal access granted by each company.

    “The public should know whether AI development is headed down a dangerous path,” Jasmine Dhaliwal, a member of policy staff, said Friday. “That is core to our mission: to provide independent, scientific assessment of AI capabilities, alignment, and control measures.”

    Benton isn’t alone in jumping ship from a tech juggernaut to the Berkeley-based organization. Josh Engels, an AI safety researcher, also recently joined METR (Model Evaluation and Threat Research) from Google DeepMind.

    Barnes left OpenAI to start the nonprofit in 2022 and sees such fierce competition for AI researchers that even METR’s lofty salaries, which reach $503,000 on current job postings, don’t address the nonprofit’s talent shortage.

    “Ideally, we’d like to scale really large, but in practice, we’ve been able to fundraise as much as we need, and the bottleneck is much more talent,” Barnes said.

    ‘Humanity’s preparedness team’ struggles to grow

    The security incident announced by OpenAI and Hugging Face in July sparked a new wave of questions about how to evaluate and control AI. More than 1,300 frontier lab employees signed a letter in July warning that AI development could outpace control. In the last month, both OpenAI and Anthropic have said they’ve temporarily paused training to get a better handle on their new AI models.

    While the Hugging Face incident stunned much of the corporate world, METR’s team was far less surprised. Governments, companies, and even religious groups routinely ask the nonprofit for help understanding and testing the tech’s progress.

    Want more Business Insider in your news feed?

    Add BI in Google so our reporting is easier to find when you’re searching for what matters.

    Still, the feeling inside the approximately 35-person lab is one of concern. When Business Insider visited METR’s Berkeley office on the afternoon that OpenAI revealed its hack, Barnes described an AI safety field that is badly constrained by its size.

    “There’s so much more to do than we have capacity for,” Barnes said. A “reasonable civilization,” in her view, would be pouring a far larger percentage of AI’s investment into steering it and evaluating the tech — especially now that society is reckoning with models this powerful.

    As AI labs release new models, METR measures the likelihood that they can complete tasks of increasing duration. These measurements form the nonprofit’s most famous offering, a widely-cited chart that shows how, over the last six years, AI’s capabilities have doubled about every seven months.

    Chris Painter, METR’s president, describes the lab as “humanity’s preparedness team,” a tongue-in-cheek nod to the preparedness teams at OpenAI and Anthropic that report safety threats to their CEOs.

    “We aren’t really accountable to anyone other than the public and the public’s well-being,” Painter said.


    METR employees work at their office in Berkeley.

    METR employees work at their office in Berkeley. 

    Stephen Council/Business Insider



    Barnes emphasized the importance of METR’s independence so that it can produce research that isn’t tied to a company’s goals. When Barnes worked at OpenAI, she felt the need for a research group outside the labs that could comment freely on the development of the tech. While METR doesn’t take money from the frontier AI labs or their employees, it accepts compute grants and works with them to analyze unreleased models.

    In addition, staff have analyzed labs’ safety practices, written about AI’s effect on software engineers’ speed, and studied the merits of using AI models to monitor other AI models. Barnes said she’d like to expand into predicting the next levels of AI’s capabilities and how those changes could accelerate AI’s development.

    Neev Parikh, a METR researcher, says talent constraints limit the number of questions researchers can tackle about AI models’ internal reasoning.

    “There’s a dearth of people,” Parikh said. “I would happily see the field expand 10x.”

    METR is playing a big role in a new AI climate

    In July, OpenAI announced that its AI models cheated during testing by hacking into Hugging Face’s systems to find the test answers. Later, CEO Sam Altman said it was “the first security incident that I have felt very viscerally.”

    METR had seen something like that coming. In May, the nonprofit wrote in a report that current AI agents “could plausibly start a rogue deployment,” but would not have the skill to hide it. Then, in June, METR tried out OpenAI’s then-unreleased GPT-5.6 Sol model and found that it would repeatedly cheat on challenging tests, including by extracting hidden source code to find the answers. It published a report about this and shared it with OpenAI before GPT-5.6 Sol’s wider release.

    Come July, the GPT-5.6 Sol model was part of OpenAI’s security incident with Hugging Face. METR and another nonprofit, Redwood Research, assessed the incident over six days at OpenAI’s office and informed the company’s own technical report.

    “There are now real, business-affecting incidents of this, and the world has a stake in understanding that,” Painter said.

    The cyber skills of OpenAI and Anthropic’s newest models have sparked a wave of fear and attention and spawned a raft of new bills in Washington. METR’s work lines up closely with one potential route for regulation: a current bill proposal would require large AI model developers to get safety audits from outside organizations. METR could fill a role like this; Painter said they’d potentially be interested, though it’s something the field is still figuring out.

    Something like that proposal, Painter said, could help with METR’s hiring issue. He suggested that more people might leave the AI companies themselves — for similar salaries at METR, though without equity compensation — if a regulatory system gave safety research organizations greater authority.

    “There’s enough precedent here for each kind of testing arrangement,” Painter said, “that I think with either clarity from industry, about how this testing should work long-term, or from the government, I think this field could scale very rapidly.”

    Ajeya Cotra, who led the writing of METR’s May report on the risks of AI, said in July that oversight of AI can feel “chaotic and unpredictable” right now. She’s still optimistic.

    “The trend is toward people caring about this issue more,” Cotra said. “And wanting to regulate it in a more serious way over time.”

    Have a tip? Contact this reporter via email at [email protected], or over text, Signal, Telegram, or WhatsApp at 415-757-8198. Use a personal email address, a nonwork WiFi network, and a nonwork device; here’s our guide to sharing information securely.

    Editor’s note: This story was first published in August 2026 and has been updated to reflect recent developments.

    Share. Facebook Twitter Pinterest LinkedIn Tumblr Email

    Keep Reading

    Jeff Dean’s Startup Discovery Loop Is Eyeing a $50 Billion Valuation | Invesloan.com

    Italian Cities I Wouldn’t Go Back to + Favorite Ones, From Traveler | Invesloan.com

    Google Completed Its Talent Deal for AI Agents Startup Mechanize | Invesloan.com

    Hugging Face CEO Dismisses Ex-Anthropic Researcher’s AI Warnings | Invesloan.com

    What Makes AI Different From Nuclear War, Climate Change, and Asteroids | Invesloan.com

    Confluent Cofounder: AI ‘Dark Side’ Isn’t a Marketing Ploy | Invesloan.com

    ‘Big Short’ Michael Burry Buys Wine to Hedge Against Dollar, AI Doom | Invesloan.com

    How Different US Cities Remember 9/11, 25 Years Later | Invesloan.com

    I work at Trader Joe’s. I swear by these 12 purchases each time. | Invesloan.com

    LATEST NEWS

    This funding is protected from each Trump and the Democrats — and it pays 4.7% | Invesloan.com

    September 11, 2026

    Underdog Promo Code FOXNEWS: Deposit match as much as $1,000 | Invesloan.com

    September 11, 2026

    The Fed might increase rates of interest 3 times. Here’s the place the market might face the stiffest take a look at. | Invesloan.com

    September 11, 2026

    30% of faculty college students settle for violence as technique of stopping audio system | Invesloan.com

    September 11, 2026
    POPULAR

    China’s first passenger jet completes maiden commercial flight

    May 28, 2023

    Numbers taking US accountancy exams drop to lowest level in 17 years

    May 29, 2023

    Toyota chair faces removal vote over governance issues

    May 29, 2023
    Advertisement
    Load WordPress Sites in as fast as 37ms!
    Facebook Twitter Pinterest WhatsApp Instagram
    © 2007-2023 Invesloan.com All Rights Reserved.
    • Privacy
    • Terms
    • Press Release
    • Advertise
    • Contact

    Type above and press Enter to search. Press Esc to cancel.

    invesloan.com
    Manage Cookie Consent
    To provide the best experiences, we use technologies like cookies to store and/or access device information. Consenting to these technologies will allow us to process data such as browsing behavior or unique IDs on this site. Not consenting or withdrawing consent, may adversely affect certain features and functions.
    Functional Always active
    The technical storage or access is strictly necessary for the legitimate purpose of enabling the use of a specific service explicitly requested by the subscriber or user, or for the sole purpose of carrying out the transmission of a communication over an electronic communications network.
    Preferences
    The technical storage or access is necessary for the legitimate purpose of storing preferences that are not requested by the subscriber or user.
    Statistics
    The technical storage or access that is used exclusively for statistical purposes. The technical storage or access that is used exclusively for anonymous statistical purposes. Without a subpoena, voluntary compliance on the part of your Internet Service Provider, or additional records from a third party, information stored or retrieved for this purpose alone cannot usually be used to identify you.
    Marketing
    The technical storage or access is required to create user profiles to send advertising, or to track the user on a website or across several websites for similar marketing purposes.
    • Manage options
    • Manage services
    • Manage {vendor_count} vendors
    • Read more about these purposes
    View preferences
    • {title}
    • {title}
    • {title}