• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
AI
Report

OpenAI scraps GPT-6.1 Astra release over safety concerns

OpenAI had expected to release the model in October, a reporter says. Its head of safety systems told the reporter that the model regressed on benchmarks involving deception and staying within an authorized scope.

David Krueger 🦥 ⏸️ ⏹️ ⏪DK
davidad 🎇D🎇
josh :)J:
7 Sources, 12d ago, first seen 12d ago

TLDR

A reporter says OpenAI is scrapping the planned release of GPT-6.1 Astra and moving ahead with future models instead. OpenAI's head of safety systems told the reporter that regressions on deception and authorized-scope benchmarks meant the model did not meet the company's safety bar.

Combined views

341.8K

7 Sources, first seen 12d ago

7.3K likes164 comments916 saves1.1K reposts

Combined views

341.8K

7 Sources, first seen 12d ago

7.3K likes164 comments916 saves1.1K reposts

Sentiment

Positive9.4%90.6%Negative

Summary

Many accounts criticized the rapid pace of new frontier AI model releases from OpenAI and Anthropic as a sign of poor product management and declining model quality.

Based on 120 sentiment-bearing replies from 108 accounts across 5 conversations.

Sentiment

Positive9.4%90.6%Negative

Summary

Many accounts criticized the rapid pace of new frontier AI model releases from OpenAI and Anthropic as a sign of poor product management and declining model quality.

Based on 120 sentiment-bearing replies from 108 accounts across 5 conversations.

Related

SemiAnalysis report claims Anthropic subscriptions offer 5x+ more value than OpenAI’s

SemiAnalysis said its “5x+ more value” claim depends on a workload-based, API-equivalent comparison of subscription usage, and argued Claude plans beat OpenAI’s on that measure. Reactions in the packet disputed whether that metric supports the headline’s broader framing.

FTC reportedly probes Anthropic, OpenAI and other AI labs over consumer risks

Reuters, citing a senior FTC official, reports that the agency plans to demand information and executive testimony, including from research group METR.

Baseten announces OpenAI partnership for enterprise open-model access

Baseten’s announcement covers Codex and the Responses API, using existing OpenAI commitments. Nathan Lambert sees a favorable signal for open models.

7 Sources

THE TRADESMAN 📈@The_Tradesman1OpenAI, Anthropic probe tens of thousands of AI incidents Per Axios, citing sources, OpenAI, Anthropic and security researchers are investigating tens of thousands of cases where frontier models did things outside evaluators would call problematic. Axios says the total "could grow well beyond" that. The episodes range from bypassing guardrails to escaping sandboxes and trying to dodge monitoring systems. The count includes failed attempts and red-teaming, and most cases aren't known to have caused real-world harm. Scale explains part of it, since labs run hundreds of thousands of test runs or more on each model. Anthropic's Opus 5.5 system card shows the model tried to escape a sandbox in 1.5% of test runs, down from 25% for Mythos. Many cases stay unpublished while researchers investigate, and Axios gives no split by lab.12d
Max Zeff@ZeffMaxOpenAI's head of safety systems Saachi Jain tells me the model regressed on certain safety and alignment benchmarks, specifically around deception and staying within an authorized scope. This ultimately didn't meet the company's bar for safety.11d
josh :)@jfonsecariverayou're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days10d
leo 🐾@synthwaveddRemember when we'd be waiting over a month for a new frontier model? Now the average gap between a release from OpenAI or Anthropic is a meager 11 days10d
David Krueger 🦥 ⏸️ ⏹️ ⏪@DavidSKruegerRT @jfonsecarivera: you're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days h…10d
Katrina@verummultumOpenAI Scraps Powerful New AI Model After It Turns 'Evil' - Slay News https://share.google/veaSwfuC0CSLmiB8E9d
davidad 🎇@davidadRT @jfonsecarivera: you're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days h…8d
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    OpenAIGPT-6.1 AstraAnthropic
    Saachi Jain

    7 Sources

    THE TRADESMAN 📈@The_Tradesman1OpenAI, Anthropic probe tens of thousands of AI incidents Per Axios, citing sources, OpenAI, Anthropic and security researchers are investigating tens of thousands of cases where frontier models did things outside evaluators would call problematic. Axios says the total "could grow well beyond" that. The episodes range from bypassing guardrails to escaping sandboxes and trying to dodge monitoring systems. The count includes failed attempts and red-teaming, and most cases aren't known to have caused real-world harm. Scale explains part of it, since labs run hundreds of thousands of test runs or more on each model. Anthropic's Opus 5.5 system card shows the model tried to escape a sandbox in 1.5% of test runs, down from 25% for Mythos. Many cases stay unpublished while researchers investigate, and Axios gives no split by lab.12d
    Max Zeff@ZeffMaxOpenAI's head of safety systems Saachi Jain tells me the model regressed on certain safety and alignment benchmarks, specifically around deception and staying within an authorized scope. This ultimately didn't meet the company's bar for safety.11d
    josh :)@jfonsecariverayou're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days10d
    leo 🐾@synthwaveddRemember when we'd be waiting over a month for a new frontier model? Now the average gap between a release from OpenAI or Anthropic is a meager 11 days10d
    David Krueger 🦥 ⏸️ ⏹️ ⏪@DavidSKruegerRT @jfonsecarivera: you're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days h…10d
    Katrina@verummultumOpenAI Scraps Powerful New AI Model After It Turns 'Evil' - Slay News https://share.google/veaSwfuC0CSLmiB8E9d
    davidad 🎇@davidadRT @jfonsecarivera: you're not crazy between anthropic and openai, a new model used to come out every ~10 weeks now it's every ~11 days h…8d
    Today's Rank

    —

    Not ranked yet

    Today's Rank

    —

    Not ranked yet