Path to Astra: critical capabilities and frontier safeguards
OpenAI announces Astra, the first model to meet the Critical cybersecurity capability threshold, incorporating enhanced safety safeguards.
OpenAI announces Astra, the first model to meet the Critical cybersecurity capability threshold, incorporating enhanced safety safeguards.
1 editorial report · 3 verified social mentions. The most authoritative report leads while later evidence completes the story.
🤖 Anthropic just dropped Claude Fable 5.1 and Mythos 5.1 Anthropic says its newest models are its most advanced yet for coding, research, and knowledge work and Fable 5.1 looks like a serious upgrade. Fable 5.1 scores 55.8% on Terminal-Bench 4.0, compared with 42.0% for Fable 5. On Terminal-Bench-Science, it hits 52.6%, more than double its predecessor. It’s not just smarter. It can deliver similar or better results at lower effort levels, while cache reads cost 75% less. Anthropic says this can cut real-world costs by around 25%, and up to 45% for highly agentic workloads. Anthropic is also rolling out stronger enterprise privacy with Enterprise Frontier Safeguards, while claiming its cybersecurity safeguards now flag benign requests about 60% less often. And Fable 5.1 isn’t limited to coding. Its research capabilities are being positioned as an early look at how AI could eventually contribute to scientific discovery. Fable 5.1 is available now. Mythos 5.1, built for cyberdefenders and life scientists, is available through trusted-access programs. Source. @aipost 🏴
Open mention🤖 Path to Astra: critical capabilities and frontier safeguards Astra is the first OpenAI model to meet the Critical cybersecurity capability threshold under the Preparedness Framework, with stronger safeguards for release. 📰 Source: OpenAI News 🔗 Link: https://openai.com/index/path-to-astra # AI # ArtificialIntelligence
Open mentionIn a Bloomberg interview, Sam Altman said Astra became powerful enough to hit OpenAI’s “c yber critical” threshold, which forced them to add new safeguards before release. Bloomberg also pressed him on AI finding zero-day exploits without human help. Altman clarified that the model they paused over that issue was a future model, not Astra itself. He also said future models will become more autonomous, which is why OpenAI is focusing heavily on monitoring, sandboxing and alignment. So the real issue isn’t just smarter AI. It’s AI that can increasingly act and work on its own. submitted by /u/didiTonic [link] [comments]
Open mention