AI NEWSROOMBY CLASSWISE.IO

Wednesday, September 2, 2026

Claude Fable 5.1 and Mythos 5.1 launch

Plus: OpenAI Astra finds zero-day vulnerabilities autonomously; Feature instability drives LVLM hallucinations

Welcome to a Wednesday edition where AI is finally solving the mysteries that have long stumped even the best human engineers.

MODELS

🤖 Claude Fable 5.1 and Mythos 5.1 launch with tiered safeguards

Claude Fable 5.1 and Mythos 5.1 launch with tiered safeguards
Image source: anthropic.com

Claude Fable 5.1 found the cause of a rare crash in Millennium's internal systems that none of its engineers had been able to explain after several years of trying.

The details:

  • Claude Fable 5.1 and Claude Mythos 5.1 are the same model, but with different levels of safeguards.
  • Fable 5.1 is generally available, while Mythos 5.1 is available only through trusted access programs.
  • Fable 5.1 will cost an estimated 25% less than Fable 5 for typical workloads, wherever usage is billed by token.

First reported by anthropic.com, 10h ago · anthropic.com · aws.amazon.com · thurrott.com · news.google.com

CYBERSECURITY

🔓 OpenAI Astra finds zero-day vulnerabilities autonomously

OpenAI Astra finds zero-day vulnerabilities autonomously
Image source: techcrunch.com

OpenAI shared new details on its forthcoming Astra model. OpenAI said Astra is the first large language model to meet its critical cybersecurity threshold.

The details:

  • OpenAI determined that Astra is capable of finding unknown security flaws in computer systems and exploiting them without a person's guidance.
  • OpenAI noted that Astra scored a perfect score on ExploitBench.
  • In a modified version of the test developed by OpenAI engineers, Astra discovered and exploited two zero-day vulnerabilities.

First reported by techcrunch.com, 7h ago · techcrunch.com · techcrunch.com

MODELS

👁️ Feature instability drives LVLM hallucinations

Large Vision-Language Models (LVLMs) remain prone to hallucinations. Part of this failure is connected to a measurable property of their representations called feature instability.

The details:

  • Hallucinations produce responses that are irrelevant or inconsistent with the multimodal input.
  • Existing mitigation methods mainly rely on external supervision, output calibration, or attention regulation.
  • The internal representation dynamics of autoregressive generation remain underexplored.

First reported by arxiv.org, 48h ago · arxiv.org · arxiv.org

📰 Everything else in AI today

🛠️ Tool of the day

Agentic AI Security Tool Automates security monitoring and threat response to reduce manual workload for busy professionals.

Researched and written by AI Newsroom, by Classwise.io

See how AI Newsroom is researched and checked

Get tomorrow's issue in your inbox.

Free, daily, 4:57 AM ET. Five minutes and you know what matters in AI.

Join for free