daily cyber × ai intelligence

index

tagged

[claude-fable-5.1]

5 editions · 4 items

September 14, 2026

  • GPT-6 Astra’s new agent benchmarks show a capability jump with large reliability caveats. Across six Vending-Bench runs, each simulating a year from a $500 starting balance, GPT-6 Astra averaged a final balance of $15,515 versus $5,422 for Claude Fable 5.1, according to The Decoder’s report on Andon Labs’ tests. On Drone-Bench, Astra’s best attempts beat the human-AI baseline on all five subtasks, including code for finding and following a specified person, but overall success remained unreliable. A separate robotics test had Astra complete 7 of 100 dual-arm tasks; MolmoAct2 completed 0 of the same 100. These were controlled evaluations, not live business or surveillance deployments, and they materially extend the initial Astra coverage (earlier coverage). · Model Capability, Evaluation & Safety

in Hermes Logs Reveal Unattended AI Post-Exploitation

September 4, 2026

Malware That Gaslights the AI Analyst

A North Korea-linked macOS implant called Gaslight embeds fake system error messages to trick AI analyzers into abandoning malware analysis while the payload executes. CISA added seven actively exploited vulnerabilities to its KEV catalog, including pre-auth flaws in SonicWall SMA 1000, JFrog Artifactory, and BerriAI LiteLLM, with post-exploitation involving reverse shells and crypto miners. ShinyHunters published stolen data from McKesson, Neogen, Elekta, and Jack Henry after extortion deadlines expired, while Shai-Hulud infostealer now targets 469 credential locations including AI tool configs. OpenAI's GPT-6 Astra crossed a critical cybersecurity threshold, finding two unknown zero-days during testing and marking what the company calls the start of the AGI era.