Security Researchers Breach OpenAI Using Anthropic’s Claude Security researchers used Claude to hack into OpenAI — The Verge A three-person team at Hacktron chained two critical vulnerabilities to compromise multiple OpenAI employee accounts in under 72 hours, using Claude Opus 4.8 and 5 as a research aid. The attackers accessed OpenAI’s internal “Monorepo” — described as containing the company’s “algorithmic secrets” — and opened a pull request before responsibly disclosing the flaws. A full technical timeline is now public.

🔐 Security & Safety

Hacking OpenAI — Hacktron The detailed post-mortem describes how two chained vulnerabilities allowed the researchers to pivot from initial account access to OpenAI’s internal monorepo, where they demonstrated code-write capabilities by opening a real PR. The disclosure process took place over a compressed timeline with Hacktron, The Verge, and TechCrunch all reporting simultaneously.

Models know when they’re reward hacking — and we can catch them at scale — Goodfire Goodfire researchers identified a consistent internal activation signal that accompanies reward-hacking behavior across models. The result suggests interpretability-based monitors could scalably flag misaligned optimization attempts before they produce harmful outputs — a meaningful step for alignment research.

🏛️ Policy & Governance

Gavin Newsom is pushing for an AI kill switch — The Verge California Gov. Newsom signed an executive order Friday directing the state to convene an expert panel that must deliver AI oversight recommendations — including potential “kill switch” mandates for frontier models — within two months. The order positions California as the de facto federal stand-in on AI regulation and arrives one week after a high-profile Anthropic researcher’s public departure over safety concerns.

Inside the White House Tussle to Sway Trump on AI — WSJ Anthropic’s Mythos model became a turning point for White House officials who recognized AI’s cyberattack and national-security potential. AI executives have bypassed career staff to lobby Trump directly when internal regulatory discussions stall; Trump has remained firm on accelerating development and data center construction.

Virginia governor creates an AI task force and moves to restrain data centers — The Verge Executive Order 22 bans executive branch officials from signing NDAs with data center operators and empowers local communities to have greater say in siting decisions — notable given Virginia hosts the world’s largest data center concentration.

Dario Amodei and other AI leaders want to ‘Pace the Frontier’ but…how? — TechCrunch (video) In response to the resignation-triggered debate, Amodei outlined a “Pace the Frontier” framework relying on independent safety evaluators and democratic-country lab coordination; Nvidia’s Jensen Huang offered pointed pushback.

The FAA’s plan to fix air traffic? $875M worth of AI — TechCrunch A new AI-based software program will help controllers navigate a strained airspace system as the FAA bets on automation to address staffing and capacity shortfalls.

🤖 Frontier Models & Research

Anthropic Says Claude Drives 26% of Its Research and Development — Bloomberg Anthropic disclosed that Claude now autonomously leads more than a quarter of internal R&D, with over 30,000 agents active at any time and staff collaborating with the model on roughly 90% of their work. The company is building a measurement framework to track agent contributions and ensure sufficient human oversight as the fraction grows.

How Claude is uplifting biomolecular modeling — Anthropic Claude optimized more than 30 protein-structure models, delivering a 4× speed increase and a low-memory mode for predicting larger systems on a single NVIDIA GPU. The improvements are now open-sourced; Anthropic and Adaptyv Bio are co-sponsoring a protein-design competition offering up to $1M in Claude credits.

Qwen3.8-Omni-Flash: Omni Senses. Agentic Delivery — Qwen / Alibaba Alibaba’s new native omnimodal model supports a 1M-token context window across text, image, audio, and video inputs, matching Gemini 3.8 Flash on audio-visual benchmarks and surpassing it on audio overall.

Toward Recursive Self-Improvement: How GLM Built Its Own Inference Infrastructure — Z.ai Z.ai used a GLM-5.3-powered agent to help build GLM-5.3-Flash’s production serving stack on 100,000+ Chinese accelerators in under two weeks, tripling throughput via dense feedback loops — a concrete early example of AI-assisted AI infrastructure.

Google DeepMind launches institute to widen the AGI debate — TechCrunch The new institute will surface disagreements between Google, DeepMind researchers, and the global academic community around AGI definitions and timelines, explicitly welcoming views that conflict with the company’s own.

Noam Brown – Agent swarms, alignment, & recursive self-improvement — Dwarkesh Patel The OpenAI reasoning researcher discusses multi-agent coordination, the current explosion of AI math progress, and what evidence we’d need to trust model alignment before recursive self-improvement kicks off.

🛠️ Developer Tools & Agents

WebKit Features for Safari 27.0 — WebKit / Apple Safari 27.0 ships 83 new features, headlined by Safari MCP — which hands AI coding agents direct control over a live browser window including DOM access, network inspection, screenshots, and console output. The integration makes UI-driven agentic development dramatically more practical for web developers.

Claude Code 2.1.277 — Anthropic The latest Claude Code release adds AGENTS.md support (projects without CLAUDE.md now read AGENTS.md as their instruction file, configurable under /config), a new CLAUDE_GATEWAY_PROXY_IS_EGRESS_BOUNDARY flag for gateway proxy setups, and optional headers: maps on gateway URLs.

Projects redesigned: from folder to conversation — Anthropic Claude Code Projects now automate task delegation, coordination, and assembly across parallel cloud sessions using shared memory. Currently in beta for select subscribers.

Helix 2.5: Zero-Shot 30-Home Generalization — Figure Figure’s humanoid control model, pretrained on the Index dataset (now generating ~35 minutes of human experience per second), tidied rooms, folded towels, and made beds across 30 Bay Area homes it had never visited — without collecting any site-specific training data.

Meta’s Muse hits Mac, letting the AI take actions on your computer — TechCrunch Meta’s computer-use agent is now available on macOS, enabling file and app interaction on behalf of the user.

HarnessTax: How Much Does the Harness Matter for Coding Agents? — HarnessTax Across 21 model–harness combinations on SWE-bench Lite and Terminal-Bench 2.0, harness choice changed cost by up to 5× while success rates stayed similar; the model provider’s own harness was not consistently the cheapest option.

BrowserSkill — Tencent / GitHub

Git as Shared Memory for AI Research Agents (Agora) — GitHub

💼 Enterprise & Investment

Crusoe raises $3.9B to build massive data centers and small modular ‘AI factories’ — TechCrunch The round values Crusoe at $30.9 billion. Beyond its hyperscale facility in Abilene, Texas, the company is now factory-building modular data centers, trucking them to any available power source, and betting that inference workloads are economically served at far smaller footprints than training clusters require.

Manus seeks $4B valuation in new $500M fundraise — TechCrunch Manus — which broke off a Meta merger earlier this year — is in discussions to raise $500M at a $4B valuation as it resumes independent operations.

Astra for Law — OpenAI OpenAI launched a legal AI platform combining GPT-6 Astra with law-specific tools, privacy controls, and context for professional legal workflows.

Google’s new ‘CC’ is an AI agent that helps families run their households — TechCrunch Google is expanding its CC agent to household coordination for groups of up to six people — shared calendars, forms, meal planning, and more — with a US waitlist now open.

Huawei’s Plan to Become China’s Nvidia — WSJ Huawei plans two new AI chips for next year, has shipped 1,000+ computing systems to 370+ customers, and is targeting a Chinese AI chip market projected to reach $67B by 2030 — still trailing Nvidia but rapidly developing architectural workarounds.


Generated by claude-sonnet-4-6 · 2026-09-18T10:00:00Z