OpenAI GPT-6 Astra Launches, Marking What OpenAI Calls the ‘AGI Era’ The Verge AI / TechCrunch AI OpenAI unveiled GPT-6 Astra, describing it as “a generational leap in capability” across cybersecurity, software engineering, science, and computer use. It is the first OpenAI model formally designated as crossing the company’s “critical cybersecurity capability threshold,” and OpenAI has positioned the launch as the beginning of the AGI era — while promising robust safeguards. The release comes alongside warnings from AI safety researchers alarmed by Astra’s underlying “recurrent depth” reasoning technique.
🚀 Frontier Models
OpenAI launches Astra, its powerful (and controversial) new model — TechCrunch AI GPT-6 Astra sets a new bar for computer and browser use, with OpenAI claiming unmatched “speed, accuracy, and safety.” The model’s use of recurrent depth — a technique enabling thinking outside standard sequential reasoning chains — has alarmed AI safety researchers, who argue it makes the model’s behavior harder to predict or audit.
Google says its new Gemini 3.8 Flash model ‘works harder’ but might cost more — The Verge AI Arriving weeks after Gemini 3.7 Flash, the 3.8 Flash release adds iterative tool-calling and extra reasoning steps on complex tasks. Introductory pricing matches 3.7 Flash at $0.75 per million input tokens, but Google hints at higher tiers for heavier workloads.
Meta’s Ava model with computer-use capabilities is in closed testing — AI Weekly Meta’s answer to Astra and Claude’s computer-use features is quietly being evaluated inside the Meta AI desktop app; no public timeline has been announced.
OpenAI GPT-Live voice model launches with sub-300ms latency — AI Weekly GPT-Live is a native voice model powering ChatGPT Voice, eliminating the text-pipeline middleman and adding emotional nuance to real-time speech.
🏢 Industry & Investment
Nvidia is buying Hugging Face for almost $13 billion — The Verge AI In the AI industry’s largest acquisition since the current wave began, Nvidia agreed to purchase Hugging Face — home to over 3 million open-source models and 18 million developers — for $12.93 billion. The deal puts the world’s dominant AI chipmaker in direct control of the ecosystem’s most-used model and dataset hub, raising immediate questions about neutrality and access for non-Nvidia hardware users.
Accel reportedly in talks to lead $1B round for Thinking Machines at $40B valuation — TechCrunch AI The high-profile AI startup, which already tops $100M in annual recurring revenue, is in advanced discussions for a round that would rank among the year’s largest venture deals.
Meta is paying to peek at how you use their latest AI model — TechCrunch AI Meta’s new Muse Spark coding-agent model comes with an explicit 95% discount in exchange for users sharing prompts and outputs to train future models.
South Korea plans $919B AI infrastructure investment targeting 18.4GW by 2035 — AI Weekly The sovereign-AI program targets 8.4GW of data-center capacity by 2029, positioning South Korea as a major state-level actor in global AI infrastructure buildout.
Palo Alto Networks paid $500M for Thrive-backed Console — TechCrunch AI
🛡️ Safety & Security
ChatGPT, Grok, and Claude all went down at the same time — The Verge AI Around 11AM ET, ChatGPT, Grok, and Claude simultaneously experienced widespread errors — an unusual triple outage affecting the three most-used AI assistants at once. All services have since recovered; the cause of the coordinated failure has not been publicly disclosed.
Abliteration.ai is making a business out of removing AI guardrails — TechCrunch AI The startup argues that giving defenders the same unconstrained tools as adversaries improves cybersecurity — a dual-use rationale that has drawn both attention and controversy.
🛠️ Developer Tools & Infrastructure
Nvidia launches free tool that links idle computers into a personal AI data center — The Verge AI Nvidia’s Personal AI Router (PAIR) is open-source software that pools home computers — including non-Nvidia machines like MacBooks — to distribute local inference workloads across tools like Ollama and LM Studio. It’s a notable push to make edge AI more accessible without new hardware.
Perplexity open-sourced Lily, a local inference engine for hybrid compute — AI Weekly Built around Qwen3.6-35B-A3B and optimized for Apple silicon, Lily underpins Perplexity Computer’s hybrid compute model.
AWS adds MiniMax models to Bedrock with 4M token context windows — AI Weekly
Claude Code 2.1.259 — Claude Code
Adds managedMcpServers for organizations to push HTTP/SSE MCP servers to all users, --permission-prompts none for unattended headless hosts, and expanded GitLab CLI (glab) command recognition.
🌐 Applied AI
Google now lets you chat with Gmail, Docs, and Keep — The Verge AI Gmail Live, Docs Live, and Keep Live bring real-time Gemini voice control to Google’s core productivity apps, letting users manage tasks, notes, and emails by speaking naturally — extending the Gemini Live paradigm across the Workspace suite.
Google says its AI weather model is getting better — The Verge AI WeatherNext 3 delivers global forecasts at unprecedented resolution, with improved precipitation prediction; results will surface in Google Search, Maps, and Gemini.
Tether AI unveils TranslatePsy-AfriSLM supporting 18 African languages fully offline — AI Weekly
Ollie is betting its focus on privacy can help it win the AI assistant race — TechCrunch AI
Generated by claude-sonnet-4-6 on 2026-09-03T10:00:00Z