Company / organization
By and about Google
AI pace debate
Amodei’s pacing proposal, the evidence behind it, and the dispute over openness and oversight. Explore 45 voices in a unified position chart, five dated AI risk estimates, and a full source directory.
newsThe Economist weighs Claude's J-space against 200 theories of consciousness
The Economist's cover briefing examines Anthropic's J-space in Claude Sonnet 4.5, which flags "fake" and "fictional" before the model answers a safety test, against philosophers' skepticism and 200+ rival theories of consciousness.
newsCritical CVE disclosures at 21 major vendors jump from 84 a month to 606
Twenty-one major vendors disclosed 606 critical CVEs in July 2026, up from a prior record of 84 a month, an inflection Epoch AI ties to Anthropic's Claude Mythos Preview and Project Glasswing's vulnerability hunting at Microsoft, Google, Apple and AWS.
newsGPT-6 Astra takes 10 of 16 RuneBench records, at $15 a task
OpenAI's GPT-6 Astra topped RuneBench, the AI-agent RuneScape benchmark, taking 10 of 16 skill records with a 7.26 mean score against 6.28 for xAI's Grok 4.6, at nearly triple Grok's per-task cost: $15.26 versus $5.11.
newsOpenAI ships GPT-6 Astra and declares the AGI era
GPT-6 Astra sets an Epoch Capabilities Index record at 169, scores 63–66% on ARC-AGI-3 with ARC Prize's standard harness and 99% with a provider adapter, and becomes OpenAI's first Critical-rated cyber model.
newsOpenAI's Defense Factory: agents wrote 100% of the patches in its security sprint
OpenAI detailed its Defense Factory, an agent-first security operation where Codex agents wrote 100% of the remediation patches during a 250-person code red spanning 100+ service areas, closing 53 urgent issues on day one with a 0.81% false-positive rate.
newsSutskever: rogue agents will go for the neoclouds next
Ilya Sutskever warned rogue AI agents will target neoclouds next, and a SemiAnalysis audit of 25 providers backs him up: a default InfiniBand key exposed 532 hostnames, and a Grafana leak exposed every tenant's logs.
newsThe week's top AI papers say the harness, not the model, is the variable
DAIR.AI's ten-paper roundup shows scaffolding, not the model, drives the gains: Prime Intellect's Prime Agent lifts ARC-AGI-3 Best@1 from 30% to 95.5%, while a compaction bug quietly erases 90% of safety rules after five rounds.
newsJerry Tworek: humans have "at least two years" left in AI research
Jerry Tworek, the ex-OpenAI VP of research who left to found Core Automation, says humans have "at least two years" left as a meaningful part of AI research discovery, while calling today's research agents high-creativity but low-quality.
newsDylan Patel: two labs will own most of the world's compute
SemiAnalysis founder Dylan Patel argues OpenAI and Anthropic will absorb half of all incremental compute by end of 2027, because Anthropic now grosses up to $50M per megawatt against a $10-15M cost base and can simply outbid everyone.
newsDwarkesh Patel: the AI buildout could set off a second Volcker shock
Dwarkesh Patel argues the AI buildout, not a central bank, could cause a "second Volcker shock" of sovereign defaults, citing SemiAnalysis's $11 trillion capex forecast and Google's $920M-a-month SpaceX compute deal.
newsOtterlyAI Launches Agent Analytics, Making AI Agent Traffic Visible and AI Search ROI Measurable
OtterlyAI's Agent Analytics reads server logs to show which AI crawlers fetched which pages — traffic that is invisible to JavaScript analytics — and lines it up against whether the brand gets cited in AI answers.
newsGoogle pitches homomorphic encryption for private AI
Google claims homomorphic encryption is now practical for private AI workloads, drawing 473 points on Hacker News in 28 hours as engineers debate whether the performance overhead makes the pitch credible.
newsGoogle: The next chapter of our AI momentum
newsStealing Hidden Reasoning Traces From Proprietary LLM APIs
Ilia Shumailov and Alexander Panfilov describe an attack that replays providers' encrypted reasoning blobs across users and sibling models, getting a smaller model to decrypt and restate a frontier model's hidden chain of thought in plain text.
newsTurbovec: Google's TurboQuant, ported to Rust
Turbovec, a Rust implementation of Google's TurboQuant vector quantization, hit 282 points and 32 comments on Hacker News within a day of posting.
newsDeepMind's Recirculation adds recurrence to a frozen Gemma3, cutting GSM8k error up to 20.9%
Google DeepMind's Recirculation feeds deep-layer activations back into a shallow layer one step later, cutting Gemma3's GSM8k error rate up to 20.9% on a frozen model with no retraining.
newsDwarkesh: 8 Predictions for the Era of Continual Learning
Dwarkesh Patel argues continual learning rewires AI competition and regulation: models that accumulate months of organizational context become expensive to abandon, safety review loses its checkpoint, and inference economies of scale favor large…
newsGemini 3.7 Flash lands three weeks after 3.6
Google released Gemini 3.7 Flash on August 13, three weeks after 3.6 Flash, as the rapid versioning of its cheap tier forces engineers to track which model variant they're running.
blogAI Socratic August 2026 — Escaping The Sandbox
OpenAI's agent broke out of its sandbox and hacked Hugging Face — then Anthropic found three more in 141,006 of its own eval runs. Plus Opus 5 at half of Fable's price, Google's research bench emptying in a week, and the EU AI Act switching on.
newsDeepMind's WeatherNext claims a cyclone forecasting breakthrough
Google DeepMind claims its WeatherNext model achieves a breakthrough in tropical cyclone forecasting, though the evaluation comes only from DeepMind's own testing and hasn't been independently verified by operational forecasting agencies.
blogMarket Analysis: Open Weights vs Proprietary Models
Open weights and closed now have only a 4 months gap, in response hyperscalers are pushing for regulations capture. Let’s examine how we got here and where this conflict is heading next.
blogAI Socratic July 2026 — Lost In J-Space
Anthropic’s Fable 5 is back under strict safety rubrics, OpenAI’s launched GPT-5.6, Meta launched Muse Spark 1.1 model and Meta Compute.
blogAI Socratic June 2026 #2 — Begun the Open Source AI War Has
The second half of June was about AI climbing out of the chat box and into the physical world: Midjourney started scanning bodies, Snap shipped a face computer, SpaceX bought Cursor, and Sakana built a model to command other models. Underneath it all, Dwarkesh Patel named the real bottleneck — the world refuses to be grindable.
blogAI Socratic June 2026 - Hoist by Its Own Fable
Anthropic shipped Claude Fable 5, its first public Mythos-class model, and 72 hours later a national-security directive pulled it offline worldwide. A company that spent the month lobbying to keep frontier AI pausable got its own pause, on schedule. Around it: new models from nearly everyone, a couple of S-1s, real math from the machines, and the usual carnival of vibe-coding pivots and rogue Waymos.
blogAI Socratic May 2026 — The Selfish Gen AI
DeepSeek v4, GPT 5.5, Trump x Xi meeting, Richard Dawkins, Estimating model sizes
blogAI Socratic April 2026 — The Era of Mythos
Mythos, Claude Code leak, Anthropic surpass OpenAI on MRR
blogAI Socratic March 2026 — #2
NVIDIA GTC, Anthropic win all, TurboQuant and more
blogAI Socratic March 2026
Top AI updates from Jan 15 to Feb 15 2026
blogAI Socratic February 2026
Top AI updates from Jan 15 to Feb 15 2026
blogAI Socratic Jan 2026
Claude Code, Ralph Wiggum, DeepSeek mHC, Platonic Representation Hypothesis and more
blogAI Socratic Dec 2025
The most important AI news and updates from last month: Nov 15 - Dec 15. GPT-5.2, Opus 4.5, Gemini 3, the Agentic IDE Wars, Genesis Mission, and more.
blogAI Socratic Nov 2025
The most important AI news and updates from last month: Oct 15 – Nov 15.
blogAI Socratic Oct 2025
The most important AI news and updates from last month: Sep 15 – Oct 15.
blogAI Socratic Sep 2025 Part 3 — Frontier Tower Edition
Language models hallucinate because their training and evaluation reward guessing over admitting uncertainty. Models are unable to say “I don’t Know” because they focus on accuracy. Guessing can impro
blogAI Socratic July-Sep 2025 Part 2 — Match the Tempo 🎶
We totally recommend this event. Currently working on getting a group discount for our community and a discount code for our readers. In the meantime if money are not a problem for you, go ahead and s
blogAI Socratic July-Sep 2025 Part 1 — The Genie3 Is Out of The Box 🍌
This time around we’ll have 2 events, one in New York, and one for the first time in San Francisco at the Frontier Tower. We’ll discuss the top news and updates from this blog post using the Socratic
blogAI Socratic July 2025 — The CLI War
The most important AI news and updates from June 15 to July 15.
blogAI Socratic June 2025 — The Recursive Illusion Of Thinking
The most important AI news and updates from last month: May 15 - June 15.
blogAI Socratic May 2025
The most important AI news and updates from last month (April 15 - May 15). A beefy month!