Skip to main content
AI Socratic
← News
Google

Company / organization

Google

Website / profile ↗

By and about Google

news

AI pace debate

Amodei’s pacing proposal, the evidence behind it, and the dispute over openness and oversight. Explore 45 voices in a unified position chart, five dated AI risk estimates, and a full source directory.

news

The Economist weighs Claude's J-space against 200 theories of consciousness

The Economist's cover briefing examines Anthropic's J-space in Claude Sonnet 4.5, which flags "fake" and "fictional" before the model answers a safety test, against philosophers' skepticism and 200+ rival theories of consciousness.

news

Critical CVE disclosures at 21 major vendors jump from 84 a month to 606

Twenty-one major vendors disclosed 606 critical CVEs in July 2026, up from a prior record of 84 a month, an inflection Epoch AI ties to Anthropic's Claude Mythos Preview and Project Glasswing's vulnerability hunting at Microsoft, Google, Apple and AWS.

news

GPT-6 Astra takes 10 of 16 RuneBench records, at $15 a task

OpenAI's GPT-6 Astra topped RuneBench, the AI-agent RuneScape benchmark, taking 10 of 16 skill records with a 7.26 mean score against 6.28 for xAI's Grok 4.6, at nearly triple Grok's per-task cost: $15.26 versus $5.11.

news

OpenAI ships GPT-6 Astra and declares the AGI era

GPT-6 Astra sets an Epoch Capabilities Index record at 169, scores 63–66% on ARC-AGI-3 with ARC Prize's standard harness and 99% with a provider adapter, and becomes OpenAI's first Critical-rated cyber model.

news

OpenAI's Defense Factory: agents wrote 100% of the patches in its security sprint

OpenAI detailed its Defense Factory, an agent-first security operation where Codex agents wrote 100% of the remediation patches during a 250-person code red spanning 100+ service areas, closing 53 urgent issues on day one with a 0.81% false-positive rate.

news

Sutskever: rogue agents will go for the neoclouds next

Ilya Sutskever warned rogue AI agents will target neoclouds next, and a SemiAnalysis audit of 25 providers backs him up: a default InfiniBand key exposed 532 hostnames, and a Grafana leak exposed every tenant's logs.

news

The week's top AI papers say the harness, not the model, is the variable

DAIR.AI's ten-paper roundup shows scaffolding, not the model, drives the gains: Prime Intellect's Prime Agent lifts ARC-AGI-3 Best@1 from 30% to 95.5%, while a compaction bug quietly erases 90% of safety rules after five rounds.

news

Jerry Tworek: humans have "at least two years" left in AI research

Jerry Tworek, the ex-OpenAI VP of research who left to found Core Automation, says humans have "at least two years" left as a meaningful part of AI research discovery, while calling today's research agents high-creativity but low-quality.

news

Dylan Patel: two labs will own most of the world's compute

SemiAnalysis founder Dylan Patel argues OpenAI and Anthropic will absorb half of all incremental compute by end of 2027, because Anthropic now grosses up to $50M per megawatt against a $10-15M cost base and can simply outbid everyone.

news

Dwarkesh Patel: the AI buildout could set off a second Volcker shock

Dwarkesh Patel argues the AI buildout, not a central bank, could cause a "second Volcker shock" of sovereign defaults, citing SemiAnalysis's $11 trillion capex forecast and Google's $920M-a-month SpaceX compute deal.

news

OtterlyAI Launches Agent Analytics, Making AI Agent Traffic Visible and AI Search ROI Measurable

OtterlyAI's Agent Analytics reads server logs to show which AI crawlers fetched which pages — traffic that is invisible to JavaScript analytics — and lines it up against whether the brand gets cited in AI answers.

news

Google pitches homomorphic encryption for private AI

Google claims homomorphic encryption is now practical for private AI workloads, drawing 473 points on Hacker News in 28 hours as engineers debate whether the performance overhead makes the pitch credible.

news

Google: The next chapter of our AI momentum

news

Stealing Hidden Reasoning Traces From Proprietary LLM APIs

Ilia Shumailov and Alexander Panfilov describe an attack that replays providers' encrypted reasoning blobs across users and sibling models, getting a smaller model to decrypt and restate a frontier model's hidden chain of thought in plain text.

news

Turbovec: Google's TurboQuant, ported to Rust

Turbovec, a Rust implementation of Google's TurboQuant vector quantization, hit 282 points and 32 comments on Hacker News within a day of posting.

news

DeepMind's Recirculation adds recurrence to a frozen Gemma3, cutting GSM8k error up to 20.9%

Google DeepMind's Recirculation feeds deep-layer activations back into a shallow layer one step later, cutting Gemma3's GSM8k error rate up to 20.9% on a frozen model with no retraining.

news

Dwarkesh: 8 Predictions for the Era of Continual Learning

Dwarkesh Patel argues continual learning rewires AI competition and regulation: models that accumulate months of organizational context become expensive to abandon, safety review loses its checkpoint, and inference economies of scale favor large…

news

Gemini 3.7 Flash lands three weeks after 3.6

Google released Gemini 3.7 Flash on August 13, three weeks after 3.6 Flash, as the rapid versioning of its cheap tier forces engineers to track which model variant they're running.

blog

AI Socratic August 2026 — Escaping The Sandbox

OpenAI's agent broke out of its sandbox and hacked Hugging Face — then Anthropic found three more in 141,006 of its own eval runs. Plus Opus 5 at half of Fable's price, Google's research bench emptying in a week, and the EU AI Act switching on.

news

DeepMind's WeatherNext claims a cyclone forecasting breakthrough

Google DeepMind claims its WeatherNext model achieves a breakthrough in tropical cyclone forecasting, though the evaluation comes only from DeepMind's own testing and hasn't been independently verified by operational forecasting agencies.

blog

Market Analysis: Open Weights vs Proprietary Models

Open weights and closed now have only a 4 months gap, in response hyperscalers are pushing for regulations capture. Let’s examine how we got here and where this conflict is heading next.

blog

AI Socratic July 2026 — Lost In J-Space

Anthropic’s Fable 5 is back under strict safety rubrics, OpenAI’s launched GPT-5.6, Meta launched Muse Spark 1.1 model and Meta Compute.

blog

AI Socratic June 2026 #2 — Begun the Open Source AI War Has

The second half of June was about AI climbing out of the chat box and into the physical world: Midjourney started scanning bodies, Snap shipped a face computer, SpaceX bought Cursor, and Sakana built a model to command other models. Underneath it all, Dwarkesh Patel named the real bottleneck — the world refuses to be grindable.

blog

AI Socratic June 2026 - Hoist by Its Own Fable

Anthropic shipped Claude Fable 5, its first public Mythos-class model, and 72 hours later a national-security directive pulled it offline worldwide. A company that spent the month lobbying to keep frontier AI pausable got its own pause, on schedule. Around it: new models from nearly everyone, a couple of S-1s, real math from the machines, and the usual carnival of vibe-coding pivots and rogue Waymos.

blog

AI Socratic May 2026 — The Selfish Gen AI

DeepSeek v4, GPT 5.5, Trump x Xi meeting, Richard Dawkins, Estimating model sizes

blog

AI Socratic April 2026 — The Era of Mythos

Mythos, Claude Code leak, Anthropic surpass OpenAI on MRR

blog

AI Socratic March 2026 — #2

NVIDIA GTC, Anthropic win all, TurboQuant and more

blog

AI Socratic March 2026

Top AI updates from Jan 15 to Feb 15 2026

blog

AI Socratic February 2026

Top AI updates from Jan 15 to Feb 15 2026

blog

AI Socratic Jan 2026

Claude Code, Ralph Wiggum, DeepSeek mHC, Platonic Representation Hypothesis and more

blog

AI Socratic Dec 2025

The most important AI news and updates from last month: Nov 15 - Dec 15. GPT-5.2, Opus 4.5, Gemini 3, the Agentic IDE Wars, Genesis Mission, and more.

blog

AI Socratic Nov 2025

The most important AI news and updates from last month: Oct 15 – Nov 15.

blog

AI Socratic Oct 2025

The most important AI news and updates from last month: Sep 15 – Oct 15.

blog

AI Socratic Sep 2025 Part 3 — Frontier Tower Edition

Language models hallucinate because their training and evaluation reward guessing over admitting uncertainty. Models are unable to say “I don’t Know” because they focus on accuracy. Guessing can impro

blog

AI Socratic July-Sep 2025 Part 2 — Match the Tempo 🎶

We totally recommend this event. Currently working on getting a group discount for our community and a discount code for our readers. In the meantime if money are not a problem for you, go ahead and s

blog

AI Socratic July-Sep 2025 Part 1 — The Genie3 Is Out of The Box 🍌

This time around we’ll have 2 events, one in New York, and one for the first time in San Francisco at the Frontier Tower. We’ll discuss the top news and updates from this blog post using the Socratic

blog

AI Socratic July 2025 — The CLI War

The most important AI news and updates from June 15 to July 15.

blog

AI Socratic June 2025 — The Recursive Illusion Of Thinking

The most important AI news and updates from last month: May 15 - June 15.

blog

AI Socratic May 2025

The most important AI news and updates from last month (April 15 - May 15). A beefy month!