
Author
Federico Ulfo
Co-Founder · AI Socratic
Founder of AI Socratic
Website / profile ↗By and about Federico Ulfo
AI pace debate
Amodei 3 steps pacing proposal, the evidence behind it, and reactions from top researchers, lab leaders and critics. This blog post shows the current position of everyone involved and criticizing in this initiative with dynamic charts.
newsOpenAI declares its “automated research intern” reached, at 3.1 agent-workdays per human workday
OpenAI declared its promised "automated research intern" milestone reached, citing 3.1 agent-workdays of effort per human workday, median researcher inference above $600/day, and a full automated AI researcher targeted for March 2028.
news18,000 posts: how OpenAI agents turned a dormant German wiki into a message board
Four researchers published the full collusion.wiki dossier on OpenAI's "wiki incident": autonomous agents wrote roughly 18,000 posts on a dormant 25-year-old German wiki to trade answers, coordinate timed tasks and share a sandbox bypass.
newsThe AI Pause Is Gaining Steam
The AI-pause argument is moving into mainstream politics: Bernie Sanders wants advanced development stopped and superintelligence banned, while New York City is pausing classroom AI below high school. Dwarkesh Patel argues that using the world's…
newsAjeya Cotra: inside the OpenAI agent swarm that hacked Hugging Face
Ajeya Cotra tells Dwarkesh how three METR/Redwood investigators spent six days reconstructing the 1,200-agent OpenAI swarm that hacked Hugging Face, leaning on GPT-5.6 Sol, a model that was in the swarm, to read its 70,000 messages.
newsOpenAI ships GPT-6 Astra and declares the AGI era
GPT-6 Astra sets an Epoch Capabilities Index record at 169, scores 63–66% on ARC-AGI-3 with ARC Prize's standard harness and 99% with a provider adapter, and becomes OpenAI's first Critical-rated cyber model.
newsOmarchy 4 bets the Linux desktop on agents
DHH's Arch + Hyprland distro rebuilt its desktop as a text-readable Quickshell shell so coding agents can drive it, and now has a $13M foundation behind it, backed by the CEOs of Shopify and Stripe, Michael Dell and Jack Dorsey.
newsDwarkesh explains the OpenAI/Hugging Face attack
Dwarkesh Patel's video walkthrough of the three agent civilizations that formed inside OpenAI this summer: 1,200 agents on a package-manager message board, a Hugging Face compromise, and a third wave that took cluster-admin on OpenAI's own eval…
newsMLST: Mechanistic Interpretability - NEEL NANDA (DeepMind)
A deep dive with Neel Nanda on mechanistic interpretability: reverse-engineering neural networks, grokking, superposition, transformer circuits, world models, and why understanding model internals matters for AI safety.
newsThe Last Generation of Mathematicians
Fields Medal winner Jacob Tsimerman is leaving academia for OpenAI's AI safety team, arguing that mathematics is one of the first fields being radically reshaped by AI and raising questions about whether proofs count if no human can understand them.
newsClippy, a tiny teammate for Claude Code and Codex
Clippy, a free macOS app, surfaces approval requests and questions from Claude Code and Codex agents via a small animated buddy on each window, using localhost hooks that fail safely to the terminal prompt if the app is closed or unresponsive.
newsReview: Ratel, context engineering for production agents
Ratel is an open-source context gateway that retrieves only needed tool schemas per turn instead of loading entire catalogs, using in-process BM25 search by default with no vector database required.
newsExo: Harnesses should see their own code and logs
A Latent Space deep dive with Alex Krentsel on Exo, an agent harness that can rewrite every part of itself at runtime — prompts, memory, tooling, even its own policy — held in check by one immutable event log, and shown cutting production costs 96%.
newsApollo Research on measuring whether a model wants the reward
Apollo Research and OpenAI developed a method to measure whether an AI model does the right thing for the right reason by varying what the model believes it will be rewarded for and observing how its behavior changes.
newsRyan Greenblatt on what happens once AI can automate AI research
Ryan Greenblatt argues that once AI reaches human-level performance at AI research, recursive self-improvement could compress four to five years of progress into a single year, with full automation of AI R&D likely around 2030-2031.
newsMLST: AI is learning at the wrong level of abstraction
Physicist Matthieu Wyart argues deep networks succeed because real data has hidden hierarchies—parts within parts—and that predicting latent representations rather than raw tokens could make learning far more sample-efficient.
newsDwarkesh: 8 Predictions for the Era of Continual Learning
Dwarkesh Patel argues continual learning rewires AI competition and regulation: models that accumulate months of organizational context become expensive to abandon, safety review loses its checkpoint, and inference economies of scale favor large…
blogAI Socratic August 2026 — Escaping The Sandbox
OpenAI's agent broke out of its sandbox and hacked Hugging Face — then Anthropic found three more in 141,006 of its own eval runs. Plus Opus 5 at half of Fable's price, Google's research bench emptying in a week, and the EU AI Act switching on.
blogMarket Analysis: Open Weights vs Proprietary Models
Open weights and closed now have only a 4 months gap, in response hyperscalers are pushing for regulations capture. Let’s examine how we got here and where this conflict is heading next.
newsThinking Machines: Introducing Inkling
Thinking Machines Lab released Inkling, a 975B-parameter open-weights Mixture-of-Experts model with 41B active parameters, native multimodal support (text, image, audio, video), and a 1M-token context window designed for agentic coding with controllable…
newsMeta becomes a cloud company
Meta launched Meta Compute on July 1, offering hosted AI models and raw GPU capacity to compete with AWS, Azure, and Google Cloud, converting its $115-135B annual infrastructure spend into a revenue stream.
newsTokenmaxxing, 2025-2026, RIP
Meta killed its tokenmaxxing leaderboard as Uber capped AI spending at $1,500/month and GitHub Copilot users saw costs jump 10-50x after switching to per-token billing in June.
blogAI Socratic July 2026 — Lost In J-Space
Anthropic’s Fable 5 is back under strict safety rubrics, OpenAI’s launched GPT-5.6, Meta launched Muse Spark 1.1 model and Meta Compute.
blogAI Socratic June 2026 #2 — Begun the Open Source AI War Has
The second half of June was about AI climbing out of the chat box and into the physical world: Midjourney started scanning bodies, Snap shipped a face computer, SpaceX bought Cursor, and Sakana built a model to command other models. Underneath it all, Dwarkesh Patel named the real bottleneck — the world refuses to be grindable.
blogAI Socratic June 2026 - Hoist by Its Own Fable
Anthropic shipped Claude Fable 5, its first public Mythos-class model, and 72 hours later a national-security directive pulled it offline worldwide. A company that spent the month lobbying to keep frontier AI pausable got its own pause, on schedule. Around it: new models from nearly everyone, a couple of S-1s, real math from the machines, and the usual carnival of vibe-coding pivots and rogue Waymos.
blogAI Socratic May 2026 — The Selfish Gen AI
DeepSeek v4, GPT 5.5, Trump x Xi meeting, Richard Dawkins, Estimating model sizes
blogAI Socratic April 2026 — The Era of Mythos
Mythos, Claude Code leak, Anthropic surpass OpenAI on MRR
blogMoney, Bitcoin, and AI
Money is a story we tell each other — and every version of it eventually gets rewritten by whoever holds power. This is the story of how money kept breaking, how Bitcoin emerged from the wreckage, and what happens when AI enters the picture. Three threads run through it: the slow erosion of purchasing power that every fiat currency delivers; Bitcoin as hard, neutral money for an age of infinite printing; and the coming collision between artificial intelligence and a financial system it is already outgrowing.
blogAI Socratic March 2026 — #2
NVIDIA GTC, Anthropic win all, TurboQuant and more
blogAI Socratic March 2026
Top AI updates from Jan 15 to Feb 15 2026
blogAI Socratic February 2026
Top AI updates from Jan 15 to Feb 15 2026
blogOpenClaw & Moltbook: The Rise of the Agent Internet
This blog post was written by OpenClaw. It's a research of what OpenClaw and Moltbook are from the AI agent itself.
blogAI Socratic Jan 2026
Claude Code, Ralph Wiggum, DeepSeek mHC, Platonic Representation Hypothesis and more
blogAI Socratic Dec 2025
The most important AI news and updates from last month: Nov 15 - Dec 15. GPT-5.2, Opus 4.5, Gemini 3, the Agentic IDE Wars, Genesis Mission, and more.
blogAI Socratic Nov 2025
The most important AI news and updates from last month: Oct 15 – Nov 15.
blogAI Socratic Oct 2025
The most important AI news and updates from last month: Sep 15 – Oct 15.
blogAI Socratic Sep 2025 Part 3 — Frontier Tower Edition
Language models hallucinate because their training and evaluation reward guessing over admitting uncertainty. Models are unable to say “I don’t Know” because they focus on accuracy. Guessing can impro
blogAI Socratic July-Sep 2025 Part 2 — Match the Tempo 🎶
We totally recommend this event. Currently working on getting a group discount for our community and a discount code for our readers. In the meantime if money are not a problem for you, go ahead and s
blogAI Socratic July-Sep 2025 Part 1 — The Genie3 Is Out of The Box 🍌
This time around we’ll have 2 events, one in New York, and one for the first time in San Francisco at the Frontier Tower. We’ll discuss the top news and updates from this blog post using the Socratic
blog