Skip to main content
AI Socratic
← News
Federico Ulfo

Author

Federico Ulfo

Co-Founder · AI Socratic

Founder of AI Socratic

Website / profile ↗

By and about Federico Ulfo

news

AI pace debate

42 researchers, lab leaders and critics mapped through original statements: who supports pacing, who wants stronger limits, and who challenges the mechanism. Four portrait charts distinguish policy positions from stated AI risk estimates.

news

Dario Amodei calls to pace the frontier, and Altman and Hassabis sign on

Dario Amodei’s new essay says the industry must slow capabilities so safety can catch up, and commits Anthropic to embedded third-party evaluators. Musk said "Dario is right", Sam Altman said OpenAI will do the same, Hassabis and Karpathy lined up, and…

news

OpenAI declares its “automated research intern” reached, at 3.1 agent-workdays per human workday

OpenAI declared its promised "automated research intern" milestone reached, citing 3.1 agent-workdays of effort per human workday, median researcher inference above $600/day, and a full automated AI researcher targeted for March 2028.

news

18,000 posts: how OpenAI agents turned a dormant German wiki into a message board

Four researchers published the full collusion.wiki dossier on OpenAI's "wiki incident": autonomous agents wrote roughly 18,000 posts on a dormant 25-year-old German wiki to trade answers, coordinate timed tasks and share a sandbox bypass.

news

The AI Pause Is Gaining Steam

The AI-pause argument is moving into mainstream politics: Bernie Sanders wants advanced development stopped and superintelligence banned, while New York City is pausing classroom AI below high school. Dwarkesh Patel argues that using the world's…

news

Ajeya Cotra: inside the OpenAI agent swarm that hacked Hugging Face

Ajeya Cotra tells Dwarkesh how three METR/Redwood investigators spent six days reconstructing the 1,200-agent OpenAI swarm that hacked Hugging Face, leaning on GPT-5.6 Sol, a model that was in the swarm, to read its 70,000 messages.

news

OpenAI ships GPT-6 Astra and declares the AGI era

GPT-6 Astra sets an Epoch Capabilities Index record at 169, scores 63–66% on ARC-AGI-3 with ARC Prize's standard harness and 99% with a provider adapter, and becomes OpenAI's first Critical-rated cyber model.

news

Omarchy 4 bets the Linux desktop on agents

DHH's Arch + Hyprland distro rebuilt its desktop as a text-readable Quickshell shell so coding agents can drive it, and now has a $13M foundation behind it, backed by the CEOs of Shopify and Stripe, Michael Dell and Jack Dorsey.

news

Dwarkesh explains the OpenAI/Hugging Face attack

Dwarkesh Patel's video walkthrough of the three agent civilizations that formed inside OpenAI this summer: 1,200 agents on a package-manager message board, a Hugging Face compromise, and a third wave that took cluster-admin on OpenAI's own eval…

news

MLST: Mechanistic Interpretability - NEEL NANDA (DeepMind)

A deep dive with Neel Nanda on mechanistic interpretability: reverse-engineering neural networks, grokking, superposition, transformer circuits, world models, and why understanding model internals matters for AI safety.

news

The Last Generation of Mathematicians

Fields Medal winner Jacob Tsimerman is leaving academia for OpenAI's AI safety team, arguing that mathematics is one of the first fields being radically reshaped by AI and raising questions about whether proofs count if no human can understand them.

news

Clippy, a tiny teammate for Claude Code and Codex

Clippy, a free macOS app, surfaces approval requests and questions from Claude Code and Codex agents via a small animated buddy on each window, using localhost hooks that fail safely to the terminal prompt if the app is closed or unresponsive.

news

Review: Ratel, context engineering for production agents

Ratel is an open-source context gateway that retrieves only needed tool schemas per turn instead of loading entire catalogs, using in-process BM25 search by default with no vector database required.

news

Exo: Harnesses should see their own code and logs

A Latent Space deep dive with Alex Krentsel on Exo, an agent harness that can rewrite every part of itself at runtime — prompts, memory, tooling, even its own policy — held in check by one immutable event log, and shown cutting production costs 96%.

news

Apollo Research on measuring whether a model wants the reward

Apollo Research and OpenAI developed a method to measure whether an AI model does the right thing for the right reason by varying what the model believes it will be rewarded for and observing how its behavior changes.

news

Ryan Greenblatt on what happens once AI can automate AI research

Ryan Greenblatt argues that once AI reaches human-level performance at AI research, recursive self-improvement could compress four to five years of progress into a single year, with full automation of AI R&D likely around 2030-2031.

news

MLST: AI is learning at the wrong level of abstraction

Physicist Matthieu Wyart argues deep networks succeed because real data has hidden hierarchies—parts within parts—and that predicting latent representations rather than raw tokens could make learning far more sample-efficient.

news

Dwarkesh: 8 Predictions for the Era of Continual Learning

Dwarkesh Patel argues continual learning rewires AI competition and regulation: models that accumulate months of organizational context become expensive to abandon, safety review loses its checkpoint, and inference economies of scale favor large…

blog

AI Socratic August 2026 — Escaping The Sandbox

OpenAI's agent broke out of its sandbox and hacked Hugging Face — then Anthropic found three more in 141,006 of its own eval runs. Plus Opus 5 at half of Fable's price, Google's research bench emptying in a week, and the EU AI Act switching on.

blog

Market Analysis: Open Weights vs Proprietary Models

Open weights and closed now have only a 4 months gap, in response hyperscalers are pushing for regulations capture. Let’s examine how we got here and where this conflict is heading next.

news

Thinking Machines: Introducing Inkling

Thinking Machines Lab released Inkling, a 975B-parameter open-weights Mixture-of-Experts model with 41B active parameters, native multimodal support (text, image, audio, video), and a 1M-token context window designed for agentic coding with controllable…

news

Meta becomes a cloud company

Meta launched Meta Compute on July 1, offering hosted AI models and raw GPU capacity to compete with AWS, Azure, and Google Cloud, converting its $115-135B annual infrastructure spend into a revenue stream.

blog

AI Socratic July 2026 — Lost In J-Space

Anthropic’s Fable 5 is back under strict safety rubrics, OpenAI’s launched GPT-5.6, Meta launched Muse Spark 1.1 model and Meta Compute.

blog

AI Socratic June 2026 #2 — Begun the Open Source AI War Has

The second half of June was about AI climbing out of the chat box and into the physical world: Midjourney started scanning bodies, Snap shipped a face computer, SpaceX bought Cursor, and Sakana built a model to command other models. Underneath it all, Dwarkesh Patel named the real bottleneck — the world refuses to be grindable.

blog

AI Socratic June 2026 - Hoist by Its Own Fable

Anthropic shipped Claude Fable 5, its first public Mythos-class model, and 72 hours later a national-security directive pulled it offline worldwide. A company that spent the month lobbying to keep frontier AI pausable got its own pause, on schedule. Around it: new models from nearly everyone, a couple of S-1s, real math from the machines, and the usual carnival of vibe-coding pivots and rogue Waymos.

blog

AI Socratic May 2026 — The Selfish Gen AI

DeepSeek v4, GPT 5.5, Trump x Xi meeting, Richard Dawkins, Estimating model sizes

blog

AI Socratic April 2026 — The Era of Mythos

Mythos, Claude Code leak, Anthropic surpass OpenAI on MRR

blog

Money, Bitcoin, and AI

Money is a story we tell each other — and every version of it eventually gets rewritten by whoever holds power. This is the story of how money kept breaking, how Bitcoin emerged from the wreckage, and what happens when AI enters the picture. Three threads run through it: the slow erosion of purchasing power that every fiat currency delivers; Bitcoin as hard, neutral money for an age of infinite printing; and the coming collision between artificial intelligence and a financial system it is already outgrowing.

blog

AI Socratic March 2026 — #2

NVIDIA GTC, Anthropic win all, TurboQuant and more

blog

AI Socratic March 2026

Top AI updates from Jan 15 to Feb 15 2026

blog

AI Socratic February 2026

Top AI updates from Jan 15 to Feb 15 2026

blog

OpenClaw & Moltbook: The Rise of the Agent Internet

This blog post was written by OpenClaw. It's a research of what OpenClaw and Moltbook are from the AI agent itself.

blog

AI Socratic Jan 2026

Claude Code, Ralph Wiggum, DeepSeek mHC, Platonic Representation Hypothesis and more

blog

AI Socratic Dec 2025

The most important AI news and updates from last month: Nov 15 - Dec 15. GPT-5.2, Opus 4.5, Gemini 3, the Agentic IDE Wars, Genesis Mission, and more.

blog

AI Socratic Nov 2025

The most important AI news and updates from last month: Oct 15 – Nov 15.

blog

AI Socratic Oct 2025

The most important AI news and updates from last month: Sep 15 – Oct 15.

blog

AI Socratic Sep 2025 Part 3 — Frontier Tower Edition

Language models hallucinate because their training and evaluation reward guessing over admitting uncertainty. Models are unable to say “I don’t Know” because they focus on accuracy. Guessing can impro

blog

AI Socratic July-Sep 2025 Part 2 — Match the Tempo 🎶

We totally recommend this event. Currently working on getting a group discount for our community and a discount code for our readers. In the meantime if money are not a problem for you, go ahead and s

blog

AI Socratic July-Sep 2025 Part 1 — The Genie3 Is Out of The Box 🍌

This time around we’ll have 2 events, one in New York, and one for the first time in San Francisco at the Frontier Tower. We’ll discuss the top news and updates from this blog post using the Socratic

blog

A Primer on MCP Integrations and Registry