Skip to main content
AI Socratic
← News
M

Company / organization

MiniMax

Website / profile ↗

By and about MiniMax

news

World Labs' Atlas generates a minute of 1440p video under exact camera control

Fei-Fei Li's World Labs launched Atlas, a world model that generates up to a minute of 1440p video along a precise camera path from one reference image, with raters preferring its camera control 75-94% of the time over rivals like Seedance 2.5.

news

Agent-generated kernels cut Qwen-Image serving latency 42.3%

Baseten engineer Brian Li reports an agentic kernel-development framework cut serving latency 42.3% on Qwen-Image and 15.2% on FLUX.2, atop an already human-tuned SGLang stack on NVIDIA B300 GPUs.

news

Chinese labs converge on one architecture: 3:1 linear attention and a 2,048-token budget

Z.ai's GLM-5.3-Flash and Alibaba's Qwen3.8-Flash-Next, released a day apart, independently converged on the same recipe: 3:1 linear attention, a 2,048-token attention budget and four-branch gated residuals, while MiniMax dissents and keeps full attention.

news

An 11x price spread for the same open-weight model on OpenRouter

Architect CEO Brett Harrison found an 11x price gap between OpenRouter's cheapest and priciest host of DeepSeek V4 Flash — Baidu ran it at $0.049 per million tokens and 124 tokens/second while 26 of 30 rivals were both pricier and slower.

news

Darkbloom's idle-Mac inference network doubles to 499 nodes in 60 hours

Darkbloom, Eigen Labs' idle-Mac inference network, grew to 499 nodes (432 hardware-attested) in 60 hours, up from 389 a day earlier, with utilization at just 8%.

news

Alibaba's Scroll drops context compaction and beats the best long-horizon agent by 37.4 points

Alibaba researchers unveiled Scroll, a context manager that skips compaction entirely and has the model write Python to retrieve what it needs, scoring 94.8% on LongMemEval_S, 73.1% on BEAM_10M and 86.7% on LOCA_256K with Qwen3.8-Max.

news

MiniMax M2.7

MiniMax released M2.7 on June 16, 2026, an open-weights model designed for autonomous task execution with multi-agent collaboration, scoring 56.2% on SWE-Pro and 1495 ELO on GDPval-AA.

blog

AI Socratic June 2026 - Hoist by Its Own Fable

Anthropic shipped Claude Fable 5, its first public Mythos-class model, and 72 hours later a national-security directive pulled it offline worldwide. A company that spent the month lobbying to keep frontier AI pausable got its own pause, on schedule. Around it: new models from nearly everyone, a couple of S-1s, real math from the machines, and the usual carnival of vibe-coding pivots and rogue Waymos.

news

MiniMax: M3

MiniMax released M3, an open-weight multimodal model with 1M-token context and a sparse attention architecture that reduces per-token compute by ~95% at full context; it scores 59.0% on SWE-Bench Pro, outperforming GPT-5.5 in the company's own testing.

blog

AI Socratic March 2026

Top AI updates from Jan 15 to Feb 15 2026

news

StepFun's Step 3.5 Flash

StepFun released Step 3.5 Flash, a sparse mixture-of-experts model with 196B total parameters but only 11B activated per token, designed to run on 128GB of memory and trained with the Muon optimizer.

news

Chinese labs accused of distilling Claude models

Anthropic says DeepSeek, Moonshot AI, and MiniMax used over 24,000 fraudulent accounts to generate 16 million Claude exchanges for model distillation.