Sakana AI and Nvidia introduce TwELL, a sparse transformer format achieving over 95% sparsity in feedforward layers that speeds up inference and training by 20% or more while reducing memory and energy use on billion-parameter models.
The weekly AI digest — models, agents, open source, research — plus a monthly round-up. Unsubscribe anytime.