D
Company / organization
DeepSeek
By and about DeepSeek
news
Gemini 2.0 Flash — Cheapest Model Yet
Google released Gemini 2.0 Flash at $0.40 per million tokens with a 1M-token context window, enabling processing of thousands of PDFs for under $1.
newsCerebras — Fastest Inference Platform
Cerebras achieves 1,200 tokens/second inference speed on DeepSeek-R1-Distill-Llama-70B, claiming 10x faster performance than comparable models and 3x faster than Groq.
newsSFT Memorizes, RL Generalizes
Supervised fine-tuning memorizes training examples while reinforcement learning generalizes to new problems, a distinction DeepSeek's approach to model training highlights.