DeepSeek reran post-training on their small model and its agent scores jumped 7x. Same architecture, same 13B active par...
reactive:deepseek-v4-flash-launch · Aakash Gupta (@aakashgupta) · 2026-07-31
(No summary yet for this item — extraction summaries are still backfilling.)