DeepSeek Ships V4.1-Flash, a 763-Billion-Parameter Model That Cuts KV-Cache Memory to a Quarter of Its Predecessor's
DeepSeek's new V4.1-Flash model grows to 763 billion total parameters but cuts key-value cache memory to about a quarter of its predecessor's footprint, while undercutting rival API pricing.