[AI]■ STORY TIMELINE
DEEPSEEK V4.1-FLASH SLASHES AI MEMORY NEEDS
Deepseek released V4.1-Flash, a multimodal model with 552 billion parameters that reduces KV cache memory to 25% of its predecessor. The model matches performance of larger competitors while using only 16 billion active parameters per token.
The Decoder+0m
Deepseek releases V4.1-Flash, a multimodal model with 552 billion parameters that cuts KV cache memory to a quarter of i…