DeepSeek Slashes AI Memory Demands Raising Risks for Chip Giants
DeepSeek's V4.1-Flash model cuts KV-cache HBM requirements by 75% and SSD needs by 87.5% during inference. While training demands remain high, dramatic inference efficiency improvements threaten the …