DeepSeek releases V4.1-Flash AI model with drastically reduced costs and memory usage
Chinese AI company DeepSeek launched its V4.1-Flash model on September 10, 2026, featuring a 552-billion-parameter architecture that cuts memory requirements by 75% and offers inference pricing as low as $0.003 per million tokens. The model claims to match or exceed performance of competitors like GPT-6 Astra, Claude Opus 5, and GPT-5.6 Sol on various benchmarks while operating at a fraction of the cost.