Geo News
Community curated by people like you
LatestAICryptoHealthWorld AffairsUS Politics
←Back to story

DeepSeek launches V4.1-Flash model with 552B parameters and a million-token context window

CryptobriefingView source article

Part of story

DeepSeek releases V4.1-Flash AI model with drastically reduced costs and memory usage

Chinese AI company DeepSeek launched its V4.1-Flash model on September 10, 2026, featuring a 552-billion-parameter architecture that cuts memory requirements by 75% and offers inference pricing as low as $0.003 per million tokens. The model claims to match or exceed performance of competitors like GPT-6 Astra, Claude Opus 5, and GPT-5.6 Sol on various benchmarks while operating at a fraction of the cost.

Sep 10, 2026·12 sourcesAI research & benchmarksCompute, chips & AI infrastructureLarge language models (LLMs)AI tools & productsAI foundation modelsAI inference (scaling)
DeepSeek releases V4.1-Flash AI model with drastically reduced costs and memory usage
00

Other sources for this story

Scmp

DeepSeek says new Flash AI model beats Kimi K3 on cyber, coding benchmarks

The-decoder

New Deepseek model V4.1-Flash cuts memory needs for AI agents

Analyticsindiamag

AIM — India's Leading AI & Data Science Media Platform

Thenews

DeepSeek launches V4.1-Flash model highlighting unmatched 400 plus token speeds

Reuters

China's DeepSeek launches V4.1-Flash model | Reuters

Businesstimes

DeepSeek’s low-cost model deals fresh blow to rivals OpenAI, Z.AI; hits memory makers

Venturebeat

DeepSeek-V4.1-Flash debuts with $0.003/1M off-peak cached-input rate and benchmarks eclipsing GPT-5.6 Sol, Claude Opus 5

Techtimes

DeepSeek V4.1-Flash Cuts Agent Memory Costs Fourfold With New Architecture

Siliconangle

DeepSeek releases V4.1-Flash, says it outperforms flagship V4-Pro - SiliconANGLE

Decrypt

DeepSeek's New Model Nearly Matches GPT-6 Astra on Design—at 1.4% of the Cost - Decrypt

Theregister

DeepSeek's new model sets a template for powerful LLMs that run lean