DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

AITopTools Editorial TeamSeptember 10, 2026

What changed

DeepSeek released DeepSeek-V4.1-Flash, an AI model that can work with up to 1 million units of text or other input at once. It also uses new memory-saving methods intended to make repeated processing of very long inputs more efficient.

What this means for you

Developers building systems for long documents or extended, multi-step tasks may be able to handle larger inputs with less strain on computing resources. The feed does not specify where the model is available or what it costs, so access remains unclear.

Related AI tools

Explore directory listings connected to the products, companies, and workflows in this story.

Related AI news

Read Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents
New AI featuresSep 15, 2026

Google Releases Gemini 3.8 Live and 3.8 Live Extended Thinking for Production Grade Voice Agents

The models are available now in Google’s Gemini API and AI Studio, allowing developers to build voice applications at $0.005 per minute for audio input. Generated audio includes Google DeepMind’s SynthID watermark, which identifies it as AI-created.

MarkTechPostSee why it matters
Read AI for everyone in every language
New AI featuresSep 15, 2026

AI for everyone in every language

The announcement does not identify a specific product, launch date, or current user access. If delivered, the work could improve language support for people whose languages are poorly served by existing translation tools.

Google AISee why it matters