Our Blog

Blog Index

DeepSeek V4.1-Flash: Open Weights, a Quarter of the Memory, and $0.003 Cache Reads Reshape AI Agent Economics

Posted on 13th Sep 2026 06:03:25 in Artificial Intelligence, Machine Learning

Tagged as: DeepSeek, AI Agents, Open Weight Models, AI Infrastructure, AI Pricing

DeepSeek released V4.1-Flash on September 10, 2026: the smallest model in a new architecture family, it cuts the KV cache to a quarter of its predecessor's HBM footprint, prices off-peak cached input at $0.003 per million tokens, and beats DeepSeek's own V4-Pro on agentic benchmarks under an MIT license.

Read More

whatsapp me