6 articles about DeepSeek
DeepSeek releases V4.1-Flash, a 552B asymmetric MoE that outperforms V4-Pro on benchmarks at lower cost, with a KV cache needing just 1/4 the HBM of the previous generation.
DeepSeek quietly ships deepseek-v4-flash-vision-exp, an experimental vision variant of its V4-Flash model that accepts image input, alongside V4-Pro and V4-Flash point updates.
DeepSeek's V4-Flash at $0.14/M input and $0.28/M output tokens topped global token usage at 7.1 trillion tokens weekly. A research firm found it over 100x cheaper to run than Claude Fable 5, with nine of the top ten models now Chinese.
DeepSeek officially released V4-Flash with enhanced autonomous agent capabilities and reduced API costs, sustaining China's aggressive frontier pace in the AI price war. The model targets agentic workloads at significantly lower prices than US competitors.
OpenAI cut GPT-5.6 Luna API prices by 80% to $0.20 per million input tokens, matching DeepSeek's Chinese pricing. The move signals intensifying competition in the AI API market and a race to the bottom on inference costs.
Chinese AI company DeepSeek launches V4 model with superior efficiency, available under MIT License.