Ahmad Al-Dahle says DeepSeek-V4's efficient 1M context is its key bet
Original titleThe most interesting thing about DeepSeek-V4 isn't the benchmarks, it's the bet: efficient ultra-long context is the precondition for tes...
Ahmad Al-Dahle argues that the most interesting part of DeepSeek-V4 is its bet on efficient ultra-long context rather than its benchmarks. He says this is the precondition for test-time scaling and long-horizon agents, and cites 27% of V3's FLOPs at 1M tokens.
The quoted DeepSeek post announces DeepSeek-V4-Pro (1.6T total, 49B active) and DeepSeek-V4-Flash (284B total, 13B active), both open-sourced with 1M context and API access.
The post argues that efficient 1M-token context, not benchmark scores, is the key bet behind DeepSeek-V4's design for test-time scaling and long-horizon agents.
Source: Ahmad Al-Dahle · x.com