DeepSeek V4 Arrives After 15 Months: Hybrid Attention Beats Expectations
DeepSeek's V4 preview, released 15 months after V3, introduces million‑token context, a hybrid attention architecture, and two model variants—Pro and Flash—offering significant compute savings, new training techniques, and performance gains that challenge top closed‑source LLMs.
