

The open-source era of 1M context intelligence
DeepSeek-V4 Preview is a new series of highly efficient MoE language models, featuring V4-Pro (1.6T params) and V4-Flash (284B params). Both models support a 1 million token context window by default, utilizing a novel hybrid attention architecture to drastically reduce compute and memory costs.
還沒有評論,來搶沙發吧!
X 上關於 DeepSeek-V4 的真實討論
去 X 發文