[AINews] DeepSeek V4 Pro (1.6T-A49B) and Flash (284B-A13B), Base and Instruct — runnable on Huawei Ascend chips
By Unknown
Continuing our coverage from yesterday, DeepSeek released V4 Pro (1.6T-A49B MoE) and Flash (284B-A13B), trained on 32T tokens with FP4, featuring 1M token context via novel Compressed Sparse Attention and Heavily Compressed Attention techniques. The models are roughly Gemini 3.1 / GPT 5.4 / Opus 4.6 level, with both Base and Instruct versions released—a rarity that sets the stage for a potential DeepSeek R2. Notably, the models run on Huawei Ascend chips, carrying significant geopolitical implications.