DeepSeek Breaks the "Impossible Trinity" of Large Models: Stronger, Faster, and Cheaper
1 hours ago
Insight: Beating AI Flash News – DeepSeek V4.1 Flash has nearly overhauled its entire architecture this time, aiming to simultaneously enhance performance, boost speed, and cut costs. Stronger: The model features 552 billion backbone parameters, plus an external Engram conditional memory module with 196 billion parameters. Pre-training utilized 45T multi-modal tokens, while post-training integrated a large volume of real agent tasks, tool environments, and failure cases. DeepSWE v1.1 scores 74.2%, surpassing Claude Opus 5 and GPT-5.6 Sol. Faster: The new CED architecture splits the 40-layer model into two halves. When reading prompts, only 8 billion parameters are activated per token, and just 16 billion during generation. Coupled with CSA2 cross-layer reuse and DSpark speculative decoding, the context window has been expanded from 4K to 1M (256x longer), with decoding computation per token increasing by only roughly a quarter. Cheaper: DeepSeek has further optimized KV cache to the maximum. The main cache is switched to FP4, plus cross-layer reuse, leaving global KV for each token at just 890 bytes, about 1/4 of V4 Flash; cache stored long-term on SSD or memory is further reduced to approximately 1/8. V4.1 Flash does not equate "stronger" simply to "more computation". While model scale continues to grow, only a small portion of parameters are adjusted each time; the context window is extended, yet the cache is compressed even smaller. Performance has improved, with speed and costs remaining uncompromised.
Apple's high-end foldable iPhone faces a major test: Can Duo unlock a new growth cycle?
7 minutes ago
Trump warns Iran not to engage in underhanded activities at nuclear facilities, with Strait of Hormuz ship attack incidents continuing to escalate.
7 minutes ago
Flop Labs unveils its latest token economics draft: no VC, no pre-sale, with a total supply of 18.1 billion tokens by Year 10.
7 minutes ago
SGX opens Bitcoin and Ethereum perpetual futures to U.S. institutions, unlocking Asian crypto liquidity.
7 minutes ago
OpenAI embroiled in another math research controversy: Accused of possibly "stealing" a major mathematical proof
7 minutes ago
China and the U.S. are currently in consultations on a framework for reciprocal tariff cuts totaling $30 billion.
7 minutes ago
Hot feeds
A trader profits $448K by monitoring #Binance's new listings!
2024.12.13 17:37:29
Last week, funds have flowed into #Bitcoin, #Ethereum, and #Hyperliquid.
2024.12.16 14:48:36
A $PEPE whale that had been dormant for 600 days transferred all 2.1T $PEPE($52M) to a new address.
2024.12.14 10:35:27
When Elon Musk tweeted about Moltbook, the meme coin MOLT experienced a short-term 30% price surge, hitting a new all-time high of $114 million.
2026.01.31 18:37:29
A smart #AI coin trader made $17.6M on $GOAT, $ai16z, $Fartcoin,$arc.
2025.01.05 16:05:18
A sniper earned 2,277 $ETH ($8.3M) trading $SHIRO within 18 hours!
2024.12.03 23:09:08
MoreHot Articles

How did I turn $1,000 into $30,000 with smart money?
2024.12.09

10 promising AI Agent cryptos
2024.12.05

The 30-Year-Old Entrepreneur Behind Virtual, a Multi-Million Dollar AI Agent Society
2025.01.22

10 smart traders specializing in MEMEcoin trading on Solana
2024.12.09

A trader lost $73.9K trading memecoins in just 3 minutes — a lesson for us all!
2024.12.13

What is $SPORE? Let us take you through the on-chain records to show you how it works.
2024.12.25

