The Chinese AI developer's V4.1 Flash model uses a new “Causal-Encoder-Decoder” architecture, built on a 552 billion-parameter framework with a Mixture-of-Experts (MoE) design. DeepSeek claims it outperforms its previous flagship while cutting inference costs and boosting speeds.