DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge. The tri-parent design yields strong reasoning performance while running roughly 20 % faster than the original R1 and more than 2× faster than R1-0528 under vLLM, giving a favorable cost-to-intelligence trade-off. The checkpoint supports contexts up to 60 k tokens in standard use (tested to ~130 k) and maintains consistent <think> token behaviour, making it suitable for long-context analysis, dialogue and other open-ended generation tasks.
Modalities
Price
Free
Context
164K
Released
Jul 8, 2025
Knowledge Cutoff
Jul 2024
Token volume and request traffic to this model over time.
DeepSeek-TNG-R1T2-Chimera is the second-generation Chimera model from TNG Tech. It is a 671 B-parameter mixture-of-experts text-generation model assembled from DeepSeek-AI’s R1-0528, R1, and V3-0324 checkpoints with an Assembly-of-Experts merge.
Yes. The pricing shown on this page for DeepSeek R1T2 Chimera is zero, so you are not charged for prompt or completion tokens. Free endpoints are rate limited — see the rate limit docs.
DeepSeek R1T2 Chimera has a 163,840 token context window.
DeepSeek R1T2 Chimera was released on July 8, 2025. Its knowledge cutoff is July 31, 2024.