A sophisticated text-based Mixture-of-Experts (MoE) model featuring 21B total parameters with 3B activated per token, delivering exceptional multimodal understanding and generation through heterogeneous MoE structures and modality-isolated routing. Supporting an extensive 131K token context length, the model achieves efficient inference via multi-expert parallel collaboration and quantization, while advanced post-training techniques including SFT, DPO, and UPO ensure optimized performance across diverse applications with specialized routing and balancing losses for superior task handling.
Modalities
Context
131K
Released
Aug 12, 2025
Knowledge Cutoff
Mar 2025
A sophisticated text-based Mixture-of-Experts (MoE) model featuring 21B total parameters with 3B activated per token, delivering exceptional multimodal understanding and generation through heterogeneous MoE structures and modality-isolated routing.
ERNIE 4.5 21B A3B has a 131,072 token context window.
ERNIE 4.5 VL 424B A47B is another text model from Baidu.
ERNIE 4.5 21B A3B was released on August 12, 2025. Its knowledge cutoff is March 31, 2025.
Token volume and request traffic to this model over time.