
Moonlight-16B-A3B-Instruct is a 16B-parameter Mixture-of-Experts (MoE) language model developed by Moonshot AI. It is optimized for instruction-following tasks with 3B activated parameters per inference. The model advances the Pareto frontier in performance per FLOP across English, coding, math, and Chinese benchmarks. It outperforms comparable models like Llama3-3B and Deepseek-v2-Lite while maintaining efficient deployment capabilities through Hugging Face integration and compatibility with popular inference engines like vLLM12.
Modalities
Context
8K
Released
Feb 28, 2025
Knowledge Cutoff
Jun 2024
Token volume and request traffic to this model over time.
Moonlight-16B-A3B-Instruct is a 16B-parameter Mixture-of-Experts (MoE) language model developed by Moonshot AI. It is optimized for instruction-following tasks with 3B activated parameters per inference. The model advances the Pareto frontier in performance per FLOP across English, coding, math, and Chinese benchmarks.
Moonlight 16B A3B Instruct has a 8,192 token context window.
Kimi K3, Kimi K2.7 Code, Kimi K2.6 and 4 more are other text models from MoonshotAI.
Moonlight 16B A3B Instruct was released on February 28, 2025. Its knowledge cutoff is June 30, 2024.