14 text
Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.
Efficiency-optimized DeepSeek V4 at 284B total / 13B active for fast, high-throughput inference.
MoonshotAI's next-generation multimodal model built for long-horizon coding and multi-agent orchestration.
Google's fourth-generation open dense model with native multimodal understanding.
Efficient vision-language MoE with linear attention and 3B active parameters.
MiniMax's most capable model with advanced reasoning and long context.
Large hybrid reasoning model supporting thinking and non-thinking modes.
OpenAI's budget-friendly reasoning model — fast and surprisingly capable.
OpenAI's compact open-weights model for efficient inference.
Qwen's primary code generation model for software development workflows.
July 2025 update to Qwen3 235B A22B with improved capabilities.
Meta's efficient 70B model matching Llama 3.1 405B on key benchmarks.
Meta's 70B model from the Llama 3.1 generation.
Compact 8B model from Meta's Llama 3.1 series.