AI provider offering hosted model inference, including privacy-oriented routes exposed through model routing marketplaces.
22 text
Z.ai's frontier reasoning model for long-context agents and software engineering.
MoonshotAI's coding-focused Kimi model for end-to-end software engineering.
Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.
Efficiency-optimized DeepSeek V4 at 284B total / 13B active for fast, high-throughput inference.
MoonshotAI's next-generation multimodal model built for long-horizon coding and multi-agent orchestration.
Zhipu AI's most capable model with a major leap in coding and long-horizon task performance.
Efficient Mixture-of-Experts Gemma 4 with only 4B active parameters per token.
Google's fourth-generation open dense model with native multimodal understanding.
Mistral's next-gen small model, unifying Magistral's reasoning, Pixtral's vision, and Devstral's coding into one system.
Compact 9B vision-language model from the Qwen3.5 family.
Efficient vision-language MoE with linear attention and 3B active parameters.
Qwen's largest MoE model with 397B total parameters, 17B active.
MiniMax's most capable model with advanced reasoning and long context.
Zhipu AI's flagship fifth-generation model.
MoonshotAI's native multimodal model with state-of-the-art visual coding.
Fast variant of GLM 4.7 for low-latency applications.
Zhipu AI's advanced model with strong reasoning and multimodal support.
Zhipu AI's GLM 4.6 generation model.
Qwen's largest vision-language model with 235B MoE architecture.
Qwen's primary code generation model for software development workflows.
July 2025 update to Qwen3 235B A22B with improved capabilities.
Latest Mistral Small — efficient 24B model for fast, cost-effective tasks.