Inference platform for hosted open-weight models, surfaced in Idapt as a serving provider when an endpoint routes through Parasail.
23 text
Z.ai's frontier reasoning model for long-context agents and software engineering.
MoonshotAI's coding-focused Kimi model for end-to-end software engineering.
MiniMax's multimodal foundation model for long-horizon agentic work.
Open-source 1.6T-parameter MoE with 49B active, built for advanced reasoning and long-horizon agents.
Efficiency-optimized DeepSeek V4 at 284B total / 13B active for fast, high-throughput inference.
MoonshotAI's next-generation multimodal model built for long-horizon coding and multi-agent orchestration.
Zhipu AI's most capable model with a major leap in coding and long-horizon task performance.
Efficient Mixture-of-Experts Gemma 4 with only 4B active parameters per token.
Google's fourth-generation open dense model with native multimodal understanding.
Efficient vision-language MoE with linear attention and 3B active parameters.
Qwen's largest MoE model with 397B total parameters, 17B active.
MiniMax's most capable model with advanced reasoning and long context.
Zhipu AI's flagship fifth-generation model.
Efficient coding agent with 80B parameters, only 3B active per token.
Qwen's largest vision-language model with 235B MoE architecture.
OpenAI's budget-friendly reasoning model — fast and surprisingly capable.
OpenAI's compact open-weights model for efficient inference.
July 2025 update to Qwen3 235B A22B with improved capabilities.
Latest Mistral Small — efficient 24B model for fast, cost-effective tasks.
Meta's high-capacity multimodal MoE model with 128 experts.
Google's open multimodal model with 128K context and 140+ language support.
Large 72B vision-language model from Qwen 2.5 generation.
Meta's efficient 70B model matching Llama 3.1 405B on key benchmarks.