4 text
Efficient Mixture-of-Experts Gemma 4 with only 4B active parameters per token.
OpenAI's budget-friendly reasoning model — fast and surprisingly capable.
OpenAI's compact open-weights model for efficient inference.
12B model from Mistral and NVIDIA — efficient and multilingual.