Efficient 30B MoE language model with 3B active parameters per token.
Add a second model to compare.
This model hasn't been benchmarked yet.