Fast variant of GLM 4.7 for low-latency applications.
Add a second model to compare.
This model hasn't been benchmarked yet.