Gemma 3 27B IT
Aside from the points mentioned in our post, the Gemma release highlights some tricks used by labs: Knowledge Distillation using a big teacher model, a (5:1) local / global attention layer ratio. The latter is a configuration outlined by Noam Shazeer during his time at Character.AI. Apart from the instruction models, Google also releases the pre-trained base model.
Relative Adoption Metric: N/A. This model predates RAM history, which begins July 5, 2025 (2025-07-05).
OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Gemma 3 27B IT fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.


