All models

Gemma 3 27B IT

byGoogleGoogle· 12 Mar 2025
General purposeMultimodal

Aside from the points mentioned in our post, the Gemma release highlights some tricks used by labs: Knowledge Distillation using a big teacher model, a (5:1) local / global attention layer ratio. The latter is a configuration outlined by Noam Shazeer during his time at Character.AI. Apart from the instruction models, Google also releases the pre-trained base model.

Specs
Params27B
LicenseGemma
Capability · Artificial Analysis
AA Index
2.0
Adoption · Hugging Face
RAM score
Relative Adoption Metric: N/A. This model predates RAM history, which begins July 5, 2025 (2025-07-05).
Hugging Face Downloads
402.4K
last 30d
16.9M
all time
HF Likes
2K

Relative Adoption Metric: N/A. This model predates RAM history, which begins July 5, 2025 (2025-07-05).

Inference · OpenRouter
Tokens/Day
below top 50 this week
Peak Tokens/Day
9.4B
28 Jun 2026
Peak Rank
#19
30 Mar 2025

OpenRouter publishes daily token totals for its ~50 most-served models. Days missing from the chart mean Gemma 3 27B IT fell below that cutoff — not zero usage. OpenRouter logs usage separately per dated model version and per variant (like ":free"); the hub combines provider variants while keeping separately cataloged releases distinct.

Related Models