Google's open-weight model family. 1B/4B/12B/27B sizes, all commercially licensable.
Gemma 3 is the best open-weight model if you're on Google Cloud or you need a tiny-but-capable model for edge deployment.
Who it's for: Google Cloud users, Vertex AI developers, hobbyists.
Four sizes cover everything from a Raspberry Pi (1B) to a single H100 (27B). Pick the model that fits your hardware and budget.
All sizes support 128K tokens โ far longer than Llama 3 (8K) or Mistral Small (32K).
4B, 12B, and 27B sizes accept images natively. Document understanding, OCR, chart parsing โ all in one model.
One-click deploy on Vertex AI, with managed scaling and enterprise SLAs. Best-in-class if you're already on GCP.
Gemma 3 is the best open-weight model if you're on Google Cloud or you need a tiny-but-capable model for edge deployment. The 1B and 4B sizes are unmatched at their weight. For pure on-prem reasoning, Llama 4 or DeepSeek R1 are smarter, but Gemma 3 wins on ecosystem polish.