Gemma 2 27B is Google's largest open-source lightweight model in the Gemma family, offering 27 billion parameters with an 8K token context window.
| Provider | Input / 1M | Output / 1M |
|---|---|---|
| $0.650 | $0.650 | |
| $0.800 | $0.800 |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Gemma 2 27B suits applications where organizations need moderate language capabilities with full control over deployment and data privacy. Its open-source nature makes it ideal for fine-tuning on domain-specific datasets, building custom applications without API costs, and scenarios requiring air-gapped or on-premises deployment. The model works well for content generation, summarization, question answering, and conversational interfaces where the 8K context window is sufficient. Organizations choosing between API convenience and deployment control often select Gemma 2 27B when data sovereignty, customization requirements, or long-term cost predictability outweigh the operational complexity of self-hosting.
Gemma 2 27B pricing varies by provider and pricing type (standard vs batch). Check the pricing table above for current rates across all providers offering hosted inference.
Gemma 2 27B excels in applications requiring local deployment, data privacy, and model customization. It's well-suited for content generation, summarization, conversational interfaces, and scenarios where organizations need to fine-tune on proprietary datasets or maintain full control over their AI infrastructure.
Yes, Gemma 2 27B is open-source with downloadable model weights, allowing local deployment without API dependencies. The 27B parameter size requires substantial hardware resources but can run on high-end consumer GPUs or server infrastructure, giving you complete control over inference and data handling.