Free Gemma model via API
Updated:
Open Gemma models for experiments and light production. This page is separate from Gemini: it is a different model line.
Read more
Gemma does not inherit Google AI Studio Gemini quotas. Read the limit on the selected host.
26 models found
| № | Model | Power | Provider | Parameters | Context | Limit | |
|---|---|---|---|---|---|---|---|
| 1 |
|
⚖️ | OpenRouter | 26B | 262K | Free tier | Open |
| 2 |
|
⚖️ | OpenRouter | 31B | 262K | Free tier | Open |
| 3 |
|
⚖️ | Together AI | 27B | 66K | Fair Use (По мере нагрузки серверов) | Open |
| 4 |
|
⚡ | Together AI | 1B | 33K | Fair Use (По мере нагрузки серверов) | Open |
| 5 |
|
⚡ | Together AI | 4B | 66K | Fair Use (По мере нагрузки серверов) | Open |
| 6 |
|
⚡ | Together AI | 2B | 8K | Fair Use (По мере нагрузки серверов) | Open |
| 7 |
|
⚡ | Together AI | 9B | 8K | Fair Use (По мере нагрузки серверов) | Open |
| 8 |
|
⚡ | Together AI | ? | 33K | Fair Use (По мере нагрузки серверов) | Open |
| 9 | Medgemma 27B Text It | ⚖️ | Together AI | 27B | 131K | Fair Use (По мере нагрузки серверов) | Open |
| 10 |
|
⚡ | Together AI | ? | 131K | Fair Use (По мере нагрузки серверов) | Open |
| 11 |
|
⚖️ | Together AI | 26B | 262K | Fair Use (По мере нагрузки серверов) | Open |
| 12 |
|
⚡ | Together AI | ? | 131K | Fair Use (По мере нагрузки серверов) | Open |
| 13 |
|
⚡ | Together AI | ? | 33K | Fair Use (По мере нагрузки серверов) | Open |
| 14 |
|
⚖️ | Together AI | 27B | ? | Fair Use (По мере нагрузки серверов) | Open |
| 15 |
|
⚖️ | Together AI | 31B | 262K | Fair Use (По мере нагрузки серверов) | Open |
| 16 |
|
⚖️ | Together AI | 12B | 262K | Fair Use (По мере нагрузки серверов) | Open |
| 17 | Codegemma 1.1 7b | ⚡ | NVIDIA Build | 7B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 18 | Codegemma 7b | ⚡ | NVIDIA Build | 7B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 19 |
|
⚡ | NVIDIA Build | 2B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 20 |
|
⚖️ | NVIDIA Build | 12B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 21 |
|
⚡ | NVIDIA Build | 4B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 22 |
|
⚖️ | NVIDIA Build | 31B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 23 | Recurrentgemma 2b | ⚡ | NVIDIA Build | 2B | ? | Free serverless inference for development; limits and model availability may change. | Open |
| 24 | Diffusiongemma 26b A4b It | ⚖️ | UnoRouter | 26B | 262K | No-card free models use shared capacity and per-model limits; availability can change. | Open |
| 25 |
|
⚖️ | UnoRouter | 26B | 262K | No-card free models use shared capacity and per-model limits; availability can change. | Open |
| 26 |
|
⚖️ | UnoRouter | 31B | 262K | No-card free models use shared capacity and per-model limits; availability can change. | Open |