Model reference · open weights
t5gemma-2 is an open-weight language model from Google. t5gemma-2-1b-1b (BF16) weighs 4.2 GB; the smallest configuration that runs it is RTX 3060 12 GB.
Summary of the google/t5gemma-2-1b-1b model card, 2026-10-01
What it is
| Released by | |
|---|---|
| Released | 2025-10-25 |
| Parameters | 2.1B |
| VRAM | 4.2 GB for the weights |
What it runs on
How much memory each request adds isn't estimated yet for this architecture. The weights need at least the cards below, plus room for the context.
| Card | Weights alone |
|---|---|
| RTX 3060 12 GB | fits |
| RTX 4060 Ti 16 GB | fits |
| RTX 3090 24 GB | fits |
| RTX 4090 24 GB | fits |
| RTX 5090 32 GB | fits |
| L40S 48 GB | fits |
| A100 80 GB | fits |
| H100 80 GB | fits |
| RTX PRO 6000 Blackwell 96 GB | fits |
| DGX Spark (GB10) 128 GB unified | fits |
| H200 141 GB | fits |
| B200 180 GB | fits |
Builds
| Build | Params | Precision | Weights | Smallest setup |
|---|---|---|---|---|
| t5gemma-2-1b-1b (above) ↗ | 2.1B | BF16 | 4.2 GB | RTX 3060 12 GB |
| t5gemma-2-270m-270m ↗ | 786M | BF16 | 1.6 GB | RTX 3060 12 GB |
| t5gemma-2-4b-4b ↗ | 8.9B | BF16 | 15.0 GB | RTX 4090 24 GB |
How it works