Gemma 4 (E2B / E4B / 12B)
#1 in small on-device llms · text, images and audio on a phone, under an apache licence google took three generations to arrive at.
verdictthe only genuinely multimodal family that runs at phone scale, now on plain apache 2.0 — the clearest default in this category.
| NVIDIA Nemotron 3 Nano 4B | Gemma 4 (E2B / E4B / 12B) | |
|---|---|---|
| price | free (nemotron open model license) | free (apache 2.0) |
| free tier | yes | yes |
| params | 3.97b | ~2b / ~4.5b / 12b |
| licence | nemotron open model license | apache 2.0 |
| ram at 4-bit | ~2.5gb | 1.5-8gb by size |
| context | 262k | 128k confirmed |
| runs on | laptop cpu, no gpu | phone to 16gb laptop |
switch foranything that needs to see or hear as well as read, on hardware you already carry.
- +text, image and audio at every size in the family
- +plain apache 2.0, replacing three generations of custom terms
- +e2b runs on phone-class hardware
- +one family spanning phone, laptop and desktop
- −12b variant needs real gpu or unified memory to feel fast
- −256k context claim unverified per size
- −trades reasoning depth for edge footprint
- −google's prohibited-use page still confuses the licensing story