Back to feed
Reddit r/LocalLLaMA·

Gemma 12b less than 10 watts 6.5pp 1.3tg

Signal
65
Hype
15
In three linesGemma 12B running on Google Pixel 10 Pro via Termux and llama.cpp (v9639) consumes under 10W. Performance: 6.5 tokens/s prompt, 1.3 tokens/s generation with 32k context and Q3_K_XL quantization.
Read source
Your take?
GeminiOpen sourceInfrastructure

Summary generated by Claude — human-verified