Back to feed
Reddit r/LocalLLaMA·

I built a iOS app to benchmark GGUF models on your iPhone/iPad

Signal
72
Hype
35
In three linesGenBench is a free iOS app to download, run, and benchmark GGUF models on iPhone/iPad using llama.cpp + Metal. Measures tok/s, first-token latency, peak memory. Global leaderboard. Supports text and vision models (MiniCPM-V). Examples: SmolLM2 1.7B ~35 tok/s on iPhone 16 Pro, Qwen2.5 3B ~20 tok/s on iPhone 15 Pro.
Read source
Your take?
Open sourceToolsBenchmarksInfrastructure

Summary generated by Claude — human-verified