Back to feed
Reddit r/LocalLLaMA·

1-bit and 1.58 bit LLM Benchmarking on Jetson Orin Nano Super | Bonsai LM

Signal
78
Hype
25
In three linesComprehensive benchmark of Bonsai LM models (1-bit and 1.58-bit, 1.7B–8B) on Jetson Orin Nano Super ($250) using llama.cpp CUDA across 4 power modes. Key findings: 25W is efficiency sweet spot for ≤4B models (47–48% faster than 15W), no thermal throttling observed, Bonsai-1.7B Q1_0 achieves 5.84 tok/J in 237 MB with 26 tok/s.
Read source
Your take?
Open sourceBenchmarksInfrastructure

Summary generated by Claude — human-verified