Back to feed
Reddit r/LocalLLaMA·

OSCAR RotationZoo - Offline Spectral Covariance-Aware Rotation for 2-bit KV Cache Quantization

Signal
72
Hype
25
In three linesOSCAR RotationZoo introduces 2-bit KV cache quantization using offline spectral covariance-aware rotation. GGUF models released for Gemma-4-12B, Qwen3-32B, and Qwen3-4B-Thinking with llama.cpp and sglang implementations.
Read source
Your take?
Open sourceInfrastructure

Summary generated by Claude — human-verified