Back to feed
Reddit r/LocalLLaMA·

I didn't know it was possible to compile llamacpp to run cuda + vulkan at the same time..

Signal
45
Hype
25
In three linesUser compiles llama.cpp with CUDA and Vulkan enabled simultaneously on W7800. Achieves +10% tokens/sec improvement in decoding with MiniMax-M3-UD-IQ2_M. Tests dual GPU accelerator combination for performance optimization.
Read source
Your take?
Open sourceInfrastructure

Summary generated by Claude — human-verified