Back to feed
Reddit r/LocalLLaMA·

My GLM-5.2-FP8 HGX-H200 SGLang docker deploy config

Signal
45
Hype
15
In three linesDocker deployment config for GLM-5.2-FP8 on HGX-H200 using SGLang. Achieves 70 tokens/s and 262k context by disabling DP and moe-a2a-backend deepep, with mem-fraction-static set to 0.83. Official vLLM recipes incompatible with H200.
Read source
Your take?
QwenCode generationInfrastructureOpen source

Summary generated by Claude — human-verified