My GLM-5.2-FP8 HGX-H200 SGLang docker deploy config
Signal
45
Hype
15
In three linesDocker deployment config for GLM-5.2-FP8 on HGX-H200 using SGLang. Achieves 70 tokens/s and 262k context by disabling DP and moe-a2a-backend deepep, with mem-fraction-static set to 0.83. Official vLLM recipes incompatible with H200.Read source
Your take?
Summary generated by Claude — human-verified