ABH3PO · 4w 50-60 for whatparameter models? Even for slightly smaller ones? Hitch @hitch 1781925568 Im getting 50tk/s with qwen3.6 35B FP8 on vllm with speculative decoding. 250k tk context window.Have to do a lot of configs to get things working right 1