cmp-170hx-vllm

(★ 9)

Known-good vLLM recipe: Qwen3.8-27B-Int8 + MTP speculative decoding on NVIDIA CMP 170HX 64GB (Hynix) — 84 tok/s, 262k ctx

File Explorer

  • LICENSE
  • README.md
  • serve-g3-int8df2.sh
  • serve-qwen38-mtp.sh
  • serve-qwen38-nightly.sh

# Use via CDN

jsDelivr

jsDelivr serves any public GitHub repository as a CDN with zero setup. Pick a version and a file to get a ready-to-paste link and snippet.

// repository documentation