Skip to content
#

hygon

Here are 12 public repositories matching this topic...

A ~2,000-line miniature of vLLM v1 with a vLLM-style Platform/op-family hardware abstraction layer. One engine, four backends: NVIDIA CUDA, Hygon DCU (DTK/HIP), Ascend NPU, Moore Threads MUSA. Real continuous batching, paged KV cache, TP, CUDA graph — runs Qwen3 end-to-end

  • Updated Oct 8, 2026
  • Python

Add this topic to your repo

To associate your repository with the hygon topic, visit your repo's landing page and select "manage topics."

Learn more