Highlights
- Pro
Popular repositories Loading
-
-
Qwen3.8-27B-NVFP4-TurboQuant
Qwen3.8-27B-NVFP4-TurboQuant PublicValidated vLLM recipe: Qwen3.8-27B-NVFP4 at full 262,144-token context on a single RTX 5090, GPU-only (no CPU offload).
-
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


