Popular repositories Loading
-
Nvidia-Tesla-v100-nvfp4-pcie
Nvidia-Tesla-v100-nvfp4-pcie PublicQwen3.8-27B in native NVFP4/FP8 on 2x PCIe Tesla V100-32GB (SM70): the PCIe runbook for v100-skinny + 1Cat-vLLM, with the 3 fixes that make it work without NVLink. 61-74 tok/s decode, MTP speculati…
-
4xV100-qwen38-flash-next-abliterated-128gb-vram
4xV100-qwen38-flash-next-abliterated-128gb-vram PublicRunbook + benchmarks: Qwen3.8-Flash-Next-ABLITERATED NVFP4 on 4× Tesla V100-32GB (reflashed SXM2→PCIe, 2+2 NVLink + PLX). 1Cat-vLLM 1.5.0, TP4 — 262,144-token context validated, 46 tok/s decode, 12…
Python 2
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.