Pinned Loading
-
nf4-triton-kernel
nf4-triton-kernel PublicAn optimized, high-performance NF4 (NormalFloat 4-bit) dequantization Triton GPU kernel achieving up to 1.41x speedup over bitsandbytes.
Python 6
-
Hyper-transformer
Hyper-transformer PublicA PyTorch research evolution of language models from standard Euclidean transformers to adaptive hybrid-manifold architectures with spiking neural networks.
Python 1
-
xmerge-engine
xmerge-engine PublicMerge LLMs across different architectures and sizes — representation-level merging, not weight-space interpolation. Merge GPT-2, OPT, DistilGPT-2, SmolLM2, and more
Python 1
If the problem persists, check the GitHub status page or contact support.