Instructions to use Efficient-Large-Model/Sol-Attn-Kernel-Source with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Kernels
How to use Efficient-Large-Model/Sol-Attn-Kernel-Source with Kernels:
# !pip install kernels from kernels import get_kernel kernel = get_kernel("Efficient-Large-Model/Sol-Attn-Kernel-Source") - Notebooks
- Google Colab
- Kaggle
Sol-Attn Kernel Builder Source
This repository contains the Hugging Face Kernel Builder packaging for
Sol-Attn.
The kernel implementation is pinned to NVIDIA's NVlabs/Sana commit
8a26fb0.
The packaged Python sources preserve that implementation. The only source
relocation required by Kernel Hub is converting internal sol_attn.* imports
to package-relative imports, so the kernel remains loadable under the
version-isolated module name assigned by kernels.get_kernel(...).
Published builds are loaded from
Efficient-Large-Model/Sol-Attn:
from kernels import get_kernel
kernel = get_kernel("Efficient-Large-Model/Sol-Attn", version=1)
out = kernel.sol_attn(
q, # Contiguous BF16 CUDA tensor [batch, tokens, heads, 128].
k, # Same shape, dtype, layout, and device as q.
v, # Same shape, dtype, layout, and device as q.
tau=1.0,
thresh_type="exact",
)
See SOURCE.md for provenance and the verification command.
- Downloads last month
- -
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support