Kernels
PyTorch
cuda
structured-linear-algebra
state-space-models
s4
mamba
toeplitz
cauchy
displacement-rank
Instructions to use chrisvoncsefalvay/hotdog with libraries, inference providers, notebooks, and local apps. Follow these links to get started.
- Libraries
- Kernels
How to use chrisvoncsefalvay/hotdog with Kernels:
# !pip install kernels from kernels import get_kernel kernel = get_kernel("chrisvoncsefalvay/hotdog") - Notebooks
- Google Colab
- Kaggle
Update README.md
Browse files
README.md
CHANGED
|
@@ -19,17 +19,7 @@ Acceleration **O**n **G**PUs) is a CUDA kernel bundle for the structured
|
|
| 19 |
linear-algebra primitives that underpin state-space models (S4/S5, Mamba),
|
| 20 |
long-convolution networks (Hyena, H3), and low-displacement-rank layers.
|
| 21 |
|
| 22 |
-
The kernel exposes
|
| 23 |
-
|
| 24 |
-
$$K_{b,j} = \sum_i \frac{v_{b,i}}{\omega_j - \lambda_i},\quad
|
| 25 |
-
y = T(c,r)\,x,\quad
|
| 26 |
-
T(c,r)\,x = b,\quad
|
| 27 |
-
u = T^{-1} e_1,\; v = T^{-1} e_n,\quad
|
| 28 |
-
C = F\,T\,F^{-1},$$
|
| 29 |
-
|
| 30 |
-
$$y = A\,x\;\text{where}\;Z_1 A - A\,Z_{-1} = G H^\top,\;G,H \in \mathbb{R}^{n\times r}$$
|
| 31 |
-
|
| 32 |
-
— with full autograd, BF16/FP16/complex-dtype paths, stream-aware dispatch,
|
| 33 |
and `torch.compile` compliance via `pt2_compliant_tag` +
|
| 34 |
`needs_fixed_stride_order`.
|
| 35 |
|
|
@@ -184,4 +174,11 @@ nix run -L --max-jobs 1 --cores 4 .#build-and-copy
|
|
| 184 |
- Source: https://github.com/chrisvoncsefalvay/hotdog
|
| 185 |
- Hub: https://hf.co/chrisvoncsefalvay/hotdog
|
| 186 |
|
| 187 |
-
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 19 |
linear-algebra primitives that underpin state-space models (S4/S5, Mamba),
|
| 20 |
long-convolution networks (Hyena, H3), and low-displacement-rank layers.
|
| 21 |
|
| 22 |
+
The kernel exposes the usual Toeplitz/Cauchy ops — with full autograd, BF16/FP16/complex-dtype paths, stream-aware dispatch,
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
| 23 |
and `torch.compile` compliance via `pt2_compliant_tag` +
|
| 24 |
`needs_fixed_stride_order`.
|
| 25 |
|
|
|
|
| 174 |
- Source: https://github.com/chrisvoncsefalvay/hotdog
|
| 175 |
- Hub: https://hf.co/chrisvoncsefalvay/hotdog
|
| 176 |
|
| 177 |
+
## Author
|
| 178 |
+
|
| 179 |
+
|
| 180 |
+
I'm [Chris von Csefalvay](chrisvoncsefalvay.com), an AI researcher specialising in post-training, and the author of _[Post-Training: A Practical Guide for
|
| 181 |
+
AI Engineers and Developers](https://posttraining.guide)_ (No Starch Press, 2026). I also write [Post-Slop](https://postslop.substack.com), a periodic diatribe about AI, and what it's doing for society. You can also find me on [LinkedIn](https://linkedin.com/in/chrisvoncsefalvay) and [X](https://x.com/epichrisis).
|
| 182 |
+
## License
|
| 183 |
+
|
| 184 |
+
MIT
|