-
Second Nature Computing
- San Francisco, California
- https://linkedin.com/blakeledden
Pinned Loading
-
sm121-kernels
sm121-kernels PublicHand-written PTX kernels for the NVIDIA DGX Spark (GB10 / SM121): flash attention, GEMM, GDN linear attention, MoE, and quantization; no runtime CUDA toolkit, just the driver. By Blake Ledden on be…
Rust 2
-
flash-attention-dao
flash-attention-dao PublicForked from Dao-AILab/flash-attention
Fast and memory-efficient exact attention
Python 2
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.