🤒
Out sick
ISTJ👋. Currently focused on exploring long-context prefilling and decoding techniques for large language models (LLMs).
-
The University of Hong Kong (HKU)
- Hong Kong
- https://jianqiaolu.github.io/
Pinned Loading
-
XunhaoLai/native-sparse-attention-triton
XunhaoLai/native-sparse-attention-triton PublicEfficient triton implementation of Native Sparse Attention.
-
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.