Software engineer at NVIDIA, building the GPU kernel libraries (CuTe DSL, FlashInfer, cuDNN) that power deep learning performance. Graduated from Nanjing University. Based in Shanghai. Indie game fans. Lifelong learner.
🐱
Kernels, LLM, and Cat
Highlights
Pinned Loading
-
-
flashinfer-ai/flashinfer
flashinfer-ai/flashinfer PublicFlashInfer: Kernel Library for LLM Serving
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.





