Software engineer at NVIDIA, building the GPU kernel libraries (CuTe DSL, FlashInfer, cuDNN) that power deep learning performance. Graduated from Nanjing University. Based in Shanghai. Indie game fan. Lifelong learner.
Pinned Loading
-
flashinfer-ai/flashinfer
flashinfer-ai/flashinfer PublicFlashInfer: Kernel Library for LLM Serving
-
NVIDIA/cudnn-frontend
NVIDIA/cudnn-frontend PubliccuDNN Frontend is NVIDIA's modern, open-source entry point to the cuDNN library and a growing collection of high-performance open-source kernels.
-
-
NVIDIA/tensor-ir
NVIDIA/tensor-ir PublicTensorIR is a lightweight NVIDIA-owned MLIR compiler frontend for expressing tensor computations and lowering them to NVIDIA CUDA Tile IR.
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.





