Master student at SJTU
Research focus: AI Infra full-stack — GPU architecture, inference, operator development
-
Shanghai Jiao Tong University
- 800 Dongchuan RD. Minhang District, Shanghai, China
Highlights
- Pro
Pinned Loading
-
vllm-project/vllm-omni
vllm-project/vllm-omni PublicA framework for efficient model inference with omni-modality models
-
-
SGEMM_CUDA
SGEMM_CUDA PublicForked from siboehm/SGEMM_CUDA
Fast CUDA matrix multiplication from scratch
Cuda
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.