Skip to content
View haitwang-cloud's full-sized avatar
🎯
Focusing
🎯
Focusing

Organizations

@SAP

Block or report haitwang-cloud

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
haitwang-cloud/README.md

I'm passionate about AI Infrastructure and Open Source, especially the journey from GPU → LLM Inference → Tokens.

I contribute to projects across the AI infrastructure stack, including vLLM, SGLang, llm-d, KServe, HAMi, and Dynamo, with a focus on Kubernetes, GPU infrastructure, inference runtimes, scheduling, routing, and LLM serving.

I also share what I learn through my 📖 blog. Always happy to connect, collaborate, and learn from the community! 🚀

Currently learning Spanish on Duolingo. ¡Hola! 🇪🇸

Pinned Loading

  1. Project-HAMi/HAMi Project-HAMi/HAMi Public

    Heterogeneous GPU Sharing on Kubernetes

    Go 4.5k 800

  2. k8s-dra-driver k8s-dra-driver Public

    Forked from kubernetes-sigs/dra-driver-nvidia-gpu

    Dynamic Resource Allocation (DRA) for NVIDIA GPUs in Kubernetes

    Go

  3. sgl-project/sglang sgl-project/sglang Public

    SGLang is a high-performance serving framework for large language models and multimodal models.

    Python 33k 8.4k

  4. kserve/kserve kserve/kserve Public

    Standardized Distributed Generative and Predictive AI Inference Platform for Scalable, Multi-Framework Deployment on Kubernetes

    Go 5.8k 1.6k