Senior AI Infrastructure Engineer | GPU Optimization & Distributed Inference Platforms | vLLM, Triton, TensorRT-LLM, NCCL, CUDA C++ | High-Scale Kubernetes
- Santa Clara, CA
- https://www.linkedin.com/in/dundysm
Pinned Loading
-
NVIDIA/gpu-operator
NVIDIA/gpu-operator PublicNVIDIA GPU Operator creates, configures, and manages GPUs in Kubernetes
-
vllm-project/production-stack
vllm-project/production-stack PublicvLLM’s reference system for K8S-native cluster-wide deployment with community-driven performance optimization
-
-
production-stack
production-stack PublicForked from vllm-project/production-stack
vLLM’s reference system for K8S-native cluster-wide deployment with community-driven performance optimization
Python
-
-
google-deepmind/mujoco_warp
google-deepmind/mujoco_warp PublicGPU-optimized version of the MuJoCo physics simulator, designed for NVIDIA hardware.
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.


