-
Huawei
- Singapore
-
11:31
(UTC +08:00) - in/jian-zhang-a699ab358
- @zhangj1an
Highlights
- Pro
Pinned Loading
-
vllm
vllm PublicForked from vllm-project/vllm
A high-throughput and memory-efficient inference and serving engine for LLMs
Python
-
vllm-omni
vllm-omni PublicForked from vllm-project/vllm-omni
A framework for efficient model inference with omni-modality models
Python
-
pytorch
pytorch PublicForked from pytorch/pytorch
Tensors and Dynamic neural networks in Python with strong GPU acceleration
Python
-
rl-kernel
rl-kernel PublicForked from RL-Align/RL-Kernel
Modern RL Post-training Infrastructure: Optimized for NVIDIA/AMD GPUs with a focus on vLLM and DeepSpeed integration, CUDA/ROCm/Triton kernels, and transparent hardware-aware scaling.
Python
If the problem persists, check the GitHub status page or contact support.



