-
AMD
- San Jose, California
-
05:32
(UTC -07:00)
-
-
hipflex Public
Flexible GPU fractionalization — run more workloads per GPU.
Rust Apache License 2.0 UpdatedJul 16, 2026 -
-
-
pytorch Public
Forked from pytorch/pytorchTensors and Dynamic neural networks in Python with strong GPU acceleration
Python Other UpdatedApr 11, 2026 -
vgpu.rs Public
Forked from NexusGPU/vgpu.rsvgpu.rs is the fractional GPU & vgpu-hypervisor implementation written in Rust
Rust Apache License 2.0 UpdatedApr 2, 2026 -
tensor-fusion Public
Forked from NexusGPU/tensor-fusionTensor Fusion is a state-of-the-art GPU virtualization and pooling solution designed to optimize GPU cluster utilization to its fullest potential.
Go Apache License 2.0 UpdatedMar 9, 2026 -
-
device-metrics-exporter Public
Forked from ROCm/device-metrics-exporterDevice Metrics Exporter exports metrics from AMD devices (GPUs) to collectors like Prometheus.
C++ Apache License 2.0 UpdatedJan 21, 2026 -
-
-
hai-sglang Public
Forked from HaiShaw/sglangSGLang is a fast serving framework for large language models and vision language models.
Python Apache License 2.0 UpdatedMay 20, 2025 -
-
iree Public
Forked from iree-org/ireeA retargetable MLIR-based machine learning compiler and runtime toolkit.
C++ Apache License 2.0 UpdatedMar 7, 2025 -
TheRock Public
Forked from ROCm/TheRockThe HIP Environment and ROCm Kit - A lightweight open source build system for HIP and ROCm
CMake Apache License 2.0 UpdatedFeb 7, 2025 -
-
triton Public
Forked from triton-lang/tritonDevelopment repository for the Triton language and compiler
C++ MIT License UpdatedJan 30, 2025 -
AKS-GitHubARC-Setup Public
Documentation for bringing up an Azure Kubernetes cluster integrated with GitHub Actions Runner Controller for IREE Project
-
ossci-fleet Public
Forked from nod-ai/ossci-fleetThe goal of the OSSCI Fleet is to provide a central mechanism to enable test automation, batch job scheduling, and developer access to a federated set of limited GPU resources
Shell Apache License 2.0 UpdatedJan 24, 2025 -
discord-cluster-manager Public
Forked from gpu-mode/kernelbotThe "leetcode"-like platform for GPU hackers
Python UpdatedJan 22, 2025 -
modded-nanogpt Public
Forked from KellerJordan/modded-nanogptNanoGPT (124M) in 3.4 minutes
Python MIT License UpdatedJan 14, 2025 -
iree-turbine Public
Forked from iree-org/iree-turbineIREE's PyTorch Frontend, based on Torch Dynamo.
Python Apache License 2.0 UpdatedDec 16, 2024 -
Megatron-LM Public
Forked from ROCm/Megatron-LMOngoing research training transformer models at scale
Python Other UpdatedDec 10, 2024 -
A high-throughput and memory-efficient inference and serving engine for LLMs
Python Apache License 2.0 UpdatedDec 5, 2024 -
Liger-Kernel Public
Forked from linkedin/Liger-KernelEfficient Triton Kernels for LLM Training
Python BSD 2-Clause "Simplified" License UpdatedDec 5, 2024 -
vllm-docs Public
Forked from powderluv/vllm-docsDocumentation for vLLM Dev Channel releases
Apache License 2.0 UpdatedNov 27, 2024 -
-
mlperf-inference Public
Forked from mlcommons/inferenceReference implementations of MLPerf™ inference benchmarks
Python Apache License 2.0 UpdatedNov 5, 2024 -
iree-kernel-benchmark Public
Becnhmark kernels performance through IREE
-