Skip to content
View yeandy's full-sized avatar

Block or report yeandy

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

An agentic system that auto-optimizes LLM workloads on AMD GPUs.

Python 142 25 Updated Aug 14, 2026

Toolkit for launching and observing MaxText training on Slurm-managed GPU clusters

Shell 29 3 Updated Jul 19, 2026

Primus-SaFE(Stability and Fault Endurance)

Go 58 4 Updated Aug 14, 2026

A high-performance acceleration library dedicated to large-scale model training on AMD GPUs

Python 68 25 Updated Aug 14, 2026

A high-performance distributed file system designed to address the challenges of AI training and inference workloads.

C++ 10,116 1,077 Updated May 7, 2026

UCCL is an efficient communication library for GPUs, covering collectives, P2P (e.g., KV cache transfer, RL weight transfer), and EP (e.g., GPU-driven)

C++ 1,490 169 Updated Aug 14, 2026

Scale-out system monitoring

Python 26 6 Updated Aug 13, 2026

A flexible and high-performance training framework designed for large-scale foundation model training on AMD GPUs

Python 119 46 Updated Aug 14, 2026

Automated KRAI X workflows for Google Cloud Platform

Python 4 Updated Apr 2, 2025
Python 591 61 Updated Jul 11, 2024

Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more

Python 36,157 3,736 Updated Aug 14, 2026

Master the command line, in one page

162,104 14,836 Updated Jun 25, 2024

Apache Beam is a unified programming model for Batch and Streaming data processing.

Java 8,639 4,623 Updated Aug 14, 2026