-
xAI | BUAA
- Sunnyvale, CA
-
16:16
(UTC -07:00) - hebiao064.github.io
- in/biao-he
- @hebiao064
Lists (1)
Sort Name ascending (A-Z)
Starred repositories
Lightweight coding agent that runs in your terminal
Optimized primitives for collective multi-GPU communication
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
🧑🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including transformers (original, xl, switch, feedback, vit, ...), optimizers (adam, adabelief, sophia, ...), ga…
🚀 Pytorch Distributed native training library for LLMs/VLMs with OOTB Hugging Face support
A Python DSL to write Nvidia PTX for Hopper and Blackwell in JAX and PyTorch
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
Engine-agnostic LLM gateway in Rust. Full OpenAI & Anthropic API compatibility across vLLM, TRT-LLM, TokenSpeed, SGLang, OpenAI, Gemini & more. Industry-first gRPC pipeline, KV cache-aware routing,…
A terminal workspace with batteries included
分享AI Infra知识&代码练习:PyTorch、vLLM/SGLang、slime/vime框架入门⚡️、性能加速🚀、大模型基础🧠、AI软硬件🔧等
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
A minimal, secure Python interpreter written in Rust for use by AI
The repository has collected a batch of noteworthy MLSys bloggers (Algorithms/Systems)
Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
[ICLR'25] BigCodeBench: Benchmarking Code Generation Towards AGI
A simple, performant, and scalable Jax LLM!
Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …
A next.js web application that integrates AI capabilities with draw.io diagrams. This app allows you to create, modify, and enhance diagrams through natural language commands and AI-assisted visual…
Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.
Utility scripts for PyTorch (e.g. Make Perfetto show some disappearing kernels, Memory profiler that understands more low-level allocations such as NCCL, ...)