Highlights
- All languages
- ActionScript
- Astro
- Batchfile
- C
- C#
- C++
- CSS
- Common Lisp
- Cuda
- Dockerfile
- EJS
- Go
- HTML
- Java
- JavaScript
- Julia
- Jupyter Notebook
- Lua
- MATLAB
- MDX
- Makefile
- Markdown
- Mustache
- OCaml
- PHP
- PLpgSQL
- Perl
- PostScript
- PureScript
- Python
- QML
- R
- Ruby
- Rust
- SCSS
- Scala
- Shell
- Smarty
- Swift
- TeX
- TypeScript
- Vim Script
- Vue
Starred repositories
A Kubernetes-native cache plane for LLM inference
Skills and eval rubrics for K-12 teachers, co-developed with Learning Commons
ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…
Accurate, large-scale, and extensible simulator for LLM inference Systems
⌥ AI Coding agent for the terminal — hash-anchored edits, optimized tool harness, LSP, Python, browser, subagents, and more
Synthetic data generation, post-training, and E2B benchmark evaluation infrastructure.
Strip AI-writing tells from papers and grant proposals (NSF/NIH), while keeping scholarly voice and tying claims to evidence. A skill for Claude Code, Codex, and MorphMind.
Python tool for converting files and office documents to Markdown.
Point it at any web page and it finds the video, extracts the stream, transcodes it and casts in real time to your TV. It even burns subtitles….
Easy Data Preparation with latest LLMs-based Operators and Pipelines.
LightLLM is a Python-based LLM (Large Language Model) inference and serving framework, notable for its lightweight design, easy scalability, and high-speed performance.
Desktop and web interface for OpenCode AI agent
Principles and Methodologies for Serial Performance Optimization (OSDI' 25)
🚀 First survey bridging LLM, RL, and Agentic Eras — 100+ papers on efficient inference from static generation to systematic reasoning and action.
[Up-to-date] Large Language Model Agent: A Survey on Methodology, Applications and Challenges
将博导十年科研经验炼化为可直接调用的 AI 技能。从 Idea 构思到论文投稿,你的 AI 科研副导师。
Fork() for AI agent microVMs. Spawn 100 children in ~100ms from a warm parent; BRANCH a live VM in ~150ms. KVM-isolated, snapshot CoW.
AI 基础知识 - GPU 架构、CUDA 编程、大模型基础及AI Agent 相关知识。
A workload for deploying LLM inference services on Kubernetes
Open Model Engine (OME) — Kubernetes operator for LLM serving, GPU scheduling, and model lifecycle management. Works with SGLang, vLLM, TensorRT-LLM, and Triton
📚 《Deep Agents 实战》—— LangChain 官方大使出品,基于 LangChain / LangGraph 生态,从零构建生产级 AI Agent 的完整指南
一个开源的GPU服务器管理平台;可以实时查看模型训练状态、GPU资源占用、模型训练日志、IP访问记录等
Summary of some awesome work for optimizing LLM inference
让 AI 可以驱动 Abaqus建模。Connect Claude, Cursor, and other MCP clients directly to your active Abaqus/CAE session.