Stars
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
NVIDIA® TensorRT™ is an SDK for high-performance deep learning inference on NVIDIA GPUs. This repository contains the open source components of TensorRT.
Optimized primitives for collective multi-GPU communication
A complete AI agency at your fingertips - From frontend wizards to Reddit community ninjas, from whimsy injectors to reality checkers. Each agent is a specialized expert with personality, processes…
Achieve state of the art inference performance with modern accelerators on Kubernetes
Open-source runtime security rules engine for MCP servers and AI agents. Detects prompt injection, command injection, jailbreaks, and data exfiltration.
learn LLM inference on Apple Silicon for systems engineers: build a tiny vLLM + Qwen
Post-training with Tinker
The simplest, fastest repository for training/finetuning medium-sized GPTs.
🚀 An open-source, hands-on curriculum bridging the gap from basic RL concepts to LLM alignment, RLVR, and advanced Agentic systems.
Skills for Real Engineers. Straight from my .agents directory.
The agent that grows with you
AI agents running research on single-GPU nanochat training automatically
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
LLM Council works together to answer your hardest questions
It is said that, Ilya Sutskever gave John Carmack this reading list of ~ 30 research papers on deep learning.
Universal and Transferable Attacks on Aligned Language Models
An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.
A collection of notebooks/recipes showcasing some fun and effective ways of using Claude.
A collection of projects designed to help developers quickly get started with building deployable applications using the Claude API
A machine learning compiler for GPUs, CPUs, and ML accelerators
TensorZero is an open-source LLMOps platform that unifies an LLM gateway, observability, evaluation, optimization, and experimentation.
Agent Laboratory is an end-to-end autonomous research workflow meant to assist you as the human researcher toward implementing your research ideas
No fortress, purely open ground. OpenManus is Coming.
Ongoing research training transformer models at scale
DeepGEMM: clean and efficient BLAS kernel library on GPU
Janus-Series: Unified Multimodal Understanding and Generation Models