Skip to content
View austinmw's full-sized avatar

Block or report austinmw

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Lab Cookbook

Python 42 2 Updated Aug 5, 2026

Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台…

Python 1,754 305 Updated Sep 23, 2026

Home for "How To Scale Your Model", a short blog-style textbook about scaling LLMs on TPUs

SCSS 1,447 207 Updated Sep 22, 2026

NanoGPT (124M) in 90 seconds

Python 5,819 892 Updated Sep 18, 2026

Open-source framework for the research and development of foundation models.

Python 3,798 303 Updated Sep 23, 2026

Puffing up reinforcement learning

C 6,470 577 Updated Sep 13, 2026

Our library for RL environments + evals

Python 4,649 681 Updated Sep 23, 2026

The open source coding agent.

TypeScript 209,655 27,663 Updated Sep 23, 2026

Warp is an agentic development environment, born out of the terminal.

Rust 65,143 5,582 Updated Sep 23, 2026

Benchmark LLMs by fighting in Street Fighter 3! The new way to evaluate the quality of an LLM

Jupyter Notebook 1,482 182 Updated Mar 21, 2025

Qwen3.6-35B-A3B-heretic NVFP4 + DFlash speculative decoding on DGX Spark (GB10/sm_121a). Source-built vLLM image + 7 patches + comprehensive deployment guide.

Python 144 15 Updated Jun 28, 2026

🪨 why use many token when few token do trick. Viral skill + proxy for coding agents that cuts 65% of tokens by talking like a caveman.

Go 107,548 6,232 Updated Sep 23, 2026

🦀🌡️ Real-time system monitor for Apple Silicon Macs (M1–M5). No sudo. TUI, JSON/Prometheus metrics server, and Rust library.

Rust 1,894 81 Updated Aug 4, 2026

Symphony turns project work into isolated, autonomous implementation runs, allowing teams to manage work instead of supervising coding agents.

Elixir 27,377 2,836 Updated Sep 15, 2026

A framework for teaching AI to write like you. Not like a better version of you. Like you.

317 46 Updated Apr 13, 2026
Python 41 7 Updated Nov 11, 2025

Atropos is a Language Model Reinforcement Learning Environments framework for collecting and evaluating LLM trajectories through diverse environments

Python 1,351 401 Updated Jul 4, 2026

🎨 NeMo Data Designer: Generate high-quality synthetic data from scratch or from seed data.

Python 2,274 211 Updated Sep 23, 2026

User-friendly AI Interface (Supports Ollama, OpenAI API, ...)

Python 152,933 22,373 Updated Sep 23, 2026

Run Coding Agents in Sandboxes. Control Them Over HTTP. Supports Claude Code, Codex, OpenCode, and Amp.

TypeScript 1,572 127 Updated Jun 19, 2026

World model reinforcement learning for multi-turn VLM agents. RL for vision framework (NeurIPS 2025).

Python 505 59 Updated Sep 5, 2026

Official repository of the 2025 paper, LLM Economist: Large Population Models and Mechanism Design in Multi-Agent Generative Simulacra.

Python 125 14 Updated Mar 11, 2026

JupyMD: Use Obsidian as a Jupyter notebook IDE

TypeScript 293 15 Updated Sep 13, 2026

Every Eval Ever is a shared schema and crowdsourced eval database. It defines a standardized metadata format for storing AI evaluation results — from leaderboard scrapes and research papers to loca…

Python 126 50 Updated Sep 20, 2026

bf16 LoRA fine-tuning of [Qwen3.5-35B-A3B](https://huggingface.co/unsloth/Qwen3.5-35B-A3B) (a 35B-total / 3B-active Mixture-of-Experts vision-language model) on a single NVIDIA DGX Spark — without …

Python 18 1 Updated Mar 12, 2026

Trio – a friendly Python library for async concurrency and I/O

Python 7,336 434 Updated Sep 21, 2026

AI agents running research on single-GPU nanochat training automatically

Python 96,668 13,522 Updated Mar 26, 2026

Framework for evaluating and improving agents

Python 5,544 1,850 Updated Sep 23, 2026

Scripts for agents, shared between my repositories.

Shell 6,648 547 Updated Sep 22, 2026

The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

JavaScript 266,142 39,775 Updated Sep 22, 2026
Next