Skip to content
View OCWC22's full-sized avatar

Highlights

  • Pro

Block or report OCWC22

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Mixture-of-experts (MoE) training megakernel for NVL72s

Python 376 35 Updated Aug 5, 2026

[ICLR'26] Stronger-MAS: A RL Framework for multi LLM agent system; [arxiv] MetaAgent-X: End-to-End Reinforcement Learning Automatic Multi-Agent Systems

Python 208 28 Updated May 15, 2026

PRO-V: An Efficient Program Generation Multi-Agent System for Automatic RTL Verification

Python 36 4 Updated Dec 11, 2025

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

Python 1,031 111 Updated Aug 4, 2026

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 2,903 231 Updated Aug 5, 2026

An agentic system that auto-optimizes LLM workloads on AMD GPUs.

Python 126 17 Updated Aug 5, 2026

A Call of Duty-quality FPS in Three.js, built from a single prompt.

JavaScript 2,850 420 Updated Jul 25, 2026

A community-maintained Python framework for creating mathematical animations.

Python 39,887 3,002 Updated Aug 5, 2026

Tenstorrent's MLIR Based Compiler. We aim to enable developers to run AI on all configurations of Tenstorrent hardware, through an open-source, general, and performant compiler.

JavaScript 339 50 Updated Aug 5, 2026

🤘 TT-NN operator library, and TT-Metalium low level kernel programming model.

C++ 1,616 563 Updated Aug 5, 2026
Cuda 44 1 Updated Aug 3, 2026

Bonsai Demo

Shell 2,163 220 Updated Aug 5, 2026

A high-performance inference engine for AI models

Rust 1,666 68 Updated Aug 5, 2026

Build a compiler to solve Anthropic's interview challenge.

Python 134 16 Updated Jul 12, 2026

Run GPT-5.6 Sol and Codex models inside Claude Code — private, localhost-only, one-command setup for macOS and Linux.

Shell 16 Updated Aug 4, 2026

The fastest Mamba2 SSD kernel for TPU written in Pallas

HTML 3 Updated Jul 31, 2026

My reasearch of losslessly compressing LLM weights.

HTML 62 4 Updated Jul 23, 2026

Alex — a local LLM proxy for all your token providers, APIs and harnesses. Route Claude, ChatGPT/Codex, Gemini, Grok, Kimi & OpenRouter subscriptions into any coding tool. Your tokens, your traces,…

Rust 77 8 Updated Jul 28, 2026

Use Codex from Claude Code to review code or delegate tasks.

JavaScript 31,385 2,124 Updated Jul 8, 2026

Hyper-fast local-first LLM router: point it at your local server, burst overflow to TrustedRouter.

Go 6 2 Updated Jul 9, 2026

Anthropic's original performance take-home, now open for you to try!

Python 4,086 931 Updated Jan 22, 2026

🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman

JavaScript 96,126 5,521 Updated Aug 4, 2026

Nsight Python is a Python kernel profiling interface based on NVIDIA Nsight Tools

Python 285 21 Updated Aug 3, 2026

Machine Learning Engineering Open Book

Python 18,522 1,185 Updated Aug 5, 2026

Any model. Any hardware. Zero compromise. Built with @ziglang / @openxla / MLIR / @bazelbuild

Zig 3,962 172 Updated Aug 3, 2026

Design visual, review-gated agent loops for Claude Code before you run them.

Python 697 64 Updated Jul 7, 2026

Skill + MCP server to turn your agent into an RLM. Load context, iterate with search/code/think tools, converge on answers.

Python 210 24 Updated Apr 11, 2026
Next