Stars
Automatic algorithm Iteration agent for competition about network scale load banlancing
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
ARIS ⚔️ (Auto-Research-In-Sleep) — Lightweight Markdown-only skills for autonomous ML research: cross-model review loops, idea discovery, and experiment automation. No framework, no lock-in — works…
An agentic skills framework & software development methodology that works.
🔬 Harness Vibe Research with Self-evolving AI Scientists
[ECCV 2026] Autoregressive Image Generation Needs Only a Few Lines of Cached Tokens
Must-read papers on improving efficiency for LLM serving clusters
OpenR: An Open Source Framework for Advanced Reasoning with Large Language Models
Bayesian optimisation & Reinforcement Learning library developed by Huawei Noah's Ark Lab
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…
wang90063 / RocAlphaGo
Forked from wrongu/RocAlphaGoAn independent, student-led replication of DeepMind's 2016 Nature publication, "Mastering the game of Go with deep neural networks and tree search" (Nature 529, 484-489, 28 Jan 2016), details of wh…
Asynchronous Methods for Deep Reinforcement Learning
scalable multi agents reinforcement learning
Reimplementation of DDPG(Continuous Control with Deep Reinforcement Learning) based on OpenAI Gym + Tensorflow
Using Keras and Deep Deterministic Policy Gradient to play TORCS