Starred repositories
AgentENV (AENV) is a distributed platform for running agent environments at scale.
AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
Spec-driven development (SDD) for AI coding assistants.
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
Build high-quality LLM apps - from prototyping, testing to production deployment and monitoring.
The absolute trainer to light up AI agents.
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
🤗 smolagents: a barebones library for agents that think in code.
Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.
Agent2Agent (A2A) is an open protocol enabling communication and interoperability between opaque agentic applications.
Build Agentic workflows, RAG pipelines, with rich AI model and tool support on one collaborative workspace. Deploy on cloud, VPC, or self-hosted, so teams move from prototype to production without …
🐫 CAMEL: The first and the best multi-agent framework. Finding the Scaling Law of Agents. https://www.camel-ai.org
🦉 OWL: Optimized Workforce Learning for General Multi-Agent Assistance in Real-World Task Automation
Qwen2.5-Omni is an end-to-end multimodal model by Qwen team at Alibaba Cloud, capable of understanding text, audio, vision, video, and performing real-time speech generation.
Train your AI self, amplify you, bridge the world
A course on aligning smol models.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
This repository has code for fine-tuning LLMs with GRPO specifically for Rust Programming using cargo as feedback
RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
A live stream development of RL tunning for LLM agents
A lightweight, powerful framework for multi-agent workflows
FlashMLA: Efficient Multi-head Latent Attention Kernels
DeepEP: an efficient expert-parallel communication library