-
Peking University
- Beijing
-
23:57
(UTC +08:00) - https://al-377.github.io/
Lists (32)
Sort Name ascending (A-Z)
Agent
multi-agentAgent-tuning
App
Fantastic appart
some beautiful thingsBooks
some booksCS
list of the basement of CSCV
computer visionDataset
Exp-Algorithm
algorithm exInterview
KG
list of knowledege graph and NLPLLM
large large modelLLM-Accelerate
each type of efficient op in LLMLLM-alignment
The methods for alignment data generationLLM-Customized
characterLLM-Eval
LLM-Inference
forwardLLM-infra
The LLm imfrastructureLLM-Integration
LLM-Knowledge
LLM-prompt
The prompt engineering of LLMLLM-reasoning
The reasoning ability of LLMLLM-Recommendation
LLM-Tool
tool learningLLM-tuning
The tools for LLM tuningLLM4AD
MAR
multi agent rlMLC
machine learning compilationPOOL
todoRL
reinforce learningSKILLS
skills mapsWorld Model
Stars
The roadmap of long-horizon agents
Framework for evaluating and improving agents
EdgeBench: Unveiling scaling laws of learning from real-world environments
A benchmark for evaluating AI agents on realistic business workflows
Harness for running and evaluating AI agents against RL environments
Benchmark self-evolving Agent upon realistic large-scale file workspaces
Implementation for: Towards On-Policy Data Evolution for Visual-Native Multimodal Deep Search Agents
Context engineering for AI Agents. Manage your LLM's context window with caching, compaction, summarization, and graceful degradation.
Trae Agent is an LLM-based agent for general purpose software engineering tasks.
🔥 A collection of the Claude Code open source
The agent benchmark that scores the full stack — harness, config, and model — not just the LLM. Trace-based scoring, reliability metrics, configuration diagnostics.
The agent that grows with you
Official PyTorch implementation of "Visually-grounded Humanoid Agents"
毛选.skill — 让毛泽东的思维框架帮你分析问题、制定策略、看透本质。7个核心心智模型 · 10条决策启发式 · 完整表达DNA。不是复读语录,是用他的认知框架帮你看问题。
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
feishu-cli 是一个功能完整的飞书开放平台命令行工具。它将飞书文档、知识库、电子表格、消息、日历、任务等操作封装为简洁的命令行接口,核心能力是 Markdown ↔ 飞书文档双向无损转换。
Create beautiful slides on the web using a coding agent's frontend skills
Toolathlon-Gym for testing AI agents real-world tool-use capabilities across diverse MCP servers.
Claw-Eval is an evaluation harness for evaluating LLM as agents. All tasks verified by humans.
Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1
Agent World Model: Infinity Synthetic Environments for Agentic Reinforcement Learning
Yunjue Agent: A Fully Reproducible, Zero-Start In-Situ Self-Evolving Agent System for Open-Ended Tasks