-
Carnegie Mellon University
- Pittsburgh, PA
Highlights
- Pro
Stars
Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, intelligent classification, and agent-based multilingual summar…
[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments
The Cradle framework is a first attempt at General Computer Control (GCC). Cradle supports agents to ace any computer task by enabling strong reasoning abilities, self-improvment, and skill curatio…
Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents
🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN
Scaling Deep Research via Reinforcement Learning in Real-world Environments.
[NeurIPS 2025] 🌐 WebThinker: Empowering Large Reasoning Models with Deep Research Capability
Multi-turn RL framework for aligning models to be tutors instead of answerers. EMNLP 2025 Oral
RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.
This repository hosts the paper “LLM Based Math Tutoring: Challenges and Dataset”, along with the accompanying dataset. It explores the performance and challenges of Large Language Models (LLMs) in…
Code for the paper "Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues", published at AIED 2025.
Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
Eko (Eko Keeps Operating) - Build Production-ready Agentic Workflow with Natural Language - eko.fellou.ai
An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. The goal of this repo is to provide the si…
Convert any URL to an LLM-friendly input with a simple prefix https://r.jina.ai/
🤗 smolagents: a barebones library for agents that think in code.
An open source deep research clone. AI Agent that reasons large amounts of web data extracted with Firecrawl