Skip to content
View tsljgj's full-sized avatar
  • Carnegie Mellon University
  • Pittsburgh, PA

Highlights

  • Pro

Block or report tsljgj

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Paper Claw sends personalized daily research digests from arXiv and beyond straight to your inbox, featuring customizable categories, intelligent classification, and agent-based multilingual summar…

Python 34 3 Updated Aug 16, 2026

[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

Python 3,082 517 Updated Aug 12, 2026
Python 156 18 Updated Jun 17, 2025

The Cradle framework is a first attempt at General Computer Control (GCC). Cradle supports agents to ace any computer task by enabling strong reasoning abilities, self-improvment, and skill curatio…

Python 2,566 265 Updated Nov 7, 2024

An MCP-based chatbot | 一个基于MCP的聊天机器人

C++ 28,930 6,647 Updated Aug 15, 2026
Python 1 Updated Jan 15, 2026
Python 257 34 Updated Apr 7, 2026

AWM: Agent Workflow Memory

Python 457 51 Updated Dec 22, 2025
Python 188 17 Updated Aug 15, 2026

Darwin Gödel Machine: Open-Ended Evolution of Self-Improving Agents

Python 2,227 439 Updated Aug 13, 2025

🚀🤖 Crawl4AI: Open-source LLM Friendly Web Crawler & Scraper. Don't be shy, join here: https://discord.gg/jP8KfhDhyN

Python 78,270 8,102 Updated Aug 15, 2026

Scaling Deep Research via Reinforcement Learning in Real-world Environments.

Python 794 54 Updated May 10, 2026

[NeurIPS 2025] 🌐 WebThinker: Empowering Large Reasoning Models with Deep Research Capability

Python 1,466 142 Updated Dec 8, 2025
Python 9 Updated Feb 26, 2025
Python 1 Updated Jun 11, 2025

Multi-turn RL framework for aligning models to be tutors instead of answerers. EMNLP 2025 Oral

Python 42 12 Updated Dec 11, 2025

RAGEN leverages reinforcement learning to train LLM reasoning agents in interactive, stochastic environments.

Python 2,770 229 Updated Jul 24, 2026

This repository hosts the paper “LLM Based Math Tutoring: Challenges and Dataset”, along with the accompanying dataset. It explores the performance and challenges of Large Language Models (LLMs) in…

58 7 Updated Aug 29, 2024

Code for the paper "Training LLM-based Tutors to Improve Student Learning Outcomes in Dialogues", published at AIED 2025.

Python 14 3 Updated Aug 22, 2025

Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL

Python 5,297 479 Updated Nov 13, 2025

Eko (Eko Keeps Operating) - Build Production-ready Agentic Workflow with Natural Language - eko.fellou.ai

TypeScript 4,949 441 Updated Mar 3, 2026

An Easy-to-use, Scalable and High-performance Agentic RL Framework based on Ray (PPO & DAPO & REINFORCE++ & VLM & TIS & vLLM & Ray & Async RL)

Python 9,917 1,002 Updated Aug 13, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,974 4,410 Updated Aug 15, 2026
Jupyter Notebook 2,785 359 Updated May 2, 2025

An AI-powered research assistant that performs iterative, deep research on any topic by combining search engines, web scraping, and large language models. The goal of this repo is to provide the si…

TypeScript 19,557 1,991 Updated Apr 11, 2026

Convert any URL to an LLM-friendly input with a simple prefix https://r.jina.ai/

TypeScript 11,872 871 Updated May 22, 2026

🤗 smolagents: a barebones library for agents that think in code.

Python 28,816 2,863 Updated Jul 21, 2026

An open source deep research clone. AI Agent that reasons large amounts of web data extracted with Firecrawl

TypeScript 6,276 745 Updated May 7, 2025

Democratizing Reinforcement Learning for LLMs

Python 5,784 606 Updated Aug 16, 2026
Next