Skip to content
View 1KE-JI's full-sized avatar

Block or report 1KE-JI

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

MLEvolve is an open-source autonomous system for end-to-end machine learning algorithm design and optimization powered by progressive search and experience-driven memory.

Python 412 56 Updated Jul 14, 2026

Agentifying Patient Dynamics within LLMs through Interacting with Clinical World Model

Python 31 Updated May 15, 2026

[ICML'26] Scaling Long-Horizon LLM Agent via Context-Folding

Python 179 12 Updated May 18, 2026

MyPhoneBench: Do Phone-Use Agents Respect Your Privacy?

Python 24 Updated Apr 3, 2026

[NeurIPS 2024] OSWorld: Benchmarking Multimodal Agents for Open-Ended Tasks in Real Computer Environments

Python 3,070 512 Updated Jul 28, 2026

The lightweight framework for building agents

Python 535 60 Updated Aug 5, 2026

The largest open-source medical AI skills library for OpenClaw🦞.

Python 2,921 410 Updated Jul 21, 2026

Scalable toolkit for efficient model reinforcement

Python 1,890 501 Updated Aug 8, 2026

Code for the paper: Modular Retrieval for Generalization and Interpretation.

Python 14 Updated Feb 11, 2026
Python 36 2 Updated Oct 31, 2025

slime is an LLM post-training framework for RL Scaling.

Python 7,813 1,127 Updated Aug 7, 2026

Memory Agent monorepo

Python 89 11 Updated Oct 9, 2025

An early research stage expert-parallel load balancer for MoE models based on linear programming.

Python 523 39 Updated Nov 19, 2025

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,192 208 Updated Jun 9, 2026
Python 3 1 Updated Oct 28, 2025

MiniMax-M2, a model built for Max coding & agentic workflows.

2,601 216 Updated Nov 13, 2025

Async pipelined version of Verl

Python 124 12 Updated Apr 8, 2025

The official repo of "WebExplorer: Explore and Evolve for Training Long-Horizon Web Agents"

Python 120 4 Updated Sep 29, 2025

A simple yet powerful agent framework that delivers with open-source models

Python 4,594 473 Updated Mar 21, 2026

🏆 Top-1 on 5+ benchmarks | Web UI | Supports MiroThinker, Claude, Kimi, OpenAI

Python 3,090 322 Updated Jul 6, 2026

Writing AI Conference Papers: A Handbook for Beginners

3,955 144 Updated Jul 16, 2025

A final sanity checklist to help your CS paper get accepted, not desk rejected.

1,608 145 Updated May 25, 2026

🪐 🔧 Model Context Protocol (MCP) Server for Jupyter.

Python 1,239 186 Updated Aug 8, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,088 1,581 Updated Aug 8, 2026

Kimi K2 is the large language model series developed by Moonshot AI team

11,099 903 Updated Jan 21, 2026

A curated list of cutting-edge research papers and resources on Long Chain-of-Thought (CoT) Reasoning with Tools.

46 4 Updated Dec 17, 2025

Reinforcing General Reasoning without Verifiers

Python 102 5 Updated Jun 24, 2025

Extrapolating RLVR to General Domains without Verifiers

Python 205 13 Updated Aug 12, 2025
Next