Skip to content
View zxyscz's full-sized avatar

Block or report zxyscz

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

command line management for Google Workspace

Python 4,247 532 Updated Jul 25, 2026
Python 84 7 Updated Mar 11, 2025

Democratizing Reinforcement Learning for LLMs

Python 5,737 596 Updated Jul 27, 2026
Python 160 37 Updated May 13, 2026

τ-Bench: A Benchmark for Tool-Agent-User Interaction in Real-World Domains

Python 1,676 425 Updated Jul 24, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 87,322 19,917 Updated Jul 27, 2026

Get JSON values quickly - JSON parser for Go

Go 15,545 906 Updated May 14, 2026

OpenClaw-RL: Train any agent simply by talking

Python 5,609 608 Updated May 23, 2026

A browser-based desktop where AI Agent operates every app through natural language.

TypeScript 1,237 162 Updated Jun 3, 2026
Python 96 8 Updated Dec 23, 2025

RL research on Android devices.

Python 1,234 118 Updated Jul 27, 2026

dLLM: Simple Diffusion Language Modeling

Python 2,656 280 Updated Jul 17, 2026

SkillsBench evaluates how well skills work and how effective agents are at using them.

PDDL 1,584 346 Updated Jul 23, 2026

The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.

Python 5,609 567 Updated Jul 24, 2026

Agent Skills to help developers using AI agents with Supabase

TypeScript 2,429 176 Updated Jul 21, 2026

A LLM-based Agent that predict its tasks proactively.

Python 638 64 Updated May 12, 2026

The agent-native LLM router for autonomous agents. 55+ models (8 free), <1ms local routing, USDC payments on Base & Solana via x402.

TypeScript 6,680 631 Updated Jul 27, 2026

Research of DeepSeek Engram Architecture based on Qwen-3 and Stable Diffusion series.

Python 735 51 Updated May 21, 2026

Persistent Context Across Sessions for Every Agent – Captures everything your agent does during sessions, compresses it with AI, and injects relevant context back into future sessions. Works with C…

JavaScript 88,722 7,710 Updated Jul 23, 2026
Python 13 Updated Apr 20, 2026
Python 303 24 Updated Jun 30, 2026

Open Visual Agentic Intelligence

2,273 316 Updated Jan 31, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 384,318 80,741 Updated Jul 27, 2026

Kimi K2 is the large language model series developed by Moonshot AI team

11,043 889 Updated Jan 21, 2026

Dr. Zero Self-Evolving Search Agents without Training Data

Python 525 65 Updated Mar 23, 2026

SimpleMem: Efficient Lifelong Memory for LLM Agents — Text & Multimodal

Python 3,659 384 Updated Jul 24, 2026

PhD Thesis work -- computational model of learning and memory in decision making in reinforcement learning tasks

Jupyter Notebook 13 6 Updated Oct 15, 2021

The code for NeurIPS 2025 paper "A-Mem: Agentic Memory for LLM Agents"

Python 929 102 Updated Mar 5, 2026

Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models

Python 4,568 350 Updated Jan 14, 2026

slime is an LLM post-training framework for RL Scaling.

Python 7,665 1,100 Updated Jul 24, 2026
Next