Skip to content
View JobQiu's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report JobQiu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Lightweight Image Video Action Generation Inference Framework

Python 2,669 254 Updated Aug 14, 2026

Scalable toolkit for efficient model reinforcement

Python 1,903 515 Updated Aug 14, 2026

MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.

Python 8,371 642 Updated Jul 6, 2026

The most powerful and modular diffusion model GUI, api and backend with a graph/nodes interface.

Python 127,594 15,017 Updated Aug 14, 2026

Paper list of multi-agent reinforcement learning (MARL)

4,870 767 Updated Feb 11, 2026

OpenSpiel is a collection of environments and algorithms for research in general reinforcement learning and search/planning in games.

C++ 5,411 1,166 Updated Aug 12, 2026

A standard API for multi-agent reinforcement learning environments, with popular reference environments and related utilities

Python 3,489 514 Updated Aug 13, 2026

BenchMARL is a library for benchmarking Multi-Agent Reinforcement Learning (MARL). BenchMARL allows to quickly compare different MARL algorithms, tasks, and models while being systematically ground…

Python 653 133 Updated Feb 7, 2026

One repository is all that is necessary for Multi-agent Reinforcement Learning (MARL)

Python 1,343 198 Updated Nov 28, 2024

verl-agent is an extension of veRL, designed for training LLM/VLM agents via RL. verl-agent is also the official code for paper "Group-in-Group Policy Optimization for LLM Agent Training"

Python 2,218 214 Updated Jun 9, 2026

Checkpoint-engine is a simple middleware to update model weights in LLM inference engines

Python 996 102 Updated Aug 12, 2026

Cognee is the open-source AI memory platform for agents. Give your AI agents persistent long-term memory across sessions with a self-hosted knowledge graph engine.

Python 30,024 2,919 Updated Aug 14, 2026

🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!

Python 54,700 7,144 Updated Aug 6, 2026

A free, open source, and extensible speech-to-text application that works completely offline.

Rust 29,511 2,614 Updated Aug 14, 2026

The absolute trainer to light up AI agents.

Python 17,481 1,535 Updated Aug 14, 2026

A collection of notebooks/recipes showcasing some fun and effective ways of using Claude.

Jupyter Notebook 51,516 6,115 Updated Aug 14, 2026

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,662 11,176 Updated Jul 22, 2026

Automate browser based workflows with AI

Python 22,754 2,140 Updated Aug 14, 2026

MineContext is your proactive context-aware AI partner(Context-Engineering+ChatGPT Pulse)

Python 5,457 407 Updated May 7, 2026

⭐AI-driven public opinion & trend monitor with multi-platform aggregation, RSS, and smart alerts.🎯 告别信息过载,你的 AI 舆情监控助手与热点筛选工具!聚合多平台热点 + RSS 订阅,支持关键词精准筛选。AI 智能筛选新闻 + AI 翻译 + AI 分析简报直推手机,也支持接入 MCP 架构…

Python 61,467 24,868 Updated Jul 17, 2026

Memori is agent-native memory infrastructure. A LLM-agnostic layer that turns agent execution and conversation into structured, persistent state for production systems. Built for enterprise, Memori…

Python 15,950 3,077 Updated Aug 14, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,784 605 Updated Aug 14, 2026

verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework

Python 22,957 4,405 Updated Aug 14, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,816 7,888 Updated Aug 14, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,064 20,699 Updated Aug 14, 2026

A programming framework for agentic AI

Python 60,424 9,106 Updated Apr 15, 2026

LLM101n: Let's build a Storyteller

37,506 2,076 Updated Aug 1, 2024

Minimal, clean code for the Byte Pair Encoding (BPE) algorithm commonly used in LLM tokenization.

Python 10,677 1,092 Updated Jul 1, 2024

ChatGPT资料汇总学习,持续更新......

4,185 383 Updated Jun 4, 2025
Next