Stars
A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone
MiMo-Audio: Audio Language Models are Few-Shot Learners
slime is an LLM post-training framework for RL Scaling.
OpenClaw-RL: Train any agent simply by talking
AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
Awesome papers for role-playing with language models
Official Code for "Coser: Coordinating LLM-Based Persona Simulation of Established Roles"
Ming - facilitating advanced multimodal understanding and generation capabilities built upon the Ling LLM.
Code repo for "Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning"
Code for our NeurIPS'24 Dataset and Benchmark paper: Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation
[EMNLP 2025] Evaluating Behavioral Alignment in Conflict Dialogue: A Multi-Dimensional Comparison of LLM Agents and Humans (EMNLP 2025 Main)
a toolkit on knowledge distillation for large language models
Post-training with Tinker
This repository contains resources for accessing the official benchmarks, codes, and checkpoints of the paper: "[**Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Ob…
MultilingualSIFT: Multilingual Supervised Instruction Fine-tuning
An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset
RoleInteract: Evaluating the Social Interaction of Role-Playing Agents
Official implementation for DenseMixer: Improving MoE Post-Training with Precise Router Gradient
XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning - - — ICLR 2025
Official repo for the paper "Scaling Synthetic Data Creation with 1,000,000,000 Personas"
BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)