Skip to content
View lvbu12's full-sized avatar

Block or report lvbu12

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A Pocket-Sized MLLM for Ultra-Efficient Image and Video Understanding on Your Phone

Python 26,175 2,049 Updated Aug 12, 2026
Python 88 8 Updated Jun 17, 2026

MiMo-Audio: Audio Language Models are Few-Shot Learners

Python 1,076 107 Updated Jun 17, 2026

slime is an LLM post-training framework for RL Scaling.

Python 8,073 1,151 Updated Aug 16, 2026

OpenClaw-RL: Train any agent simply by talking

Python 5,638 609 Updated May 23, 2026

AI agent toolkit: unified LLM API, agent loop, TUI, coding agent CLI

TypeScript 91,805 11,380 Updated Aug 17, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,504 81,217 Updated Aug 17, 2026

Awesome papers for role-playing with language models

230 11 Updated Nov 3, 2024

Official Code for "Coser: Coordinating LLM-Based Persona Simulation of Established Roles"

Python 217 13 Updated Apr 2, 2026

A Systematic Survey of Deep Research

323 17 Updated Jan 1, 2026

Ming - facilitating advanced multimodal understanding and generation capabilities built upon the Ling LLM.

Jupyter Notebook 667 59 Updated Jul 27, 2026

Code repo for "Harnessing Negative Signals: Reinforcement Distillation from Teacher Data for LLM Reasoning"

Python 34 2 Updated Jul 25, 2025

Code for our NeurIPS'24 Dataset and Benchmark paper: Cooperation, Competition, and Maliciousness: LLM-Stakeholders Interactive Negotiation

Jupyter Notebook 54 15 Updated Nov 11, 2024

[EMNLP 2025] Evaluating Behavioral Alignment in Conflict Dialogue: A Multi-Dimensional Comparison of LLM Agents and Humans (EMNLP 2025 Main)

Python 4 Updated Nov 8, 2025

a toolkit on knowledge distillation for large language models

Python 454 44 Updated Aug 17, 2026

Post-training with Tinker

Python 4,023 512 Updated Aug 16, 2026

This repository contains resources for accessing the official benchmarks, codes, and checkpoints of the paper: "[**Breaking Language Barriers in Multilingual Mathematical Reasoning: Insights and Ob…

Python 49 5 Updated Jul 29, 2024

MultilingualSIFT: Multilingual Supervised Instruction Fine-tuning

Python 97 6 Updated Aug 15, 2023

An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset

Python 27 6 Updated Jan 19, 2025
7 Updated Nov 16, 2023

RoleInteract: Evaluating the Social Interaction of Role-Playing Agents

Python 70 5 Updated Oct 12, 2024

Official implementation for DenseMixer: Improving MoE Post-Training with Precise Router Gradient

Python 68 7 Updated Aug 3, 2025

Contexts Optical Compression

Python 23,812 2,200 Updated Jan 27, 2026
Jupyter Notebook 227 14 Updated Dec 23, 2025
Python 17 5 Updated Feb 3, 2025

XLand-100B: A Large-Scale Multi-Task Dataset for In-Context Reinforcement Learning - - — ICLR 2025

Python 85 11 Updated Feb 13, 2025

Official repo for the paper "Scaling Synthetic Data Creation with 1,000,000,000 Personas"

Python 1,637 131 Updated Feb 19, 2025

BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)

HTML 8,278 756 Updated Oct 16, 2024
Python 372 39 Updated May 17, 2024
Next