Skip to content
View jinyangwu's full-sized avatar

Block or report jinyangwu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 16 1 Updated Aug 8, 2026

From Proprietary to Open-Source: Bridging the Distribution Gap via Multi-Agent Protocol Distillation in Agentic Search

43 2 Updated Jul 21, 2026
Python 220 8 Updated Jul 17, 2026

[ACL 2026 Best Paper Candidate (SAC Highlight / Oral)] Official implementation of "Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism".

Python 3 1 Updated Jun 29, 2026

A collection of useful Github repositories. Github项目精选。

HTML 121 19 Updated Oct 11, 2025

AcadHomepage: A Modern and Responsive Academic Personal Homepage

SCSS 2,890 5,808 Updated Aug 8, 2026
Python 108 2 Updated Jul 1, 2026

Official Repository of Orchestra-o1: Omnimodal Agent Orchestration

Python 70 3 Updated Jun 15, 2026

将博导十年科研经验炼化为可直接调用的 AI 技能。从 Idea 构思到论文投稿,你的 AI 科研副导师。

Python 5,283 362 Updated Aug 7, 2026

🎵 将音乐搬进AI终端 · 打造Coding专属智能DJ · 听歌新范式 · 交互式音乐 |A retro pixel DJ for Claude Code / Codex CLI / Terminal with Apple Music / QQ Music / Local Music — plays music, shows synced lyrics, and 😱 panics when…

Python 113 2 Updated Jun 30, 2026

This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".

Python 88 5 Updated Apr 14, 2026

This is the official repo for the paper "AMO-Bench: Large Language Models Still Struggle in High School Math Competitions".

Python 178 4 Updated Feb 6, 2026

Official code for "Self-Distilled Agentic Reinforcement Learning"

Python 329 25 Updated Aug 9, 2026

Official code for "KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation"

Python 75 2 Updated Jun 13, 2026

Official code for "SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization"

Python 360 16 Updated Aug 7, 2026

主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题

HTML 14,861 1,460 Updated Jun 14, 2026

Extremely Long-Horizon Agentic Tasks Requiring Active Acting and Inductive Reasoning

Python 33 2 Updated Feb 9, 2026

ICLR2026 SAFER: Risk-Constrained Sample-then-Filter in Large Language Models

Python 7 1 Updated Feb 14, 2026

SafeGround: Know When to Trust GUI Grounding Models via Uncertainty Calibration

Python 9 2 Updated Jun 19, 2026

😼 优雅地使用基于 clash/mihomo 的代理环境

Shell 14,380 1,599 Updated Jul 14, 2026

Draft-Target Disaggregation LLM Serving System via Parallel Speculative Decoding.

Python 214 32 Updated Mar 18, 2026

Official Repo for Open-Reasoner-Zero

Python 2,098 121 Updated Jun 2, 2025

Sky-T1: Train your own O1 preview model within $450

Python 3,398 344 Updated Jul 12, 2025

Minimal reproduction of DeepSeek R1-Zero

Python 13,219 1,580 Updated Feb 27, 2026

Official code for the paper, "Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning"

Python 173 7 Updated Oct 23, 2025

s1: Simple test-time scaling

Python 6,663 756 Updated Jun 25, 2025

🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]

Python 1,240 107 Updated Nov 17, 2025

[NeurIPS 2024] The official implementation of paper: Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs.

Python 137 10 Updated Mar 21, 2025
Next