Skip to content
View jinyangwu's full-sized avatar

Block or report jinyangwu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 176 6 Updated Jul 17, 2026

[ACL 2026 Best Paper Candidate (SAC Highlight / Oral)] Official implementation of "Double: Breaking the Acceleration Limit via Double Retrieval Speculative Parallelism".

Python 3 1 Updated Jun 29, 2026

A collection of useful Github repositories. Github项目精选。

HTML 120 19 Updated Oct 11, 2025

AcadHomepage: A Modern and Responsive Academic Personal Homepage

SCSS 2,865 5,785 Updated Jul 19, 2026
Python 100 3 Updated Jul 1, 2026

Official Repository of Orchestra-o1: Omnimodal Agent Orchestration

Python 68 3 Updated Jun 15, 2026

将博导十年科研经验炼化为可直接调用的 AI 技能。从 Idea 构思到论文投稿,你的 AI 科研副导师。

Python 4,310 300 Updated Jul 16, 2026

🎵 将音乐搬进AI终端 · 打造Coding专属智能DJ · 听歌新范式 · 交互式音乐 |A retro pixel DJ for Claude Code / Codex CLI / Terminal with Apple Music / QQ Music / Local Music — plays music, shows synced lyrics, and 😱 panics when…

Python 114 2 Updated Jun 30, 2026

This is the official repo for the paper "General365: Benchmarking General Reasoning in LLMs under High Difficulty and Diversity".

Python 85 5 Updated Apr 14, 2026

This is the official repo for the paper "AMO-Bench: Large Language Models Still Struggle in High School Math Competitions".

Python 160 4 Updated Feb 6, 2026

Official code for "Self-Distilled Agentic Reinforcement Learning"

Python 313 24 Updated Jul 23, 2026

Official code for "KnowU-Bench: Towards Interactive, Proactive, and Personalized Mobile Agent Evaluation"

Python 76 2 Updated Jun 13, 2026

Official code for "SKILL0: In-Context Agentic Reinforcement Learning for Skill Internalization"

Python 355 15 Updated Jul 22, 2026

主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题

HTML 14,754 1,459 Updated Jun 14, 2026

Extremely Long-Horizon Agentic Tasks Requiring Active Acting and Inductive Reasoning

Python 33 2 Updated Feb 9, 2026

ICLR2026 SAFER: Risk-Constrained Sample-then-Filter in Large Language Models

Python 7 1 Updated Feb 14, 2026

SafeGround: Know When to Trust GUI Grounding Models via Uncertainty Calibration

Python 9 2 Updated Jun 19, 2026

😼 优雅地使用基于 clash/mihomo 的代理环境

Shell 14,211 1,586 Updated Jul 14, 2026

Draft-Target Disaggregation LLM Serving System via Parallel Speculative Decoding.

Python 211 32 Updated Mar 18, 2026

Official Repo for Open-Reasoner-Zero

Python 2,096 121 Updated Jun 2, 2025

Sky-T1: Train your own O1 preview model within $450

Python 3,396 344 Updated Jul 12, 2025

Minimal reproduction of DeepSeek R1-Zero

Python 13,203 1,582 Updated Feb 27, 2026

Official code for the paper, "Stop Summation: Min-Form Credit Assignment Is All Process Reward Model Needs for Reasoning"

Python 172 7 Updated Oct 23, 2025

s1: Simple test-time scaling

Python 6,660 756 Updated Jun 25, 2025

🔍 Search-o1: Agentic Search-Enhanced Large Reasoning Models [EMNLP 2025]

Python 1,240 106 Updated Nov 17, 2025

[NeurIPS 2024] The official implementation of paper: Chain of Preference Optimization: Improving Chain-of-Thought Reasoning in LLMs.

Python 137 10 Updated Mar 21, 2025

[NIPS'25 Spotlight] Mulberry, an o1-like Reasoning and Reflection MLLM Implemented via Collective MCTS

Python 1,244 113 Updated Jan 16, 2026

[CVPR 2024 Oral] InternVL Family: A Pioneering Open-Source Alternative to GPT-4o. 接近GPT-4o表现的开源多模态对话模型

Python 10,101 788 Updated Sep 22, 2025

[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.

Python 24,941 2,775 Updated Aug 12, 2024
Next