Skip to content
View dongyh20's full-sized avatar

Block or report dongyh20

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

This is the code repository of paper "SFT Conflicts, RL Coexists"

Python 33 2 Updated Aug 4, 2026

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

Python 184 9 Updated Aug 3, 2026

FlashKDA: high-performance Kimi Delta Attention kernels

Cuda 1,208 114 Updated Jul 30, 2026

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 3,170 265 Updated Aug 12, 2026

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

Python 1,069 118 Updated Aug 7, 2026

Open Frontier Intelligence

8,395 651 Updated Aug 6, 2026

Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence

Python 75 1 Updated Aug 2, 2026

A generalist video MLLM built for fine-grained motion, long-form reasoning, temporal grounding, and online proactive response.

Python 436 7 Updated Jul 27, 2026

[ECCV2026] ViQ: Text-Aligned Visual Quantized Representations at Any Resolution

Python 82 5 Updated Jul 1, 2026

Measuring frontier coding agents on original, long-horizon engineering tasks

Python 1,361 88 Updated Aug 6, 2026

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

84 Updated Jul 22, 2026

From Vision-Language-Action Models to a Real-World Robot Learning Stack

Python 265 21 Updated Aug 4, 2026

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

Python 895 61 Updated Aug 12, 2026

Agents' Last Exam

Python 941 58 Updated Aug 5, 2026

Kimi Code CLI — The Starting Point for Next-Gen Agents

TypeScript 6,463 1,025 Updated Aug 12, 2026

A collection of skills for AI financial analysis.

JavaScript 3,152 366 Updated Aug 5, 2026

My learning notes for ML SYS.

HTML 6,861 477 Updated Aug 12, 2026

[ECCV 2026] Official code of GEM: Generative Supervision Helps Embodied Intelligence

Python 91 1 Updated May 30, 2026

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects

Jupyter Notebook 309 16 Updated Jun 11, 2026

Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.

5,988 293 Updated Jun 23, 2026

Can Language Models Rebuild Programs From Scratch?

Python 887 62 Updated Jul 26, 2026

Beyond SFT-to-RL: Pre-alignment via Black-BoxOn-Policy Distillation for Multimodal RL

Python 99 2 Updated May 6, 2026

A benchmark for evaluating LLMs on Chinese traditional fortune telling — Bazi (八字) and Ziwei Doushu (紫微斗数).

Python 2,281 329 Updated May 9, 2026

📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥

2,162 102 Updated Aug 6, 2026

Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Curso…

JavaScript 62,790 10,314 Updated Aug 7, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,714 405 Updated Aug 13, 2026

Reference code for the Meta-Harness paper.

Python 1,400 138 Updated Jul 11, 2026

Terrarium: Multi-turn data engine for evaluating and optimizing LLM agents in living environments.

Python 55 2 Updated Jul 14, 2026

🦞 ClawMark: A Living-World Benchmark for Multi-Day, Multimodal Coworker Agents

Python 122 11 Updated May 28, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

201,904 20,716 Updated Apr 20, 2026
Next