Skip to content
View dongyh20's full-sized avatar

Block or report dongyh20

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

PerceptionBench: Evaluating Atomic Visual Perception in Multimodal Large Language Models

Python 181 9 Updated Aug 3, 2026

FlashKDA: high-performance Kimi Delta Attention kernels

Cuda 1,192 112 Updated Jul 30, 2026

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 3,032 248 Updated Aug 8, 2026

MoonEP: A Perfectly Balanced Expert Parallelism Library via Dynamic Redundant Experts

Python 1,054 114 Updated Aug 7, 2026

Open Frontier Intelligence

8,287 633 Updated Aug 6, 2026

Apple-π: Benchmarking Thinking with Video Towards Law-Grounded Physical Intelligence

Python 73 1 Updated Aug 2, 2026

A generalist video MLLM built for fine-grained motion, long-form reasoning, temporal grounding, and online proactive response.

Python 416 6 Updated Jul 27, 2026

[ECCV2026] ViQ: Text-Aligned Visual Quantized Representations at Any Resolution

Python 82 5 Updated Jul 1, 2026

Measuring frontier coding agents on original, long-horizon engineering tasks

Python 1,338 86 Updated Aug 6, 2026

S-Agent: Spatial Tool-Use Elicits Reasoning for Spatial Intelligence

84 Updated Jul 22, 2026

From Vision-Language-Action Models to a Real-World Robot Learning Stack

Python 262 21 Updated Aug 4, 2026

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

Python 890 61 Updated Aug 9, 2026

Agents' Last Exam

Python 940 57 Updated Aug 5, 2026

Kimi Code CLI — The Starting Point for Next-Gen Agents

TypeScript 6,228 987 Updated Aug 8, 2026

A collection of skills for AI financial analysis.

JavaScript 3,137 366 Updated Aug 5, 2026

My learning notes for ML SYS.

HTML 6,843 474 Updated Aug 8, 2026

[ECCV 2026] Official code of GEM: Generative Supervision Helps Embodied Intelligence

Python 91 1 Updated May 30, 2026

PhysX-Omni: Unified Simulation-Ready Physical 3D Generation for Rigid, Deformable, and Articulated Objects

Jupyter Notebook 308 16 Updated Jun 11, 2026

Skill package for ML/CV/NLP paper writing, curated and adapted from Prof. Peng Sida's open notes for Codex, Claude Code, and Gemini.

5,910 289 Updated Jun 23, 2026

Can Language Models Rebuild Programs From Scratch?

Python 884 62 Updated Jul 26, 2026

Beyond SFT-to-RL: Pre-alignment via Black-BoxOn-Policy Distillation for Multimodal RL

Python 98 2 Updated May 6, 2026

A benchmark for evaluating LLMs on Chinese traditional fortune telling — Bazi (八字) and Ziwei Doushu (紫微斗数).

Python 2,270 330 Updated May 9, 2026

📰 Must-read papers and blogs on LLM based Long Context Modeling 🔥

2,160 101 Updated Aug 6, 2026

Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Curso…

JavaScript 62,632 10,293 Updated Aug 7, 2026

SenseNova-U series: Native Unified Paradigm with NEO-unify from the First Principles

Python 4,597 397 Updated Aug 7, 2026

Reference code for the Meta-Harness paper.

Python 1,383 136 Updated Jul 11, 2026

Terrarium: Multi-turn data engine for evaluating and optimizing LLM agents in living environments.

Python 54 2 Updated Jul 14, 2026

🦞 ClawMark: A Living-World Benchmark for Multi-Day, Multimodal Coworker Agents

Python 120 11 Updated May 28, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

200,930 20,647 Updated Apr 20, 2026

The agent that grows with you

Python 227,953 44,771 Updated Aug 10, 2026
Next