Skip to content
View AberHu's full-sized avatar

Block or report AberHu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[SIGGRAPH‘2026] PEAR :Pixel-aligned Expressive humAn mesh Recovery

Python 305 28 Updated Aug 1, 2026

2026 好用的付费机场推荐

3,120 147 Updated Aug 5, 2026

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python 41,373 3,294 Updated Aug 8, 2026

RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Models

Jupyter Notebook 857 82 Updated Jul 22, 2026

from vibe coding to agentic engineering - practice makes claude perfect

HTML 64,164 6,376 Updated Aug 8, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

200,653 20,621 Updated Apr 20, 2026

同事.skill、老板.skill、前任.skill、自己.skill、永生.skill、女娲.skill……

3,564 316 Updated Jun 28, 2026

Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1

Python 73,562 11,925 Updated Jul 28, 2026

Evaluate and improve models and agents using environments

Python 1,094 259 Updated Aug 8, 2026

Scalable toolkit for efficient model reinforcement

Python 1,890 500 Updated Aug 8, 2026

VILA is a family of state-of-the-art vision language models (VLMs) for diverse multimodal AI tasks across the edge, data center, and cloud.

Python 3,850 330 Updated Mar 12, 2026

LEAKED SYSTEM PROMPTS FOR CHATGPT, CLAUDE, GEMINI, GROK, PERPLEXITY, CURSOR, LOVABLE, REPLIT, AND MORE! - AI SYSTEMS TRANSPARENCY FOR ALL! 👐

46,801 9,518 Updated Jul 24, 2026

Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.

Python 1,511 173 Updated Aug 7, 2026

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,208 3,239 Updated Jun 11, 2026

Mobile-Agent: The Powerful GUI Agent Family

Python 9,057 908 Updated Jul 7, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,086 1,582 Updated Aug 8, 2026

A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related webs…

Python 1,951 68 Updated Jul 31, 2026

The official implement of VITA, VITA15, LongVITA, VITA-Audio, VITA-VLA, and VITA-E.

Python 164 4 Updated Oct 28, 2025

Fully Open Framework for Democratized Multimodal Training

Python 1,165 78 Updated Aug 8, 2026

[CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations

Python 592 20 Updated Apr 22, 2026

PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.

Java 28,269 2,698 Updated Aug 7, 2026

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,476 131 Updated Aug 1, 2026

​​Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation

Python 7,582 1,329 Updated May 22, 2026

🦛 CHONK docs with Chonkie ✨ — The lightweight ingestion library for fast, efficient and robust RAG pipelines

Python 4,653 351 Updated Aug 8, 2026

Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.

Jupyter Notebook 3,934 282 Updated Apr 23, 2026

E2M converts various file types (doc, docx, epub, html, htm, url, pdf, ppt, pptx, mp3, m4a) into Markdown. It’s easy to install, with dedicated parsers and converters, supporting custom configs. E2…

Jupyter Notebook 1,294 74 Updated Sep 8, 2024

An open-source implementaion for fine-tuning Qwen-VL series by Alibaba Cloud.

Python 1,951 220 Updated Jul 25, 2026

Repair invalid JSON documents

TypeScript 2,392 88 Updated Jul 3, 2026

Video-R1: Reinforcing Video Reasoning in MLLMs [🔥the first paper to explore R1 for video]

Python 886 46 Updated Dec 14, 2025

Official repository of 'Visual-RFT: Visual Reinforcement Fine-Tuning' & 'Visual-ARFT: Visual Agentic Reinforcement Fine-Tuning'’

Jupyter Notebook 2,266 110 Updated Oct 29, 2025
Next