Skip to content
View AberHu's full-sized avatar

Block or report AberHu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[SIGGRAPH‘2026] PEAR :Pixel-aligned Expressive humAn mesh Recovery

Python 308 29 Updated Aug 1, 2026

2026 好用的付费机场推荐

3,154 148 Updated Aug 5, 2026

Academic Research Skills for Claude Code: research → write → review → revise → finalize

Python 42,138 3,355 Updated Aug 11, 2026

RynnBrain 1.1: Towards More Capable and Generalizable Embodied Foundation Models

Jupyter Notebook 860 82 Updated Aug 10, 2026

from vibe coding to agentic engineering - practice makes claude perfect

HTML 64,373 6,390 Updated Aug 12, 2026

A single CLAUDE.md file to improve Claude Code behavior, derived from Andrej Karpathy's observations on LLM coding pitfalls.

201,777 20,704 Updated Apr 20, 2026

同事.skill、老板.skill、前任.skill、自己.skill、永生.skill、女娲.skill……

3,610 318 Updated Aug 10, 2026

Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1

Python 73,951 11,984 Updated Aug 12, 2026

Evaluate and improve models and agents using environments

Python 1,108 263 Updated Aug 12, 2026

Scalable toolkit for efficient model reinforcement

Python 1,900 511 Updated Aug 12, 2026

VILA is a family of state-of-the-art vision language models (VLMs) for diverse multimodal AI tasks across the edge, data center, and cloud.

Python 3,852 330 Updated Mar 12, 2026

LEAKED SYSTEM PROMPTS FOR CHATGPT, CLAUDE, GEMINI, GROK, PERPLEXITY, CURSOR, LOVABLE, REPLIT, AND MORE! - AI SYSTEMS TRANSPARENCY FOR ALL! 👐

46,863 9,527 Updated Jul 24, 2026

Model compression toolkit engineered for enhanced usability, comprehensiveness, and efficiency.

Python 1,522 172 Updated Aug 7, 2026

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,252 3,244 Updated Jun 11, 2026

Mobile-Agent: The Powerful GUI Agent Family

Python 9,074 908 Updated Jul 7, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,134 1,586 Updated Aug 12, 2026

A comprehensive list of papers for the definition of World Models and using World Models for General Video Generation, Embodied AI, and Autonomous Driving, including papers, codes, and related webs…

Python 1,954 69 Updated Jul 31, 2026

The official implement of VITA, VITA15, LongVITA, VITA-Audio, VITA-VLA, and VITA-E.

Python 164 4 Updated Oct 28, 2025

Fully Open Framework for Democratized Multimodal Training

Python 1,171 78 Updated Aug 12, 2026

[CVPR 2026] SpatialVID: A Large-Scale Video Dataset with Spatial Annotations

Python 595 20 Updated Apr 22, 2026

PDF Parser for AI-ready data. Automate PDF accessibility. Open-source.

Java 28,360 2,706 Updated Aug 10, 2026

A Survey of Reinforcement Learning for Large Reasoning Models

TeX 2,478 132 Updated Aug 1, 2026

​​Unlimited-length talking video generation​​ that supports image-to-video and video-to-video generation

Python 7,615 1,332 Updated May 22, 2026

🦛 CHONK docs with Chonkie ✨ — The lightweight ingestion library for fast, efficient and robust RAG pipelines

Python 4,664 351 Updated Aug 8, 2026

Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.

Jupyter Notebook 3,947 282 Updated Apr 23, 2026

E2M converts various file types (doc, docx, epub, html, htm, url, pdf, ppt, pptx, mp3, m4a) into Markdown. It’s easy to install, with dedicated parsers and converters, supporting custom configs. E2…

Jupyter Notebook 1,294 74 Updated Sep 8, 2024

An open-source implementaion for fine-tuning Qwen-VL series by Alibaba Cloud.

Python 1,956 220 Updated Jul 25, 2026

Repair invalid JSON documents

TypeScript 2,393 88 Updated Jul 3, 2026

Video-R1: Reinforcing Video Reasoning in MLLMs [🔥the first paper to explore R1 for video]

Python 886 46 Updated Dec 14, 2025

Official repository of 'Visual-RFT: Visual Reinforcement Fine-Tuning' & 'Visual-ARFT: Visual Agentic Reinforcement Fine-Tuning'’

Jupyter Notebook 2,268 110 Updated Oct 29, 2025
Next