Skip to content
View np-csu's full-sized avatar
🙂
thinking and programming
🙂
thinking and programming
  • Central South University
  • China

Block or report np-csu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Swift 9 Updated Jun 17, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,120 1,586 Updated Aug 11, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.

Rust 195,044 109,200 Updated Aug 6, 2026

[CVPR 2026] PromptEnhancer is a prompt-rewriting tool, refining prompts into clearer, structured versions for better image generation.

Python 3,751 325 Updated Jun 10, 2026

📹 A more flexible framework that can generate videos at any resolution and creates videos from images.

Python 2,193 169 Updated Aug 11, 2026

We introduce BabyVision, a benchmark revealing the infancy of AI vision.

Python 237 10 Updated Jan 13, 2026

omo/lazycodex: The coding agent for tokenmaxxers;the one and only agent harness for complex codebases. For your Codex, for your OpenCode

TypeScript 67,709 5,531 Updated Aug 11, 2026

基于深度强化学习的开源自动因子工厂。

Python 3,002 3,099 Updated Jun 12, 2026

Official implementation of "WorDepth: Variational Language Prior for Monocular Depth Estimation"

Python 48 4 Updated Feb 4, 2025

Get 10X more out of Claude Code, Codex or any coding agent

Rust 27,744 2,951 Updated Apr 24, 2026

[ECCV 2026] Towards Scalable Pre-training of Visual Tokenizers for Generation

Python 500 16 Updated Apr 15, 2026

A Cross-Platform Backend for High-Performance Sparse Convolutions

Python 142 48 Updated Jun 25, 2026

A Plugin-Based Multi-Agent System for In-Editor Academic Writing, Review, and Editing

TypeScript 1,524 72 Updated Jul 3, 2026

[ArXiv 2025] MobileI2V: Fast and High-Resolution Image-to-Video on Mobile Devices

Python 91 5 Updated May 20, 2026

Kandinsky 5.0: A family of diffusion models for Video & Image generation

Python 808 62 Updated Aug 7, 2026

Lightweight Image Video Action Generation Inference Framework

Python 2,638 251 Updated Aug 11, 2026

The repository provides code for running inference and finetuning with the Meta Segment Anything Model 3 (SAM 3), links for downloading the trained model checkpoints, and example notebooks that sho…

Python 11,282 1,703 Updated Jul 31, 2026

[NeurIPS 2025] Official implementation of ScaleDiff: Higher-Resolution Image Synthesis via Efficient and Model-Agnostic Diffusion

Python 7 Updated Oct 31, 2025

PyTorch implementation of JiT https://arxiv.org/abs/2511.13720

Python 2,490 166 Updated Dec 8, 2025

[CVPR 2026] Scaling Spatial Intelligence with Multimodal Foundation Models

Python 294 16 Updated May 14, 2026

Local UI to run and train LLMs and diffusion models, including Kimi K3, MiniMax-H3, Gemma 4, Qwen3.6, DeepSeek-V4, FLUX and more.

Python 70,208 6,331 Updated Aug 11, 2026

Depth Anything 3

Python 6,093 676 Updated Jul 27, 2026

Automatically crawl arXiv papers daily and summarize them using AI. Illustrating them using GitHub Pages.

JavaScript 2,935 1,089 Updated Aug 11, 2026

Cambrian-S: Towards Spatial Supersensing in Video

Python 564 20 Updated Apr 3, 2026

[NeurIPS 2025 Oral]Infinity⭐️: Unified Spacetime AutoRegressive Modeling for Visual Generation

Python 775 28 Updated Apr 16, 2026

Interactive, editable docs designed for coding agents

TypeScript 1,658 120 Updated Jan 19, 2026

collection of diffusion model papers categorized by their subareas

2,223 102 Updated Mar 16, 2026

Open-source AI penetration testing tool to find and fix your app’s vulnerabilities.

Python 51,202 5,484 Updated Aug 11, 2026

[WACV'25 Oral] Fine-Tuning Image-Conditional Diffusion Models is Easier than You Think

Python 521 22 Updated Jul 9, 2026
Next