-
Nanyang Technological University
Stars
Comprehensive open-source library of AI research and engineering skills for any AI model. Package the skills and your claude code/codex/gemini agent will be an AI research agent with full horsepowe…
The agent that grows with you
AI agents running research on single-GPU nanochat training automatically
HY-WU (Part I): An Extensible Functional Neural Memory Framework and An Instantiation in Text-Guided Image Editing
[ICLR'26] Grasp Any Region: Towards Precise, Contextual Pixel Understanding for Multimodal LLMs
Official implementation of "Generating Aligned Pseudo-Supervision from Non-Aligned Data for Image Restoration in Under-Display Camera"
[ECCV'24] Kalman-Inspired Feature Propagation for Video Face Super-Resolution
国内首个占据栅格网络全栈课程《从BEV到Occupancy Network,算法原理与工程实践》,包含端侧部署。Surrounding Semantic Occupancy Perception Course for Autonomous Driving (docs, ppt and source code) 在线课程主页:http://111.229.117.200:8100/ (作者独立搭建)
Official Repo For OMG-LLaVA and OMG-Seg codebase [CVPR-24 and NeurIPS-24]
This list of writing prompts covers a range of topics and tasks, including brainstorming research ideas, improving language and style, conducting literature reviews, and developing research plans.
🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.
[CVPR 2023] Collaborative Diffusion
Code for Text2Performer. Paper: Text2Performer: Text-Driven Human Video Generation
[SIGGRAPH Asia 2024] ReVersion: Diffusion-Based Relation Inversion from Images
[ICCV 2023] MOSE: A New Dataset for Video Object Segmentation in Complex Scenes
MulimgViewer is a multi-image viewer that can open multiple images in one interface, which is convenient for image comparison and image stitching.
[CVPR 2023] E3DGE: An encoder-based 3D StyleGAN inversion framework.
A simple CV made for graduate school applications and new graduate students. The CV is made using LateX and ModernCV. Adaption by Armin Hadzic. Original by Xavier Danaux
Code for Text2Human (SIGGRAPH 2022). Paper: Text2Human: Text-Driven Controllable Human Image Generation
PyTorch implementations of Generative Adversarial Networks.
🔎 🖼️ 🔥PyTorch Toolbox for Image Quality Assessment, including PSNR, SSIM, LPIPS, FID, NIQE, NRQM(Ma), MUSIQ, TOPIQ, NIMA, DBCNN, BRISQUE, PI and more...
[CVPR 2022 Oral] Crafting Better Contrastive Views for Siamese Representation Learning
Code for C2-Matching (CVPR2021). Paper: Robust Reference-based Super-Resolution via C2-Matching.