Skip to content
View AlmoonYsl's full-sized avatar

Highlights

  • Pro

Block or report AlmoonYsl

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
13 stars written in Jupyter Notebook
Clear filter

A latent text-to-image diffusion model

Jupyter Notebook 73,332 10,579 Updated Jun 18, 2024

Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.

Jupyter Notebook 19,816 1,834 Updated Jan 30, 2026

Grounded SAM: Marrying Grounding DINO with Segment Anything & Stable Diffusion & Recognize Anything - Automatically Detect , Segment and Generate Anything

Jupyter Notebook 17,710 1,595 Updated Sep 5, 2024

NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.

Jupyter Notebook 11,578 838 Updated Aug 21, 2026

Reference PyTorch implementation and models for DINOv3

Jupyter Notebook 11,223 932 Updated Jul 15, 2026

Taming Transformers for High-Resolution Image Synthesis

Jupyter Notebook 6,520 1,220 Updated Jul 30, 2024

Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation

Jupyter Notebook 4,464 734 Updated Jun 22, 2024

A data generation pipeline for creating semi-realistic synthetic multi-object videos with rich annotations such as instance segmentation masks, depth maps, and optical flow.

Jupyter Notebook 2,806 281 Updated May 21, 2026

Official repository of 'Visual-RFT: Visual Reinforcement Fine-Tuning' & 'Visual-ARFT: Visual Agentic Reinforcement Fine-Tuning'’

Jupyter Notebook 2,267 110 Updated Oct 29, 2025

A suite of image and video neural tokenizers

Jupyter Notebook 1,732 93 Updated Feb 11, 2025

[CVPR 2024] LMDrive: Closed-Loop End-to-End Driving with Large Language Models

Jupyter Notebook 929 77 Updated Apr 14, 2025

[AAAI2024] Far3D: Expanding the Horizon for Surround-view 3D Object Detection

Jupyter Notebook 199 20 Updated Dec 13, 2023

[NeurIPS 2023] Dynamo-Depth: Fixing Unsupervised Depth Estimation for Dynamical Scenes

Jupyter Notebook 89 6 Updated Jan 2, 2025