Skip to content
View WongJiayi's full-sized avatar
🎯
Focusing
🎯
Focusing
  • Friedrich-Alexander-Universität Erlangen-Nürnberg
  • Erlangen, Germany

Block or report WongJiayi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

[CVPR2026 Highlight] Cubic Discrete Diffusion: Discrete Visual Generation on High-Dimensional Representation Tokens https://arxiv.org/abs/2603.19232

Python 63 1 Updated Apr 10, 2026

DiffVP: Differential Visual Semantic Prompting for LLM-Based CT Report Generation

2 Updated Mar 17, 2026

[CVPR 2025🔥] Enhancing Video VAE by Wavelet-Driven Energy Flow for Latent Video Diffusion Model

Python 206 8 Updated May 11, 2025

🏛️ 三省六部制 · OpenClaw Multi-Agent Orchestration System — 9 specialized AI agents with real-time dashboard, model config, and full audit trails

Python 16,345 1,711 Updated Aug 3, 2026

A better open-source alternative to Typeless

Rust 55 5 Updated Aug 9, 2026

Diffusion model for synthetic 3D CT scan video generation — IDEA Lab, FAU Erlangen-Nürnberg

Python 4 Updated Mar 16, 2026

Auto-regressive 3D CT volume generation using Latent Video Flow Matching with a Spatial-Temporal DiT (STDiT), conditioned on CT report embeddings.

Python 4 Updated Mar 16, 2026

Image-to-Image Translation in PyTorch

Python 25,217 6,568 Updated Aug 6, 2025

Code for NeurIPS 2024 paper - The GAN is dead; long live the GAN! A Modern Baseline GAN - by Huang et al.

Python 870 46 Updated Jan 23, 2025

You Only Denoise once or Average (YODA) - A diffusion-based 2.5D medical image translation model with noise-supression

Python 14 1 Updated Jul 7, 2026

3D U-Net model for volumetric semantic segmentation written in pytorch

Jupyter Notebook 2,417 564 Updated Dec 16, 2025

Repo for MedSyn: Text-guided Anatomy-aware Synthesis of High-Fidelity 3D CT Images

Jupyter Notebook 60 3 Updated Jul 12, 2024
Jupyter Notebook 205 35 Updated Jul 20, 2026

A python package to streamline evaluation of unconditional image generation models

Python 17 1 Updated Apr 14, 2026

PyTorch re-implementation of FlowTok: Flowing Seamlessly Across Text and Image Tokens

Python 18 3 Updated Nov 26, 2025

This repo contains the code for 1D tokenizer and generator

Jupyter Notebook 1,172 69 Updated Mar 20, 2025
Python 24 3 Updated Dec 17, 2024

Muon is an optimizer for hidden layers in neural networks

Python 2,771 128 Updated May 24, 2026

Open-Sora: Democratizing Efficient Video Production for All

Python 29,258 3,003 Updated Apr 9, 2026

Caption free adapter that maps DINOv3 image embeddings into CLIP space so you can do zero-shot text -> image or image -> text with CLIP’s text tower

Python 49 1 Updated Sep 18, 2025

[CVPR 2022] StyleGAN-V: A Continuous Video Generator with the Price, Image Quality and Perks of StyleGAN2

Python 392 39 Updated Apr 19, 2023

[ECCV 2024] Official PyTorch implementation of RoPE-ViT "Rotary Position Embedding for Vision Transformer"

Python 468 16 Updated Oct 29, 2025

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 163,521 34,169 Updated Aug 10, 2026

Learnable Fourier Features for Multi-Dimensional Spatial Positional Encoding

Python 56 9 Updated Sep 30, 2024

An implementation of 1D, 2D, and 3D positional encoding in Pytorch and TensorFlow

Python 616 35 Updated Oct 23, 2024

Config files for my GitHub profile.

1 Updated Mar 25, 2026