Skip to content
View SOTAMak1r's full-sized avatar
🎯
Focusing
🎯
Focusing

Block or report SOTAMak1r

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Awesome MiniMax-H3

169 6 Updated Aug 12, 2026

MAGI-2-preview: Scaling Video Generation Models Efficiently

Python 458 12 Updated Aug 6, 2026

Piecewise-Taylor Attention

Python 41 2 Updated Aug 4, 2026

Open-source Dreamer world-model implementation in JAX

Python 350 26 Updated Aug 5, 2026

Open-source agentic video editing skills — an AI-powered alternative to Opus Clip and CapCut

Python 10 Updated Aug 11, 2026

Matrix-Game 3.5: Enhancing Real-Time Streaming Interactive World Models with Patch Memory

Python 132 3 Updated Jul 27, 2026

UniRL is a Framework for Unified Multimodal Model Reinforcement Learning

Python 895 61 Updated Aug 13, 2026

Training library for Megatron-based models with bidirectional Hugging Face conversion capability

Python 859 451 Updated Aug 13, 2026

Official Code of NAVA: Native Audio-Visual Alignment for Generation.

Python 220 25 Updated Jun 30, 2026

Kandinsky 5.0: A family of diffusion models for Video & Image generation

Python 808 62 Updated Aug 7, 2026

A Distributed Attention Towards Linear Scalability for Ultra-Long Context, Heterogeneous Data Training

Python 908 65 Updated Aug 11, 2026

The first MoE-based framework for unified and scalable world modeling.

Python 23 1 Updated Jul 16, 2026

Scaling Mixture-of-Experts Video Pretraining for Embodied Intelligence

Python 918 45 Updated Aug 5, 2026

Infinite Worlds with Versatile Interactions

Python 1,504 107 Updated Jul 14, 2026

Open-source native multimodal pretraining — without catastrophic forgetting.

Python 15 1 Updated Jul 2, 2026

Code for MIRA: Multiplayer Interactive World Models with Representation Autoencoders

Python 498 29 Updated Aug 3, 2026

Claude Code plugin: automated code review loop with Codex

Shell 716 45 Updated Mar 15, 2026

KVAE-Audio: a continuous full-band audio waveform autoencoder

Python 116 6 Updated Aug 10, 2026

Humanizer 的汉化版本,Claude Code Skills,旨在消除文本中 AI 生成的痕迹。

15,176 1,030 Updated Jan 19, 2026

Agent skill that removes signs of AI-generated writing from text

Python 35,330 3,160 Updated Jul 22, 2026
Python 67 Updated Aug 7, 2026

[NeurIPS 2025] PyTorch implementation of [ThinkSound], a unified framework for generating audio from any modality, guided by Chain-of-Thought (CoT) reasoning.

Python 1,375 82 Updated Apr 3, 2026

"ViMax: Agentic Video Generation (Director, Screenwriter, Producer, and Video Generator All-in-One)"

Python 11,903 1,780 Updated Jul 29, 2026

Boogu-Image-0.1 is an Apache-2.0 open-source image generation and editing model family that delivers near-closed-source performance with an order of magnitude less data.

Python 948 56 Updated Jul 23, 2026

Code release for "i1: A Simple and Fully Open Recipe for Strong Text-to-Image Models"

Python 264 15 Updated Aug 12, 2026

AI turns documents or topics into real, native PowerPoint decks—with native shapes, transitions and animations, data-backed charts and tables on demand, audio narration from speaker notes, and supp…

Python 45,853 3,730 Updated Aug 12, 2026

World Model Self-Distillation project website

19 3 Updated Jun 15, 2026

Ideogram 4: Open image model at the forefront of design

Python 2,721 279 Updated Jun 30, 2026

From Automated Idea Factory to Realization

Shell 1,385 124 Updated Jul 18, 2026
Next