Lists (1)
Sort Name ascending (A-Z)
Starred repositories
😘 让你“爱”上 GitHub,解决访问时图裂、加载慢的问题。(无需安装)
A curated list of state-of-the-art research in embodied AI, focusing on vision-language-action (VLA) models, vision-language navigation (VLN), and related multimodal learning approaches.
1 min voice data can also be used to train a good TTS model! (few shot voice cloning)
Robust Speech Recognition via Large-Scale Weak Supervision
PyTorch version of Stable Baselines, reliable implementations of reinforcement learning algorithms.
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
Open-source book with Modern CUDA Learn Notes for Beginners, includes FP16/BF16, FP8, HGEMM, FlashAttention, CuTe, etc.
openvla / openvla
Forked from TRI-ML/prismatic-vlmsOpenVLA: An open-source vision-language-action model for robotic manipulation.
moojink / openvla-oft
Forked from openvla/openvlaFine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
Machine learning metrics for distributed, scalable PyTorch applications.
Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.
Seamless operability between C++11 and Python
Open-Sora: Democratizing Efficient Video Production for All
一个支持windows/linux/mac的文本编辑器,目标是做中国人自己的编辑器,来自中国。
Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.
💥 Fast State-of-the-Art Tokenizers optimized for Research and Production
Semantic segmentation models with 500+ pretrained convolutional and transformer-based backbones.
📚A curated list of Awesome LLM/VLM Inference Papers with Codes: Flash-Attention, Paged-Attention, WINT8/4, Parallelism, etc.🎉
Medusa: Simple Framework for Accelerating LLM Generation with Multiple Decoding Heads
This project aim to reproduce Sora (Open AI T2V model), we wish the open source community contribute to this project.
Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama mode…
Aligning pretrained language models with instruction data generated by themselves.
Ongoing research training transformer models at scale