Skip to content
View zzxslp's full-sized avatar

Highlights

  • Pro

Block or report zzxslp

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

CMU 11-711 Advanced NLP https://cmu-l3.github.io/anlp-spring2026/

Jupyter Notebook 119 14 Updated Mar 31, 2026

Procedural generation with diffusion models (SIGGRAPH '26)

Python 1,254 76 Updated Jul 21, 2026

Fine-grained computation offload for off-the-shelf servers in tens of lines — paper (arXiv:2607.02630), code, and every measurement. Overlap accelerator offloads (GPU/HSM/inference) with other requ…

C 7 Updated Jul 8, 2026

DreamX-World: A General-Purpose Interactive World Model

Python 735 48 Updated Jul 23, 2026

基于Python的开源量化交易平台开发框架

Python 44,031 12,277 Updated Jul 28, 2026

Modular Cognitive Architecture Emerges in Large Language Models

Python 66 7 Updated Jul 24, 2026

100M tokens. Infinite compute. Lowest val loss wins.

Python 518 79 Updated Jul 3, 2026

The official repo of EurekaClaw

Python 697 72 Updated Jun 13, 2026

Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1

Python 72,643 11,784 Updated Jul 28, 2026

Autoresearch for GPU kernels. Give it any PyTorch model, go to sleep, wake up to optimized Triton kernels.

Python 1,488 155 Updated Mar 19, 2026

AI-Driven Scientific and Algorithmic Discovery

Python 591 85 Updated Jun 14, 2026

40+ tips for getting the most out of Claude Code, from basics to advanced - includes a custom status line script and Claude Code running itself in a container. Also includes the dx plugin: skills f…

HTML 9,475 748 Updated Jul 25, 2026

World Model Inference Engine

Python 176 28 Updated Jun 5, 2026

A zero-to-one guide on scaling modern transformers with n-dimensional parallelism.

Python 127 9 Updated Dec 29, 2025

Advancing Open-source World Models

Python 4,304 397 Updated Jul 9, 2026

Efficient Long-context Language Model Training by Core Attention Disaggregation

Python 106 7 Updated Apr 7, 2026

Miles is an enterprise-facing reinforcement learning framework for LLM and VLM post-training, forked from and co-evolving with slime.

Python 1,810 333 Updated Jul 29, 2026

Official JAX implementation of End-to-End Test-Time Training for Long Context

Python 627 47 Updated Feb 15, 2026

Repo for Qwen Image Finetune

Jupyter Notebook 249 27 Updated Jun 8, 2026

Sutskever 30 implementations inspired by https://papercode.vercel.app/ | For Agents, use https://github.com/pageman/Sutskever-Agent | Polyglot / Multi-Backed version at https://github.com/pageman/s…

Jupyter Notebook 4,205 546 Updated Mar 15, 2026

Machine Learning Systems

Python 27,636 3,387 Updated Jul 28, 2026

My learning notes for ML SYS.

HTML 6,789 471 Updated Jul 26, 2026

HY-World 1.5: A Systematic Framework for Interactive World Modeling with Real-Time Latency and Geometric Consistency

Python 1,564 141 Updated Jun 10, 2026

Type annotations and runtime checking for shape and dtype of JAX/NumPy/PyTorch/etc. arrays. https://docs.kidger.site/jaxtyping/

Python 1,847 93 Updated Jul 8, 2026

Based on Nano-vLLM, a simple replication of vLLM with self-contained paged attention and flash attention implementation

Python 944 150 Updated Jul 22, 2026

Simple and efficient DeepSeek V3 SFT using pipeline parallel and expert parallel, with both FP8 and BF16 trainings

Python 118 18 Updated Jul 27, 2025
Python 1,743 202 Updated Nov 15, 2025

This repository contains a curated collection of 300+ case studies from over 80 companies, detailing practical applications and insights into machine learning (ML) system design. The contents are o…

10,852 1,666 Updated Aug 5, 2025

Minimal PDF creation library. <400 LOC, zero dependencies, makes real PDFs.

TypeScript 1,873 74 Updated May 13, 2026

[ICLR2025, ICML2025, NeurIPS2025 Spotlight] Quantized Attention achieves speedup of 2-5x compared to FlashAttention, without losing end-to-end metrics across language, image, and video models.

Cuda 3,522 447 Updated Jan 17, 2026
Next