Skip to content
View JingChunzhen's full-sized avatar
  • Beijing

Block or report JingChunzhen

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

AI agents running research on single-GPU nanochat training automatically

Python 93,429 13,273 Updated Mar 26, 2026

Static suckless single batch CUDA-only qwen3-0.6B mini inference engine

Cuda 555 47 Updated Sep 8, 2025

Examples for Recommenders - easy to train and deploy on accelerated infrastructure.

Python 296 79 Updated Aug 6, 2026

flash attention tutorial written in python, triton, cuda, cutlass

Cuda 532 55 Updated Jan 20, 2026

Efficient triton implementation of Native Sparse Attention.

Python 283 20 Updated May 23, 2025

Exploring Applications of GRPO

Python 252 34 Updated Aug 25, 2025

一个手把手教你从零开始编写GPT并训练大语言模型的教程

Jupyter Notebook 102 9 Updated Jan 20, 2025

🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!

Python 54,463 7,088 Updated Aug 6, 2026

Self-paced bootcamp on Generative AI. Tutorials on ML fundamentals, Ollama, LLMs, RAGs, LangChain, LangGraph, Fine-tuning, DSPy & AI Agents (CrewAI), (Using ChatGPT, gpt-oss, Claude, Qwen, Gemma, L…

Jupyter Notebook 931 290 Updated Jun 20, 2026

This is a repository used by individuals to experiment and reproduce the pre-training process of LLM.

Python 504 73 Updated May 1, 2025

GPU Kernels

Cuda 225 25 Updated Apr 27, 2025

100 days of building GPU kernels!

Cuda 625 78 Updated Apr 27, 2025

Course to get into Large Language Models (LLMs) with roadmaps and Colab notebooks.

81,541 9,493 Updated Feb 5, 2026

Simple RL training for reasoning

Python 3,872 285 Updated Dec 23, 2025

🧑‍🏫 60+ Implementations/tutorials of deep learning papers with side-by-side notes 📝; including transformers (original, xl, switch, feedback, vit, ...), optimizers (adam, adabelief, sophia, ...), ga…

Python 67,296 6,749 Updated Jan 22, 2026

Machine Learning Journal for Intermediate to Advanced Topics.

Jupyter Notebook 2,366 254 Updated Sep 8, 2025

test tensorrt c++/python api, c++/python plugins.

C++ 9 2 Updated Jul 23, 2021

Writing an OS in 1,000 lines.

C 3,556 286 Updated Jul 16, 2026

RAFT contains fundamental widely-used algorithms and primitives for machine learning and information retrieval. The algorithms are CUDA-accelerated and form building blocks for more easily writing …

Cuda 1,036 244 Updated Aug 7, 2026

[🔥updating ...] AI 自动量化交易机器人(完全本地部署) AI-powered Quantitative Investment Research Platform. 📃 online docs: https://ufund-me.github.io/Qbot ✨ :news: qbot-mini: https://github.com/Charmve/iQuant

Jupyter Notebook 18,279 2,573 Updated Mar 11, 2026

搜索、推荐、广告、用增等工业界实践文章收集(来源:知乎、Datafuntalk、技术公众号)

HTML 4,540 484 Updated Aug 8, 2026

本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)

HTML 24,875 2,840 Updated Jul 19, 2026

how to optimize some algorithm in cuda.

Cuda 3,189 289 Updated Aug 7, 2026

Ongoing research training transformer language models at scale, including: BERT & GPT-2

Python 2,258 368 Updated Aug 14, 2025

The Triton Inference Server provides an optimized cloud and edge inferencing solution.

Python 10,910 1,823 Updated Aug 7, 2026

RPC framework based on C++ Workflow. Supports SRPC, Baidu bRPC, Tencent tRPC, thrift protocols.

C++ 2,137 408 Updated Mar 24, 2026

Modern CUDA Learn Notes with PyTorch for Beginners, 200+ CUDA Kernels, Tensor Cores, HGEMM, FA-2 MMA.

Cuda 11,732 1,234 Updated Aug 6, 2026

Transformer Explained Visually: Learn How LLM Transformer Models Work with Interactive Visualization

JavaScript 8,366 925 Updated Jun 6, 2026

Repository hosting code for "Actions Speak Louder than Words: Trillion-Parameter Sequential Transducers for Generative Recommendations" (https://arxiv.org/abs/2402.17152).

Python 1,958 407 Updated Aug 4, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,511 20,434 Updated Aug 8, 2026
Next