Stars
Fully open reproduction of DeepSeek-R1
主要记录大语言大模型(LLMs) 算法(应用)工程师相关的知识及面试题
SGLang is a high-performance serving framework for large language models and multimodal models.
Implement a ChatGPT-like LLM in PyTorch from scratch, step by step
llama3 implementation one matrix multiplication at a time
DeepSeek-V2: A Strong, Economical, and Efficient Mixture-of-Experts Language Model
Emscripten: An LLVM-to-WebAssembly Compiler
Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more
Official implementation of AnimateDiff.
A lite C++ AI toolkit: 100+ models with MNN, ORT and TRT, including Det, Seg, Stable-Diffusion, Face-Fusion.
A Cloud Native Batch System (Project under CNCF)
🐜🐜🐜 ants is the most powerful and reliable pooling solution for Go.
Development repository for the Triton language and compiler
Get up and running with Kimi, GLM, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Common source, scripts and utilities for creating Triton backends.
Swap GPT for any LLM by changing a single line of code. Xinference lets you run open-source, speech, and multimodal models on cloud, on-prem, or your laptop — all through one unified, production-re…
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…
Latent Consistency Models: Synthesizing High-Resolution Images with Few-Step Inference
A curated list of practical guide resources of LLMs (LLMs Tree, Examples, Papers)
trholding / llama2.c
Forked from karpathy/llama2.cLlama 2 Everywhere (L2E)