Skip to content
View weijietong's full-sized avatar

Organizations

@RoaringBitmap

Block or report weijietong

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

OnPair: Short Strings Compression for Fast Random Access

Rust 7 1 Updated Jan 31, 2026

A fast framework for writing baseline compiler back-ends in C++

LLVM 670 37 Updated Aug 5, 2026

OpenResty's Branch of LuaJIT 2

C 1,449 252 Updated Sep 17, 2026

Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.

Python 31,356 3,797 Updated Sep 17, 2026

Composable transformations of Python+NumPy programs: differentiate, vectorize, JIT to GPU/TPU, and more

Python 36,318 3,793 Updated Sep 20, 2026

A machine learning compiler for GPUs, CPUs, and ML accelerators

C++ 4,543 940 Updated Sep 21, 2026

Implement a reasoning LLM in PyTorch from scratch, step by step

Jupyter Notebook 5,259 825 Updated Sep 17, 2026

Build smaller, faster, and more secure desktop and mobile applications with a web frontend.

Rust 111,219 4,010 Updated Sep 21, 2026

AnyBlox runtime and tooling

C 42 1 Updated Sep 4, 2025

An extensible, state-of-the-art framework for columnar compression, and the fastest FOSS columnar file format. Formerly at @spiraldb, now an Incubation Stage project at LFAI&Data, part of the Linux…

Rust 3,207 225 Updated Sep 21, 2026

An easy-to-use, header-only C++ wrapper for Linux' perf event API

C++ 145 23 Updated May 18, 2026
C++ 30 4 Updated Nov 7, 2025

GPU-native composable analytics engine

C++ 1,074 128 Updated Sep 18, 2026

Goal: Enable awesome tooling for Bazel users of the C language family.

Python 917 194 Updated Aug 11, 2025

The universal proxy platform

Go 38,179 4,624 Updated Sep 20, 2026

[TMLR 2025] Efficient Reasoning Models: A Survey

Python 320 23 Updated Jun 26, 2026

Vector (and Scalar) Quantization, in Pytorch

Python 4,006 339 Updated Sep 2, 2026

Official repository of the xLSTM.

Python 2,204 186 Updated Sep 7, 2026

A library for accelerating Transformer models on NVIDIA GPUs, including using 8-bit and 4-bit floating point (FP8 and FP4) precision on Hopper, Ada and Blackwell GPUs, to provide better performance…

Python 3,546 831 Updated Sep 19, 2026

The hub for EleutherAI's work on interpretability and learning dynamics

Jupyter Notebook 2,941 224 Updated Nov 15, 2025

Modeling, training, eval, and inference code for OLMo

Python 6,682 800 Updated Nov 24, 2025

InkFuse - An Experimental Database Runtime Unifying Vectorized and Compiled Query Execution.

C++ 58 4 Updated May 13, 2024

CUDA Templates and Python DSLs for High-Performance Linear Algebra

C++ 10,462 2,094 Updated Sep 17, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,680 2,765 Updated Sep 21, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,938 9,179 Updated Sep 14, 2026

Local UI to run and train LLMs and diffusion models. Supports GGUF, MLX, Qwen3.8, DeepSeek-V4, MiniMax-H3, Gemma 4, FLUX and more.

Python 76,505 6,992 Updated Sep 21, 2026

The AI developer platform. Use Weights & Biases to train and fine-tune models, and manage models from experimentation to production.

Python 11,255 899 Updated Sep 20, 2026

A Flexible Framework for Experiencing Heterogeneous LLM Inference/Fine-tune Optimizations

Python 19,527 1,577 Updated Sep 19, 2026

FP16xINT4 LLM inference kernel that can achieve near-ideal ~4x speedups up to medium batchsizes of 16-32 tokens.

Python 1,149 92 Updated Sep 4, 2024

A concise but complete full-attention transformer with a set of promising experimental features from various papers

Python 5,948 519 Updated Sep 21, 2026
Next