Skip to content
View woruyu's full-sized avatar

Organizations

@llvm

Block or report woruyu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Github mirror of trition-lang/triton repo.

MLIR 187 67 Updated Aug 14, 2026

Exercises for Learning MLIR (Originally written for PPoPP 2026)

C++ 108 8 Updated Jul 21, 2026

A Python framework for GPU-accelerated simulation, robotics, and machine learning.

Python 6,995 589 Updated Aug 13, 2026

Virtual whiteboard for sketching hand-drawn like diagrams

TypeScript 129,556 14,859 Updated Aug 13, 2026

Development repository for the Triton language and compiler

MLIR 19,941 3,108 Updated Aug 14, 2026

A compiler for homomorphic encryption

MLIR 764 150 Updated Aug 14, 2026

The Torch-MLIR project aims to provide first class support from the PyTorch ecosystem to the MLIR ecosystem.

C++ 1,885 720 Updated Aug 13, 2026

cuTile is a programming model for writing parallel kernels for NVIDIA GPUs

Python 2,126 143 Updated Aug 12, 2026

This repository contains tutorials and examples for Triton Inference Server

Python 858 155 Updated Aug 7, 2026

The Modular Platform (includes MAX & Mojo)

Mojo 26,788 2,913 Updated Aug 13, 2026

A retargetable MLIR-based machine learning compiler and runtime toolkit.

C++ 3,889 980 Updated Aug 13, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,006 20,667 Updated Aug 14, 2026

TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…

Python 14,379 2,660 Updated Aug 14, 2026

MII makes low-latency and high-throughput inference possible, powered by DeepSpeed.

Python 2,109 191 Updated Jun 30, 2025

A scalable generative AI framework built for researchers and developers working on Large Language Models, Multimodal, and Speech AI (Automatic Speech Recognition and Text-to-Speech)

Python 18,122 3,554 Updated Aug 13, 2026

Ongoing research training transformer models at scale

Python 17,422 4,358 Updated Aug 14, 2026

MLIR For Beginners tutorial

C++ 1,342 142 Updated Jul 18, 2025
Python 39 27 Updated Aug 14, 2026

[Deprecated] ⭐️ TT-NN Compiler for PyTorch 2 ⭐️ Enables running PyTorch models on Tenstorrent hardware using eager or compile path

Python 62 28 Updated Feb 24, 2026

Learn project of generate Dwarf debug symbol with llvm library.

C++ 1 Updated Oct 17, 2025
C++ 1 Updated Dec 18, 2022

Enabling PyTorch on XLA Devices (e.g. Google TPU)

C++ 2,799 571 Updated May 27, 2026

A machine learning compiler for GPUs, CPUs, and ML accelerators

C++ 4,468 894 Updated Aug 14, 2026

Repo for AI Compiler team. The intended purpose of this repo is for implementation of a PJRT device.

Python 74 32 Updated Aug 14, 2026

a lightweight and blazingly fast WebAssembly compiler and runtime library designed for use on resource-constrained embedded systems while seamlessly scaling to high-powered desktop and server systems

C++ 23 7 Updated Aug 12, 2026

Backward compatible ML compute opset inspired by HLO/MHLO

MLIR 683 214 Updated Aug 13, 2026

Temporary downstream RISC-V LLVM tree. You almost certainly want upstream LLVM instead (see https://llvm.org/docs/GettingStarted.html)

C++ 4 4 Updated Jul 16, 2023

LLVM Code Generation, published by Packt

C++ 280 67 Updated May 14, 2026

This is the second repo for the book "LLVM Code Generation". This will be linked to the main repo for this title.

LLVM 46 30 Updated Aug 9, 2026
Next