Skip to content
View yxanul's full-sized avatar

Block or report yxanul

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

JAX backend for SGL

Python 333 121 Updated Aug 6, 2026

Open Frontier Intelligence

8,242 625 Updated Aug 6, 2026

TPU inference for vLLM, with unified JAX and PyTorch support.

Python 403 280 Updated Aug 8, 2026

A simple, performant, and scalable Jax LLM!

Python 2,382 582 Updated Aug 8, 2026

Fast Polar Decomposition for Muon

Python 171 13 Updated Jul 2, 2026

A batteries-included framework for building web apps

Rust 4,339 160 Updated Aug 8, 2026

A Quirky Assortment of CuTe Kernels

Python 1,095 150 Updated Aug 8, 2026

Tokamax: A GPU and TPU kernel library.

Python 259 46 Updated Aug 8, 2026

[Tech Report] Expanded Hyper-Connections

60 1 Updated Jul 21, 2026

Instant, Concurrent, Secure & Lightweight Sandbox for AI Agents.

Rust 11,000 1,018 Updated Aug 7, 2026

🌟 [My TouchBar My rules]. The Touch Bar Customisation App for your MacBook Pro

Swift 4,319 231 Updated May 14, 2026

Zmeu — a systems programming language with its own syntax, compile-time ownership & borrow checking, structured concurrency, and an MLIR-to-LLVM native backend.

Rust 1 Updated Jun 29, 2026

DeepSpec: a full-stack codebase for training and evaluating speculative decoding algorithms

Python 6,906 645 Updated Jul 9, 2026

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,539 7,747 Updated Aug 8, 2026

LMCache: Supercharge Your LLM with the Fastest KV Cache Layer

Python 11,069 1,678 Updated Aug 8, 2026

Automated alignment adjustment for LLMs — direct steering, LoRA, and MoE expert-granular abliteration, optimized via multi-objective Optuna TPE.

Python 190 38 Updated Aug 6, 2026

A PyTorch native library for training speculative decoding models

Python 223 59 Updated Aug 8, 2026

Fully automatic censorship removal for language models

Python 27,208 2,946 Updated Aug 7, 2026

Blazingly fast pusher drop-in replacement written in rust

Rust 789 59 Updated Aug 8, 2026

Compare command performance using Linux hardware counters (perf_event_open) — Linux-only CLI

Rust 2 1 Updated Jun 7, 2026

Full Transformer into a custom chip. microGPT in RTL, generating names on a Virtex-5 FPGA at ~56k tokens/second.

Verilog 627 103 Updated Jun 25, 2026

An implementation of a LRU cache

Rust 830 129 Updated Aug 3, 2026

A high-performance ASCII video rendering engine featuring real-time WebSocket binary streaming and an isolated compiler for serverless static generation. Built for low-latency 30 FPS playback on HT…

Python 2,590 298 Updated Aug 8, 2026

Automatically generates Rust FFI bindings to C (and some C++) libraries.

Rust 5,260 822 Updated Jul 28, 2026

Puffing up reinforcement learning

C 6,238 528 Updated Aug 4, 2026

An extremely fast Python type checker and language server, written in Rust.

Python 19,421 323 Updated Aug 7, 2026

The LLVM Project is a collection of modular and reusable compiler and toolchain technologies.

LLVM 39,691 18,150 Updated Aug 8, 2026

Simple, fast, safe, compiled language for developing maintainable software. Compiles itself in <1s with zero library dependencies. Supports automatic C => V translation. https://vlang.io

V 37,786 2,268 Updated Aug 8, 2026

Unofficial ChatGPT desktop app for Linux (formerly the Codex app), built locally from OpenAI’s official macOS app. Includes Chat, Work, and Codex. Packages for Debian/Ubuntu (.deb), Fedora/openSUSE…

JavaScript 3,504 441 Updated Aug 8, 2026

Fast stackful fibers with a NUMA-aware work-stealing scheduler

C++ 318 12 Updated Aug 6, 2026
Next