Skip to content
View ved1beta's full-sized avatar
🌴
On vacation
🌴
On vacation

Organizations

@axolotl-ai-cloud @dylo-oss @sanshins

Block or report ved1beta

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

DFlash: Block Diffusion for Flash Speculative Decoding

Python 6,118 433 Updated Aug 18, 2026

A lightweight inference engine supporting speculative speculative decoding (SSD).

Python 1,006 80 Updated May 10, 2026

ARC Relay — WebSocket relay server for agent remote control by Axolotl AI

Python 47 6 Updated Mar 30, 2026

Visual Causal Flow

Python 3,425 307 Updated Feb 3, 2026

High-performance Rust extensions for Axolotl (no OOM for large datasets) - drop-in acceleration for existing installations.

Python 3 Updated Jul 2, 2026

Nano vLLM

Python 15,607 2,638 Updated Apr 26, 2026
Python 1 Updated Oct 26, 2025

Disposable Temp mail service

Go 43 Updated Sep 16, 2026

[MLSys 2024 Best Paper Award] AWQ: Activation-aware Weight Quantization for LLM Compression and Acceleration

Python 3,639 326 Updated Jul 17, 2025

A safetensors extension to efficiently store sparse quantized tensors on disk

Python 1 Updated Jul 28, 2025

inference engine for LLMs

Python 2 1 Updated Nov 21, 2025

The Prime Intellect CLI provides a powerful command-line interface for managing GPU resources across various providers

Python 1 Updated Dec 3, 2025

Tensors and Dynamic neural networks in Python with strong GPU acceleration

Python 1 Updated Sep 10, 2026

🚀 Efficient implementations of state-of-the-art linear attention models in Torch and Triton

Python 1 Updated Jun 7, 2025

Finetune Qwen3, Llama 4, TTS, DeepSeek-R1 & Gemma 3 LLMs 2x faster with 70% less memory! 🦥

Python 1 Updated Dec 15, 2025

Utils for Unsloth

Python 1 Updated Aug 16, 2025

🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.

Python 1 Updated Dec 2, 2025

Efficient implementations of state-of-the-art sequence modeling architectures—using PyTorch and Triton.

Python 2 Updated Sep 22, 2025

Small scale distributed training of sequential deep learning models, built on Numpy and MPI.

Python 3 Updated Apr 25, 2026

FlashInfer: Kernel Library for LLM Serving

Cuda 6,502 1,494 Updated Sep 24, 2026

PyTorch native quantization and sparsity for training and inference

Python 1 Updated Feb 1, 2026
Python 1 1 Updated Jun 6, 2025

A community curated list of Rust Language streamers

751 40 Updated Jan 7, 2024

Extremely fast Query Engine for DataFrames, written in Rust

Rust 39,852 3,128 Updated Sep 24, 2026

Node.js dependency tracing utility

JavaScript 1,672 185 Updated Aug 18, 2026

[Hackintosh] Configuration for Lenovo Thinkpad P50.

ASL 49 12 Updated Sep 14, 2023

Autonomous coding agent as an SDK, IDE extension, or CLI assistant.

TypeScript 69,197 7,504 Updated Sep 24, 2026

HelixDB is an OLTP graph database with native vector and full-text search built in Rust on Object Storage.

Rust 6,087 370 Updated Sep 23, 2026

Rust full node implementation of the Fuel v2 protocol.

Rust 56,831 2,862 Updated Sep 22, 2026
Next