Lists (2)
Sort Name ascending (A-Z)
Stars
The best-benchmarked open-source AI memory system. And it's free.
The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
Reference implementations of MLPerf® inference benchmarks
The Triton TensorRT-LLM Backend
TensorRT LLM provides users with an easy-to-use Python API to define Large Language Models (LLMs) and supports state-of-the-art optimizations to perform inference efficiently on NVIDIA GPUs. Tensor…
A high-throughput and memory-efficient inference and serving engine for LLMs
AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.
Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"
Stable Diffusion with Core ML on Apple Silicon
A Python framework for GPU-accelerated simulation, robotics, and machine learning.
PyTriton is a Flask/FastAPI-like interface that simplifies Triton's deployment in Python environments.
Open source cross-platform compiler for compute-intensive loops used in AI algorithms, from Microsoft Research
Evals is a framework for evaluating LLMs and LLM systems, and an open-source registry of benchmarks.
Pretrain, finetune ANY AI model of ANY size on 1 or 10,000+ GPUs with zero code changes.
Running large language models on a single GPU for throughput-oriented scenarios.
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
A latent text-to-image diffusion model
Examples demonstrating available options to program multiple GPUs in a single node or a cluster
AITemplate is a Python framework which renders neural network into high performance CUDA/HIP C++ code. Specialized for FP16 TensorCore (NVIDIA GPU) and MatrixCore (AMD GPU) inference.
Common source, scripts and utilities for creating Triton backends.