Large Language Model Text Generation Inference
-
Updated
Mar 21, 2026 - Python
Large Language Model Text Generation Inference
BELLE: Be Everyone's Large Language model Engine(开源中文对话大模型)
🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading
Go package implementing Bloom filters, used by many important systems
[NeurIPS 2023] LLM-Pruner: On the Structural Pruning of Large Language Models. Support Llama-3/3.1, Llama-2, LLaMA, BLOOM, Vicuna, Baichuan, TinyLlama, etc.
Fast Inference Solutions for BLOOM
🩹Editing large language models within 10 seconds⚡
💬 Chatbot web app + HTTP and Websocket endpoints for LLM inference with the Petals client
Spicetify theme inspired by Microsoft's Fluent Design, Always up-to-date!, A Powerful Theme to Calm your Eyes While Listening to Your Favorite Beats
Two sample projects built for the FourthBrain Generative AI Workshop
Effects created for ReShade
OpenGL C++ Graphics Engine
Crosslingual Generalization through Multitask Finetuning
DirectX 11 Renderer written in C++11
To associate your repository with the bloom topic, visit your repo's landing page and select "manage topics."