Skip to content
View ggerganov's full-sized avatar
🌴
AFK until 27th July
🌴
AFK until 27th July

Sponsors

Organizations

@ggml-org

Block or report ggerganov

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

ggml speech-to-text inference for 16+ model families

C++ 1,571 57 Updated Jul 25, 2026

Standalone C++/GGML runtime for ThinkSound text->sound-effect generation

C++ 35 1 Updated Jul 13, 2026

Port of Nvidia LocateAnything-3B on ggml

C++ 375 43 Updated Jul 22, 2026

Vim-fork focused on extensibility and usability

Vim Script 101,342 6,989 Updated Jul 25, 2026

github action to speedup building using ccache

TypeScript 181 73 Updated Jul 20, 2026

Visualizer for neural network, deep learning and machine learning models

JavaScript 33,261 3,168 Updated Jul 25, 2026

Fast state-of-the-art image and video segmentation in portable C/C++

C++ 340 37 Updated Apr 10, 2026

Mount Hugging Face Buckets and repos as local filesystems. No download, no copy, no waiting.

Rust 769 61 Updated Jul 12, 2026

Portable C++17 implementation of ACE-Step 1.5 AI Music Generator using GGML. Text + lyrics in, stereo 48kHz MP3 or WAV out. Runs on CPU, CUDA, ROCm, Metal, Vulkan.

C++ 374 75 Updated Jul 21, 2026

A C++17 single-file header-only wrapper for llama.cpp

C++ 30 Updated Jul 20, 2026

Run AI models locally on your machine with node.js bindings for llama.cpp. Enforce a JSON schema on the model output on the generation level

TypeScript 2,144 208 Updated Jul 20, 2026

A free, open source, and extensible speech-to-text application that works completely offline.

Rust 27,484 2,381 Updated Jul 25, 2026

A cosy home for your LLMs.

Swift 1,405 90 Updated Jul 25, 2026

Audio playback and capture library written in C, in a single source file.

C 7,056 582 Updated Jul 20, 2026

Local LLM-assisted text completion for Qt Creator.

C++ 65 9 Updated Jun 21, 2026

Simple GUI around whisper.cpp for voice-to-text on Linux

Python 77 15 Updated Apr 22, 2026

Local LLM-assisted text completion for Qt Creator.

C++ 70 16 Updated Jun 21, 2026

MLPerf Client is a benchmark for Windows, Linux and macOS, focusing on client form factors in ML inference scenarios.

C++ 86 8 Updated Jul 14, 2026

Emacs package for LLM-assisted code/text completion

Emacs Lisp 43 2 Updated May 22, 2026

Lemonade helps users discover and run local AI apps by serving optimized LLMs right from their own GPUs and NPUs. Join our discord: https://discord.gg/5xXzkMu8Zk

C++ 5,116 420 Updated Jul 25, 2026

The application performs real-time inference on audio from an ALSA capture device

C++ 39 1 Updated Jun 19, 2025

TTS support with GGML

C++ 244 33 Updated Oct 5, 2025
Python 539 59 Updated Jun 11, 2026

LLM plugin for interacting with llama-server models

Python 31 6 Updated May 28, 2025

Running any GGUF SLMs/LLMs locally, on-device in Android

Kotlin 871 144 Updated Jun 21, 2026

DINOv2 inference engine written in C/C++ using ggml and OpenCV.

C++ 99 7 Updated May 6, 2025

Qwen3 is the large language model series developed by Qwen team, Alibaba Cloud.

Python 27,425 2,026 Updated Jan 9, 2026

Real-time webcam demo with SmolVLM and llama.cpp server

HTML 5,563 895 Updated May 12, 2025
Next