Skip to content
View io1261's full-sized avatar

Block or report io1261

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Diffusion model(SD,Flux,Wan,Qwen Image,Z-Image,...) inference in pure C/C++

C++ 6,746 733 Updated Aug 12, 2026
Python 5,888 357 Updated Aug 13, 2026

The easiest way to use Ollama in .NET

C# 1,397 187 Updated Jul 24, 2026

Scale computer-use 2.0 with open-source drivers, cross-OS fleets, and benchmarks for training, evaluation, and data generation.

HTML 21,335 1,469 Updated Aug 14, 2026

A unified AI model hub for aggregation & distribution. It supports cross-converting various LLMs into OpenAI-compatible, Claude-compatible, or Gemini-compatible formats. A centralized gateway for p…

Go 45,149 10,677 Updated Aug 14, 2026

Open-source GEO content engineering and multi-site distribution system with AI tasks, RAG/semantic chunking, analytics, GEOFlow Agent and WordPress target publishing.

PHP 3,231 745 Updated Aug 11, 2026

π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.

Rust 89,993 11,960 Updated Aug 14, 2026

🎙️ 「大模型」从0训练0.1B能听能说能看的全模态Omni模型!A 0.1B Omni model trained from scratch, capable of listening, speaking, and seeing!

Python 2,321 271 Updated Aug 6, 2026

Browser Harness | Self-healing harness that enables LLMs to complete any task.

Python 16,684 1,581 Updated Aug 3, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 14,176 1,628 Updated Aug 13, 2026

[ECCV 2026] 🔥 Official impl. of "DreamLite: A Lightweight On-Device Unified Model for Image Generation and Editing".

Python 746 50 Updated Aug 8, 2026

An open-source AI Voice Agent that integrates with Asterisk/FreePBX using Audiosocket/RTP technology

Python 1,176 257 Updated Aug 11, 2026

C++ ggml runtime hub for multilingual ASR and TTS models: Cohere Transcribe, Parakeet TDT, Voxtral, Canary 1B v2, etc, plus universal forced alignment, and more

C++ 541 90 Updated Aug 13, 2026

Open-source LLM-based ASR model family for Chinese, dialect, accent, and multilingual speech, with FunASR, vLLM, streaming, and llama.cpp runtimes.

C 1,474 146 Updated Jul 24, 2026

The headless browser for AI agents and web scraping

Rust 21,374 1,536 Updated Aug 14, 2026

🚀 World's largest GPT Image 2 prompt library, updated daily — 2000+ curated prompts with preview images, 16 languages. OpenAI's next-gen image model with pixel-perfect text rendering, cross-image c…

TypeScript 9,304 848 Updated Aug 14, 2026
Python 72 4 Updated May 4, 2026

Port of Funasr's Sense-voice model in C/C++

C 571 77 Updated Dec 19, 2025

Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key

Python 11,725 1,078 Updated Mar 22, 2026

Self-evolving agent: grows skill tree from 3.3K-line seed, achieving full system control with 6x less token consumption

Python 13,772 1,601 Updated Aug 13, 2026

MOSS-TTS-Nano is an open-source multilingual tiny speech generation model from MOSI.AI and the OpenMOSS team. With only 0.1B parameters, it is designed for realtime speech generation, can run direc…

Python 4,155 529 Updated Jul 26, 2026

MOSS‑TTS Family is an open‑source speech and sound generation model family from MOSI.AI and the OpenMOSS team. It is designed for high‑fidelity, high‑expressiveness, and complex real‑world scenario…

Python 3,989 353 Updated Jul 26, 2026

High-Quality Voice Cloning TTS for 600+ Languages

Python 9,102 1,500 Updated Aug 10, 2026

Qwen3-ASR is an open-source series of ASR models developed by the Qwen team at Alibaba Cloud, supporting stable multilingual speech/music/song recognition, language detection and timestamp prediction.

Python 3,368 335 Updated Jun 26, 2026

LLM inference in C/C++

C++ 2,269 388 Updated Aug 13, 2026

CLI proxy that reduces LLM token consumption by 60-90% on common dev commands. Single Rust binary, zero dependencies

Rust 76,099 4,788 Updated Aug 13, 2026

PersonaPlex code.

Python 10,352 1,444 Updated Mar 2, 2026

Simple DirectMedia Layer

C 16,328 2,915 Updated Aug 11, 2026

Omni inference in C/C++

C++ 243 72 Updated Aug 11, 2026
Next