Skip to content
View ishine's full-sized avatar
  • gerzz.inc
  • shanghai

Block or report ishine

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Language model tokenization at GB/s

Rust 3,925 204 Updated Aug 6, 2026

A general framework for distilling human-created multimodal resources into reusable, executable skills that AI agents can browse, compose, and run, validated across diverse domains including web, P…

Python 369 33 Updated Jul 17, 2026
Python 6 2 Updated Feb 20, 2026

AI 环绕声智能上混系统 — 基于 Demucs 深度学习模型,将普通立体声音频分离为人声/贝斯/鼓/乐器 4 音轨,智能路由到 5.1/7.1 多声道布局。支持 NVIDIA CUDA / AMD ROCm / Intel Arc 显卡加速,Gradio WebUI 一键操作,输出 WAV/FLAC/AAC 环绕声文件。

Python 4 Updated Jun 3, 2026

🛸 A modern, cross-platform C++ webview library

C++ 908 57 Updated Jul 28, 2026

A multi-instrument music transcription model developed by Kyutai and Mirelo.

Python 977 125 Updated Aug 5, 2026

Eureka-Audio: A 1.7B lightweight audio–language model that matches 7B–30B models on ASR, audio understanding, and paralinguistic reasoning.

Python 42 5 Updated Apr 11, 2026

Chrome extension & CLI to let agents control your browser. Runs Playwright snippets in a stateful sandbox. Available as CLI or MCP

HTML 3,747 168 Updated Jul 27, 2026

Official JAX implementation of End-to-End Test-Time Training for Long Context

Python 631 47 Updated Feb 15, 2026

MiroThinker is a deep research agent optimized for complex research and prediction tasks. Our latest models, MiroThinker-1.7, achieves 74.0 and 75.3 on the BrowseComp and BrowseComp Zh, respectively.

Python 8,367 645 Updated Jul 6, 2026

This repository contains tools to download, crawl, and process French political speeches from the vie-publique.fr public dataset. It allows for the collection of speech metadata and the scraping of…

Jupyter Notebook 3 1 Updated Jan 4, 2026

General plug-and-play inference library for Recursive Language Models (RLMs), supporting various sandboxes.

Python 5,362 873 Updated Jun 26, 2026

Qwen-UI-Agent: Towards Next-Generation Real-World Centric Foundation GUI Agent

Jupyter Notebook 1,899 182 Updated Aug 5, 2026

DeepTutor: Lifelong Personalized Tutoring. https://deeptutor.info/.

Python 32,710 4,271 Updated Aug 6, 2026

Playwright is a framework for Web Testing and Automation. It allows testing Chromium, Firefox and WebKit with a single API.

TypeScript 94,097 6,225 Updated Aug 6, 2026

A compact implementation of SGLang, designed to demystify the complexities of modern LLM serving systems.

Python 4,696 772 Updated May 17, 2026

An open-source, code-first Python toolkit for building, evaluating, and deploying sophisticated AI agents with flexibility and control.

Python 21,020 3,807 Updated Aug 6, 2026

A collection of sample agents built with Agent Development Kit (ADK)

Python 10,057 2,786 Updated Aug 5, 2026
Python 49 4 Updated Mar 29, 2026

A highly optimized engine for neutts-air model to generate minutes of audio in seconds. Over 200x realtime on modern hardware!

Python 119 12 Updated Nov 24, 2025

A family of efficient speech models for multilingual phone recognition

Python 70 12 Updated Jul 18, 2026

Official code for "F5R-TTS: Improving Flow-Matching based Text-to-Speech with Group Relative Policy Optimization"

Python 170 18 Updated Mar 3, 2026

Open Audio Watermarking Tool

Python 531 48 Updated Dec 22, 2025
Python 7 3 Updated Oct 24, 2025

Generate audio signals corresponding to moving sources/receivers in a shoebox-shaped room (Python)

Python 11 1 Updated Nov 14, 2025

[ASRU 2025] Omni-R1: Do You Really Need Audio to Fine-Tune Your Audio LLM?

Python 47 1 Updated Nov 21, 2025
Python 11,838 814 Updated Feb 9, 2026

A new dataset that includes long audio, captions of local audio events, and temporal boundaries

14 Updated Mar 26, 2026

SpikeMamba presents a novel integration of spiking neural networks (SNNs) with the Mamba state space model architecture, investigating the potential for biologically-inspired temporal dynamics in l…

Python 6 1 Updated Sep 9, 2025

Resources to develop programming and software development skills

HTML 27 10 Updated Sep 21, 2023
Next