Skip to content
View natowi's full-sized avatar
:octocat:
I may be slow to respond.
:octocat:
I may be slow to respond.

Organizations

@alicevision

Block or report natowi

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

Build local voice agents with open-source models

Python 12,013 1,474 Updated Aug 10, 2026

This is a ComfyUI custom node implementation of 'PersonaLive: Expressive Portrait Image Animation for Live Streaming'.

Python 121 16 Updated Jan 25, 2026

The repository provides code for running inference with the Meta Segment Anything Audio Model (SAM-Audio), links for downloading the trained model checkpoints, and example notebooks that show how t…

Python 3,592 325 Updated May 26, 2026

[CVPR2026]We present FlashPortrait, an end-to-end video diffusion transformer capable of synthesizing ID-preserving, infinite-length videos while achieving up to 6$\times$ acceleration in inference…

Python 482 38 Updated Feb 21, 2026
Python 11,856 814 Updated Feb 9, 2026

Open-Source Frontier Voice AI

Python 52,330 5,899 Updated Jul 24, 2026

SoTA open-source TTS

Python 25,950 3,473 Updated Jul 21, 2026

Codes for automatic point-cloud-to-BIM conversion

Python 125 32 Updated Aug 10, 2026

Cross-platform E57 file viewer to list and view stored point clouds, images and metadata.

C++ 22 2 Updated Jul 11, 2026

Xst Reader is an open source viewer for Microsoft Outlook’s .ost and .pst files, written entirely in C#. To download an executable of the current version, go to the releases tab.

C# 680 134 Updated Sep 11, 2023

ComfyUI wrapper for sam-3d-body

Python 327 32 Updated Jul 31, 2026

[AAAI'24] NeuSurf: On-Surface Priors for Neural Surface Reconstruction from Sparse Input Views

Python 87 2 Updated Jan 6, 2025

TTS model capable of streaming conversational audio in realtime.

Python 1,162 99 Updated Nov 29, 2025

HyMPS will be a platform-indipendent software suite for advanced audio/video contents production.

325 20 Updated Jul 14, 2026

🐬DeepChat - A smart assistant that connects powerful AI to your personal world

TypeScript 6,209 712 Updated Aug 10, 2026

Speech-to-text, text-to-speech, speaker diarization, speech enhancement, source separation, and VAD using next-gen Kaldi with onnxruntime without Internet connection. Support embedded systems, Andr…

C++ 14,085 1,621 Updated Aug 10, 2026

A free, open source, and extensible speech-to-text application that works completely offline.

Rust 29,201 2,573 Updated Aug 10, 2026

Local Lens is a privacy-first, AI-powered photo organizer for your PC. Sort and group photos by faces, dates, and locations—all locally, with no cloud upload. Enjoy a modern, intuitive UI and keep …

Python 149 13 Updated Jul 24, 2026

The Privacy First PDF Toolkit

JavaScript 14,584 1,216 Updated Aug 9, 2026

Epson Printer Configuration tool and waste ink counter resetter

Python 609 117 Updated Dec 30, 2025

[CVPR 2025] MIDI: Multi-Instance Diffusion for Single Image to 3D Scene Generation

Python 937 72 Updated Jun 12, 2025

[AAAI 2025] GigaGS: Scaling up Planar-Based 3D Gaussians for Large Scene Surface Reconstruction

151 3 Updated Sep 11, 2024

3D Gaussian Flats: Hybrid 2D/3D Photometric Scene Reconstruction

Python 65 1 Updated Nov 26, 2025

🔄 [ECCV‘24] Pytorch implementation of 'Surface Reconstruction from 3D Gaussian Splatting via Local Structural Hints'

Python 124 5 Updated Jan 19, 2026

ComfyUI plugin for submitting workflows to Thinkbox Deadline for distributed rendering

Python 34 4 Updated Jun 25, 2026

BillionMail gives you open-source MailServer, NewsLetter, Email Marketing — fully self-hosted, dev-friendly, and free from monthly fees. Join the discord: https://discord.gg/asfXzBUhZr

Go 15,401 1,666 Updated Jun 11, 2026

[3DV 2026] ViSTA-SLAM: Visual SLAM with Symmetric Two-view Association

Python 274 14 Updated Apr 5, 2026

VibeVoice: Expressive, longform conversational speech synthesis. (Community fork)

Python 1,420 630 Updated Aug 8, 2026

A collection of MCP servers.

92,049 14,277 Updated Aug 3, 2026
Next