Skip to content
View nik-ko's full-sized avatar

Organizations

@umr-ds

Block or report nik-ko

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

teddyCloud is an open source server replacement for the Boxine Cloud

C 975 72 Updated Jul 20, 2026

A feature-rich command-line audio/video downloader

Python 180,183 15,369 Updated Jul 23, 2026

Dockerfile for WhisperX: Automatic Speech Recognition with Word-Level Timestamps and Speaker Diarization (Dockerfile, CI image build and test)

Dockerfile 454 47 Updated Jul 19, 2026

Script for cropping scanned pages in pdf

Python 2 1 Updated Apr 21, 2021

OpenParliamentTV-Tools for parsing parliamentary data

Python 3 1 Updated Jul 14, 2026

A one stop repository for generative AI research updates, interview resources, notebooks and much more!

HTML 28,455 5,837 Updated Jul 22, 2026

Audio Annotation Tool for ML development

TypeScript 91 17 Updated Jul 8, 2026

High accurate text detection (OCR) Javascript/Typescript library that runs on Node.js, Browser, React Native and C++. Based on PaddleOCR and ONNX runtime

C++ 201 35 Updated Apr 22, 2026

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 86,272 11,085 Updated Jul 22, 2026

ocr-docker is small, Flask powerd web app, helps us to extract text from images and pdf document using OCR

CSS 76 20 Updated Mar 3, 2025

Fine-tune and evaluate Whisper models for Automatic Speech Recognition (ASR) on custom datasets or datasets from huggingface.

Python 365 93 Updated May 23, 2023

Label Studio is a multi-type data labeling and annotation tool with standardized output format

TypeScript 27,921 3,640 Updated Jul 24, 2026

WhisperX: Automatic Speech Recognition with Word-level Timestamps (& Diarization)

Python 23,268 2,357 Updated Jul 13, 2026

telegram bot for self-hosted local inference of stable diffusion, text-to-speech and large language models, such as llama3

Python 41 8 Updated May 13, 2024

Create a visual search engine using tensorflow serving, elasticsearch, vuejs and nginx.

Python 50 22 Updated Mar 21, 2019

Provider for Google Calendar

JavaScript 326 36 Updated Jun 24, 2026

DeepFaceLab is the leading software for creating deepfakes.

Python 19,302 920 Updated Nov 13, 2024

Visualization toolbox for Sound Event Detection

Python 122 29 Updated Feb 26, 2024

The official gpt4free repository | various collection of powerful language models | opus 4.6 gpt 5.3 kimi 2.5 deepseek v3.2 gemini 3

Python 66,492 13,528 Updated Jul 26, 2026

Semantic Image Similarity Search in Elasticsearch

Python 31 3 Updated Nov 4, 2024

Fusion of feature pyramids for nucleus segmentation and cell segmentation

Python 10 Updated Sep 15, 2023

How to Copy Text from Images ? Answer is TextSnatcher !. Perform OCR operations in seconds on Linux Desktop.

Vala 1,388 54 Updated Mar 20, 2024

Energy Logger 4000 utility

Python 21 11 Updated Oct 30, 2017

Virtual Video Device for Background Replacement with Deep Semantic Segmentation

C++ 746 85 Updated Jan 4, 2023

This repository provides a starter code for using tensorboard via tensorflow for visualising embeddings

Python 14 8 Updated Apr 4, 2018

1 Line of code data quality profiling & exploratory data analysis for Pandas and Spark DataFrames.

Python 13,654 1,796 Updated Apr 22, 2026