Lists (2)
Sort Name ascending (A-Z)
Starred repositories
Robust Speech Recognition via Large-Scale Weak Supervision
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
🚀Clone a voice in 5 seconds to generate arbitrary speech in real-time
Blind&Invisible Watermark ,图片盲水印,提取水印无须原图!
Use Microsoft Edge's online text-to-speech service from Python WITHOUT needing Microsoft Edge or Windows or an API key
💮 amazing QRCode generator (supporting animated gif) - amazing 二维码生成器(支持 gif 动态图片二维码)
Free and open-source map hosting solution with custom styles for websites and apps, using OpenStreetMap data
Code for robust monocular depth estimation described in "Ranftl et. al., Towards Robust Monocular Depth Estimation: Mixing Datasets for Zero-shot Cross-dataset Transfer, TPAMI 2022"
各种脚本 -- 关于 虾米 xiami.com, 百度网盘 pan.baidu.com, 115网盘 115.com, 网易音乐 music.163.com, 百度音乐 music.baidu.com, 360网盘/云盘 yunpan.cn, 视频解析 flvxz.com, bt torrent ↔ magnet, ed2k 搜索, tumblr 图片下载, unzip
A Tensorflow implementation of AnimeGAN for fast photo animation ! This is the Open source of the paper 「AnimeGAN: a novel lightweight GAN for photo animation」, which uses the GAN framwork to trans…
A Blender add-on to import models from google maps
mmd_tools is a blender addon for importing Models and Motions of MikuMikuDance.
VRM Importer, Exporter and Utilities for Blender 2.93 to 5.2
Translate images to unseen domains in the test time with few example images.
This repo contains the projects: 'Virtual Normal', 'DiverseDepth', and '3D Scene Shape'. They aim to solve the monocular depth estimation, 3D scene reconstruction from single image problems.
获取B站直播推流码,支持开关播,管理直播标题、分区,显示弹幕和礼物。
A simple deep learning library for estimating a set of tags and extracting semantic feature vectors from given illustrations.