Stars
CortermTerminal2 是一个面向 AI Coding Agent 的远程终端客户端,让你可以在手机上连接服务器,继续操作 Claude Code、Codex、Copilot 等终端任务。
Use Garry Tan's exact Claude Code setup: 23 opinionated tools that serve as CEO, Designer, Eng Manager, Release Manager, Doc Engineer, and QA
angelandy / sub2api
Forked from Wei-Shaw/sub2apiSub2API-CRS2 一站式开源中转服务,让 Claude、Openai 、Gemini、Antigravity订阅统一接入,支持拼车共享,更高效分摊成本,原生工具无缝使用。
🔥 Clone and recreate any website as a modern React app in seconds
The context API to search, scrape, and interact with the web at scale. 🔥
vue and ffmpeg based tool for video clips. 使用vue(vue3) + ffmpeg + wasm 实现纯前端音视频编辑,功能包括:视频剪辑、音频剪辑、音频合成裁剪、音波展示、视频抽帧、gif抽帧、帧播放器、字幕、贴图、时间轴、素材轨道
A YAML-based Playwright automation testing framework designed for Claude Code
MuseTalk: Real-Time High Quality Lip Synchorization with Latent Space Inpainting
Implement a chat robot interaction interface similar to chatGPT. Build using React and Flask, without investing too much effort in the development of chat components. 实现类似chatPGT的聊天机器人交互界面。使用React、…
[ICLR'24 spotlight] An open platform for training, serving, and evaluating large language model for tool learning.
基于ChatGLM-6B、ChatGLM2-6B、ChatGLM3-6B模型,进行下游具体任务微调,涉及Freeze、Lora、P-tuning、全参微调等
[CVPR 2022] Thin-Plate Spline Motion Model for Image Animation.
from livespeechportraits paper with tts
DiffSinger: Singing Voice Synthesis via Shallow Diffusion Mechanism (SVS & TTS); AAAI 2022; Official code
本项目基于SadTalkers实现视频唇形合成的Wav2lip。通过以视频文件方式进行语音驱动生成唇形,设置面部区域可配置的增强方式进行合成唇形(人脸)区域画面增强,提高生成唇形的清晰度。使用DAIN 插帧的DL算法对生成视频进行补帧,补充帧间合成唇形的动作过渡,使合成的唇形更为流畅、真实以及自然。
Open-source search database for full-text, vector, and hybrid search with real-time indexing and SQL.
基于标贝数据继续训练,同时对原本的FastSpeech2模型做了改进,引入了韵律表征以及韵律预测模块,使中文发音更生动且富有节奏
PaddleFormers is an easy-to-use library of pre-trained large language model zoo based on PaddlePaddle.
HuggingLLM, Hugging Future.
Easy-to-use and powerful LLM and SLM library with awesome model zoo.
深度学习入门课、资深课、特色课、学术案例、产业实践案例、深度学习知识百科及面试题库The course, case and knowledge of Deep Learning and AI
go-stash is a high performance, free and open source server-side data processing pipeline that ingests data from Kafka, processes it, and then sends it to ElasticSearch.
Recognize tables from images and restore them into word.
🎨 🎨 深度学习 卷积神经网络教程 :图像识别,目标检测,语义分割,实例分割,人脸识别,神经风格转换,GAN等🎨🎨 https://dataxujing.github.io/CNN-paper2/