Skip to content
View lynchying's full-sized avatar

Block or report lynchying

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.

Python 5,488 618 Updated Aug 15, 2026

该仓库尝试整理推荐系统领域的一些经典算法模型

Jupyter Notebook 2,187 416 Updated Oct 15, 2023

整理开源的中文大语言模型,以规模较小、可私有化部署、训练成本较低的模型为主,包括底座模型,垂直领域微调及应用,数据集与教程等。

22,737 2,134 Updated May 10, 2026

🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.

Python 164,117 34,249 Updated Aug 15, 2026

🤗 The largest hub of ready-to-use datasets for AI models with fast, easy-to-use and efficient data manipulation tools

Python 21,840 3,350 Updated Aug 12, 2026

Open-source super AI assistant & Agent Harness. Plans tasks, runs tools and skills, self-evolves with memory and knowledge. Multi-model, multi-channel. Lightweight, extensible, one-line install. (f…

Python 46,514 10,315 Updated Aug 15, 2026

The official Meta Llama 3 GitHub site

Python 29,257 3,526 Updated Jan 26, 2025

Ingest, parse, and optimize any data format ➡️ from documents to multimedia ➡️ for enhanced compatibility with GenAI frameworks

Python 7,806 663 Updated Dec 12, 2025

[CVPR 2023] SadTalker:Learning Realistic 3D Motion Coefficients for Stylized Audio-Driven Single Image Talking Face Animation

Python 14,014 2,682 Updated Jun 26, 2024

RoBERTa中文预训练模型: RoBERTa for Chinese

Python 2,791 408 Updated Jul 22, 2024

A PyTorch-based knowledge distillation toolkit for natural language processing

Python 1,708 243 Updated May 8, 2023

Revisiting Pre-trained Models for Chinese Natural Language Processing (MacBERT)

718 61 Updated Apr 19, 2026

中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)

Python 18,935 1,839 Updated Apr 19, 2026

Collections of resources from Joint Laboratory of HIT and iFLYTEK Research (HFL)

Markdown 377 40 Updated Mar 9, 2023

Unsupervised text tokenizer for Neural Network-based text generation.

C++ 12,024 1,376 Updated Aug 15, 2026

Pre-Trained Chinese XLNet(中文XLNet预训练模型)

Python 1,646 278 Updated Apr 19, 2026

🆙 Upscayl - #1 Free and Open Source AI Image Upscaler for Linux, MacOS and Windows.

TypeScript 48,295 2,423 Updated Aug 4, 2026

text2vec, text to vector. 文本向量表征工具,把文本转化为向量矩阵,实现了Word2Vec、RankBM25、Sentence-BERT、CoSENT等文本表征、文本相似度计算模型,开箱即用。

Python 4,976 428 Updated Feb 14, 2026

State-of-the-Art Embeddings, Retrieval, and Reranking

Python 19,010 2,855 Updated Aug 14, 2026

总结梳理自然语言处理工程师(NLP)需要积累的各方面知识,包括面试题,各种基础知识,工程能力等等,提升核心竞争力

Python 7,509 1,197 Updated Aug 24, 2022

Topic Modelling for Humans

Python 16,479 4,408 Updated Nov 1, 2025

CDeC-Net: Composite Deformable Cascade Network for Table Detection in Document Images

Python 134 33 Updated Sep 11, 2025

This repository contains the code and implementation details of the CascadeTabNet paper "CascadeTabNet: An approach for end to end table detection and structure recognition from image-based documents"

Python 1,549 425 Updated Aug 27, 2021

A collection of original, innovative ideas and algorithms towards Advanced Literate Machinery. This project is maintained by the OCR Team in the Language Technology Lab, Tongyi Lab, Alibaba Group.

C++ 1,834 197 Updated Mar 17, 2026

CnOCR: Awesome Chinese/English OCR Python toolkits based on PyTorch. It comes with 20+ well-trained models for different application scenarios and can be used directly after installation. 【基于 PyTor…

Python 3,766 535 Updated Jul 5, 2026

PaddleFormers is an easy-to-use library of pre-trained large language model zoo based on PaddlePaddle.

Python 12,989 2,194 Updated Aug 15, 2026

2nd solution of ICDAR 2021 Competition on Scientific Literature Parsing, Task B.

Python 470 106 Updated Jul 4, 2022

Re-implementation of MASTER by mmocr

Python 90 18 Updated Sep 9, 2021

Different python scripts used in the OCR4all workflow.

Python 4 2 Updated May 23, 2023
Next