Stars
Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.
Multilingual Document Layout Parsing in a Single Vision-Language Model
The list of NLP paper and news I've checked. There might be short description of them (abstract) in Korean.
Generate text line images for training deep learning OCR models
A collection of localized (Korean) AWS AI/ML workshop materials for hands-on labs.
A simple screen parsing tool towards pure vision based GUI agent
DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception
Official inference framework for 1-bit LLMs
OCR, layout analysis, reading order, table recognition in 90+ languages
📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.
Optical Character Recognition (OCR) is a powerful technology that enables machines to recognize and extract text from images or scanned documents. OCR finds applications in various fields, includin…
Text recognition (optical character recognition) with deep learning methods, ICCV 2019
Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.
Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.
언어모델을 학습하기 위한 공개 한국어 instruction dataset들을 모아두었습니다.
한국어 자연어처리를 위한 파이썬 라이브러리입니다. 단어 추출/ 토크나이저 / 품사판별/ 전처리의 기능을 제공합니다.
Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.
LangChain 공식 Document, Cookbook, 그 밖의 실용 예제를 바탕으로 작성한 한국어 튜토리얼입니다. 본 튜토리얼을 통해 LangChain을 더 쉽고 효과적으로 사용하는 방법을 배울 수 있습니다.
Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.
A high-throughput and memory-efficient inference and serving engine for LLMs
A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.
📚 Tech blogs & talks by companies that run Kafka in production
Chat with your database or your datalake (SQL, CSV, parquet). PandasAI makes data analysis conversational using LLMs and RAG.
StableLM: Stability AI Language Models
Running large language models on a single GPU for throughput-oriented scenarios.
Pecab: Pure python Korean morpheme analyzer based on Mecab
Source code for end-to-end dialogue model from the MultiWOZ paper (Budzianowski et al. 2018, EMNLP)