Skip to content
View uyeongjae's full-sized avatar

Organizations

@boostcampaitech2

Block or report uyeongjae

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Unlimited OCR Works: Welcome the Era of One-shot Long-horizon Parsing.

Python 23,688 2,430 Updated Jul 29, 2026

Multilingual Document Layout Parsing in a Single Vision-Language Model

Python 9,067 802 Updated Mar 24, 2026

The list of NLP paper and news I've checked. There might be short description of them (abstract) in Korean.

Astro 38 1 Updated Aug 12, 2026

Generate text line images for training deep learning OCR models

Python 914 175 Updated May 17, 2026

A collection of localized (Korean) AWS AI/ML workshop materials for hands-on labs.

Jupyter Notebook 307 167 Updated Aug 4, 2026

A simple screen parsing tool towards pure vision based GUI agent

Jupyter Notebook 25,253 2,226 Updated Jul 20, 2026

DocLayout-YOLO: Enhancing Document Layout Analysis through Diverse Synthetic Data and Global-to-Local Adaptive Perception

Python 2,248 170 Updated Apr 14, 2025

Official inference framework for 1-bit LLMs

C++ 40,077 3,695 Updated Jul 27, 2026

OCR, layout analysis, reading order, table recognition in 90+ languages

Python 21,270 1,526 Updated Jul 23, 2026

📄 Awesome OCR multiple programing languages toolkits based on ONNX Runtime, OpenVINO, MNN, PaddlePaddle, TensorRT and PyTorch.

Python 7,485 698 Updated Aug 14, 2026

Optical Character Recognition (OCR) is a powerful technology that enables machines to recognize and extract text from images or scanned documents. OCR finds applications in various fields, includin…

Python 20 6 Updated Jul 30, 2023

Text recognition (optical character recognition) with deep learning methods, ICCV 2019

Jupyter Notebook 3,942 1,134 Updated Mar 4, 2024

Ready-to-use OCR with 80+ supported languages and all popular writing scripts including Latin, Chinese, Arabic, Devanagari, Cyrillic and etc.

Python 29,907 3,597 Updated Dec 5, 2025

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,592 11,176 Updated Jul 22, 2026

언어모델을 학습하기 위한 공개 한국어 instruction dataset들을 모아두었습니다.

Python 472 30 Updated Apr 13, 2025

한국어 자연어처리를 위한 파이썬 라이브러리입니다. 단어 추출/ 토크나이저 / 품사판별/ 전처리의 기능을 제공합니다.

Python 986 184 Updated Mar 10, 2026

Local UI to run and train LLMs and diffusion models, including Qwen3.8, Kimi K3, MiniMax-H3, Gemma 4, DeepSeek-V4, FLUX and more.

Python 71,078 6,406 Updated Aug 14, 2026

Go ahead and axolotl questions

Python 12,352 1,402 Updated Aug 13, 2026

KSS: Korean String processing Suite

Python 471 63 Updated Nov 13, 2025

LangChain 공식 Document, Cookbook, 그 밖의 실용 예제를 바탕으로 작성한 한국어 튜토리얼입니다. 본 튜토리얼을 통해 LangChain을 더 쉽고 효과적으로 사용하는 방법을 배울 수 있습니다.

Jupyter Notebook 2,041 733 Updated Aug 18, 2025

Open-source desktop app for local LLMs. Text, vision, tool-calling, OpenAI/Anthropic-compatible API. 100% private.

Python 47,544 5,979 Updated Jun 2, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,995 20,657 Updated Aug 14, 2026

A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.

Python 2,936 220 Updated Sep 30, 2023

📚 Tech blogs & talks by companies that run Kafka in production

981 81 Updated Aug 2, 2026

Chat with your database or your datalake (SQL, CSV, parquet). PandasAI makes data analysis conversational using LLMs and RAG.

Python 23,738 2,342 Updated Oct 28, 2025

StableLM: Stability AI Language Models

Jupyter Notebook 15,685 1,003 Updated Apr 8, 2024

Running large language models on a single GPU for throughput-oriented scenarios.

Python 9,358 592 Updated Oct 28, 2024

Pecab: Pure python Korean morpheme analyzer based on Mecab

Python 172 14 Updated Apr 27, 2024

Source code for end-to-end dialogue model from the MultiWOZ paper (Budzianowski et al. 2018, EMNLP)

Python 955 206 Updated Apr 18, 2026

List of Korean pre-trained language models.

188 15 Updated Aug 31, 2023
Next