Skip to content
View hitluobin's full-sized avatar
  • Tencent
  • China

Block or report hitluobin

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Algorithm powering the For You feed on X

Rust 26,953 4,590 Updated May 15, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 88,510 20,433 Updated Aug 8, 2026

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Python 42,885 4,922 Updated Aug 7, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,086 1,582 Updated Aug 8, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 73,912 9,042 Updated Aug 6, 2026

ValueCell is a community-driven, multi-agent platform for financial applications.

Python 10,962 1,814 Updated Mar 9, 2026

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,206 3,237 Updated Jun 11, 2026

本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)

HTML 24,875 2,840 Updated Jul 19, 2026

Source code for the X Recommendation Algorithm

Scala 73,660 13,260 Updated Sep 8, 2025

a TensorFlow-based distributed training framework optimized for large-scale sparse data.

C++ 334 71 Updated Apr 10, 2026

DeepRec is a high-performance recommendation deep learning framework based on TensorFlow. It is hosted in incubation in LF AI & Data Foundation.

C++ 1,198 360 Updated Jan 21, 2025

TePDist (TEnsor Program DISTributed) is an HLO-level automatic distributed system for DL models.

C++ 96 11 Updated Apr 22, 2023

A framework for large scale recommendation algorithms.

Python 2,353 385 Updated Apr 15, 2026

Text Normalization & Inverse Text Normalization

Python 807 115 Updated Jul 29, 2026

A fast, distributed, high performance gradient boosting (GBT, GBDT, GBRT, GBM or MART) framework based on decision tree algorithms, used for ranking, classification and many other machine learning …

C++ 18,672 4,049 Updated Aug 8, 2026

Graphormer is a general-purpose deep learning backbone for molecular modeling.

Python 2,468 373 Updated Jun 12, 2026

Feature engineering is the process of using domain knowledge to extract features from raw data via data mining techniques. These features can be used to improve the performance of machine learning …

Jupyter Notebook 803 276 Updated Jun 29, 2025

The official repository for ERNIE 4.5 and ERNIEKit – its industrial-grade development toolkit based on PaddlePaddle.

Python 7,735 1,445 Updated Jul 24, 2026

FinRL®: Financial Reinforcement Learning. 🔥

Jupyter Notebook 15,950 3,455 Updated Jul 13, 2026

Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!

Python 9,481 1,198 Updated Apr 2, 2024

A deep matching model library for recommendations & advertising. It's easy to train models and to export representation vectors which can be used for ANN search.

Python 2,436 542 Updated Apr 18, 2026

An Industrial Graph Neural Network Framework

C++ 1,341 266 Updated Jul 4, 2025

A framework for cleaning Chinese dialog data

Python 274 30 Updated May 14, 2021

Fast, efficiently stored Trie for Python. Uses libdatrie.

Cython 548 93 Updated Jun 1, 2026

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,256 11,157 Updated Jul 22, 2026

Chinese version of GPT2 training code, using BERT tokenizer.

Python 7,600 1,683 Updated Apr 25, 2024

The pytorch re-implement of the official efficientdet with SOTA performance in real time and pretrained weights.

Jupyter Notebook 5,241 1,245 Updated Oct 24, 2021

高质量中文预训练模型集合:最先进大模型、最快小模型、相似度专门模型

Python 810 94 Updated Jul 8, 2020

超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M

C++ 12,331 2,281 Updated May 18, 2026

A Large-Scale Chinese Cross-Domain Task-Oriented Dialogue Dataset

Python 722 119 Updated Jun 17, 2024
Next