Skip to content
View hitluobin's full-sized avatar
  • Tencent
  • China

Block or report hitluobin

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Algorithm powering the For You feed on X

Rust 31,590 5,178 Updated Aug 14, 2026

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,231 20,794 Updated Aug 17, 2026

DeepSpeed is a deep learning optimization library that makes distributed training and inference easy, efficient, and effective.

Python 42,946 4,934 Updated Aug 17, 2026

Use PEFT or Full-parameter to CPT/SFT/DPO/GRPO 600+ LLMs (Qwen3.6, DeepSeek-V4, GLM-5.1, InternLM3, Llama4, ...) and 300+ MLLMs (Qwen3-VL, Qwen3-Omni, InternVL3.5, Ovis2.5, GLM4.5v, Gemma4, Llava, …

Python 15,243 1,603 Updated Aug 17, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,155 9,072 Updated Aug 13, 2026

ValueCell is a community-driven, multi-agent platform for financial applications.

Python 11,001 1,814 Updated Mar 9, 2026

"AI-Trader: 100% Fully-Automated Agent-Native Trading"

Python 21,432 3,275 Updated Jun 11, 2026

本项目旨在分享大模型相关技术原理以及实战经验(大模型工程化、大模型应用落地)

HTML 24,899 2,840 Updated Jul 19, 2026

Source code for the X Recommendation Algorithm

Scala 73,794 13,269 Updated Sep 8, 2025

a TensorFlow-based distributed training framework optimized for large-scale sparse data.

C++ 334 71 Updated Apr 10, 2026

DeepRec is a high-performance recommendation deep learning framework based on TensorFlow. It is hosted in incubation in LF AI & Data Foundation.

C++ 1,197 359 Updated Jan 21, 2025

TePDist (TEnsor Program DISTributed) is an HLO-level automatic distributed system for DL models.

C++ 96 11 Updated Apr 22, 2023

A framework for large scale recommendation algorithms.

Python 2,354 385 Updated Apr 15, 2026

Text Normalization & Inverse Text Normalization

Python 809 114 Updated Jul 29, 2026

A fast, distributed, high performance gradient boosting (GBT, GBDT, GBRT, GBM or MART) framework based on decision tree algorithms, used for ranking, classification and many other machine learning …

C++ 18,688 4,052 Updated Aug 14, 2026

Graphormer is a general-purpose deep learning backbone for molecular modeling.

Python 2,469 373 Updated Jun 12, 2026

Feature engineering is the process of using domain knowledge to extract features from raw data via data mining techniques. These features can be used to improve the performance of machine learning …

Jupyter Notebook 807 276 Updated Jun 29, 2025

The official repository for ERNIE 4.5 and ERNIEKit – its industrial-grade development toolkit based on PaddlePaddle.

Python 7,736 1,445 Updated Jul 24, 2026

FinRL®: Financial Reinforcement Learning. 🔥

Jupyter Notebook 16,027 3,465 Updated Jul 13, 2026

Robust Video Matting in PyTorch, TensorFlow, TensorFlow.js, ONNX, CoreML!

Python 9,485 1,195 Updated Apr 2, 2024

A deep matching model library for recommendations & advertising. It's easy to train models and to export representation vectors which can be used for ANN search.

Python 2,435 541 Updated Apr 18, 2026

An Industrial Graph Neural Network Framework

C++ 1,340 266 Updated Jul 4, 2025

A framework for cleaning Chinese dialog data

Python 275 30 Updated May 14, 2021

Fast, efficiently stored Trie for Python. Uses libdatrie.

Cython 547 93 Updated Jun 1, 2026

Turn any PDF or image document into structured data for your AI. A powerful, lightweight OCR toolkit that bridges the gap between images/PDFs and LLMs. Supports 100+ languages.

Python 87,770 11,182 Updated Jul 22, 2026

Chinese version of GPT2 training code, using BERT tokenizer.

Python 7,598 1,683 Updated Apr 25, 2024

The pytorch re-implement of the official efficientdet with SOTA performance in real time and pretrained weights.

Jupyter Notebook 5,240 1,245 Updated Oct 24, 2021

高质量中文预训练模型集合:最先进大模型、最快小模型、相似度专门模型

Python 810 94 Updated Jul 8, 2020

超轻量级中文ocr,支持竖排文字识别, 支持ncnn、mnn、tnn推理 ( dbnet(1.8M) + crnn(2.5M) + anglenet(378KB)) 总模型仅4.7M

C++ 12,332 2,279 Updated May 18, 2026

A Large-Scale Chinese Cross-Domain Task-Oriented Dialogue Dataset

Python 726 119 Updated Jun 17, 2024
Next