Skip to content
View qiguanjie's full-sized avatar

Highlights

  • Pro

Block or report qiguanjie

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Python tool for converting files and office documents to Markdown.

Python 173,933 12,704 Updated Jul 29, 2026

14MB foundation model for tiny devices; phones, wearables, smart home, and robots.

Python 6,081 406 Updated Aug 15, 2026

Bash is all you need - A nano claude code–like 「agent harness」, built from 0 to 1

Python 74,311 12,033 Updated Aug 15, 2026

夫子•明察司法大模型是由山东大学、浪潮云、中国政法大学联合研发,以 ChatGLM 为大模型底座,基于海量中文无监督司法语料与有监督司法微调数据训练的中文司法大模型。该模型支持法条检索、案例分析、三段论推理判决以及司法对话等功能,旨在为用户提供全方位、高精准的法律咨询与解答服务。

Python 387 25 Updated Jul 30, 2025

[中文法律大模型] DISC-LawLLM: an intelligent legal system powered by large language models (LLMs) to provide a wide range of legal services.

Python 947 93 Updated May 27, 2025

LexiLaw - 中文法律大模型

Python 1,038 146 Updated Mar 12, 2026

用文本编辑器剪视频

Python 7,784 824 Updated Oct 5, 2024

Instruct-tune LLaMA on consumer hardware

Jupyter Notebook 18,908 2,174 Updated Jul 29, 2024

Code for loralib, an implementation of "LoRA: Low-Rank Adaptation of Large Language Models"

Python 13,738 921 Updated Dec 17, 2024

Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"

Python 71 10 Updated Feb 27, 2024

Code & Data for our Paper "Alleviating Hallucinations of Large Language Models through Induced Hallucinations"

Python 1 Updated Dec 27, 2023

The agent engineering platform.

Python 144,295 24,022 Updated Aug 15, 2026

Mirror of SRILM

Roff 61 25 Updated Aug 11, 2020
Python 8 Updated Jul 13, 2021

Unsupervised text tokenizer for Neural Network-based text generation.

C++ 12,024 1,376 Updated Aug 16, 2026

Inference code for Llama models

Python 59,557 9,791 Updated Jan 26, 2025

Best practice for training LLaMA models in Megatron-LM

Python 666 56 Updated Jan 2, 2024

A No-Recurrence Sequence-to-Sequence Model for Speech Recognition

Python 378 66 Updated Jul 21, 2022

Text Normalization & Inverse Text Normalization

Python 809 114 Updated Jul 29, 2026

中文标点符号模型,可以给文本添加标点符号。

Python 145 23 Updated Dec 24, 2024

基于PaddlePaddle实现端到端中文语音识别,从入门到实战,超简单的入门案例,超实用的企业项目。支持当前最流行的DeepSpeech2、Conformer、Squeezeformer模型

Python 872 129 Updated Dec 17, 2025

CDCPP: Cross-Domain Chinese Punctuation Prediction

11 5 Updated Feb 27, 2022

Robust Speech Recognition via Large-Scale Weak Supervision

Python 107,333 13,038 Updated Jul 28, 2026

Data repository for pretrained NLP models and NLP corpora.

Python 1,058 142 Updated Mar 16, 2018

Open-source speech recognition toolkit for training, inference, streaming ASR, VAD, punctuation, speaker diarization pipelines, and OpenAI-compatible/MCP serving.

Python 19,855 1,990 Updated Aug 14, 2026

FunASR-powered video transcription, subtitle generation, and LLM-assisted clipping tool with a local Gradio UI.

Python 6,143 736 Updated Aug 3, 2026

VITS: Conditional Variational Autoencoder with Adversarial Learning for End-to-End Text-to-Speech

Python 7,889 1,382 Updated Dec 6, 2023

Universal Romanizer that can convert any unicode script to roman (latin) script

Perl 249 23 Updated Jul 26, 2024

Unsupervised phone and word segmentation using dynamic programming on self-supervised VQ features.

Jupyter Notebook 39 8 Updated May 5, 2026
Python 1 2 Updated Oct 25, 2022
Next