Skip to content
View junhuihe-hjh's full-sized avatar

Block or report junhuihe-hjh

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Extracted system prompts from Anthropic - Claude Fable 5, Opus 5, Claude Design, Claude Code. OpenAI - ChatGPT GPT-5.6-Sol, Codex. Google - Gemini 3.5 Flash, 3.1 Pro, Antigravity. xAI - Grok, Curso…

JavaScript 62,799 10,315 Updated Aug 7, 2026

🚀 A curated collection of papers focusing on LLM-based quantitative trading.

224 25 Updated Jul 24, 2026

Official JAX implementation of End-to-End Test-Time Training for Long Context

Python 643 48 Updated Feb 15, 2026

武汉大学博士/硕士学位论文latex模板(包含插图索引、表格索引、中英文缩略语对照、主要符号表等,字体格式以及排版修正)

TeX 32 4 Updated Feb 18, 2025

Lightweight coding agent that runs in your terminal

Rust 105,620 16,019 Updated Aug 13, 2026

The best ChatGPT that $100 can buy.

Python 57,160 7,917 Updated Aug 2, 2026

Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞

TypeScript 386,110 81,157 Updated Aug 13, 2026

An agent-first mobile runtime built on OpenClaw: control apps, learn reusable skills, and connect phones.

TypeScript 132 23 Updated Jul 24, 2026

A general memory system for agents, powered by deep-research

Python 859 86 Updated Mar 14, 2026

Official implementation of Accelerating Prefilling for Long-Context Inference via Sparse Pattern Sharing

Python 3 Updated Dec 16, 2025

Unofficial implementations of block/layer-wise pruning methods for LLMs.

Jupyter Notebook 78 19 Updated Apr 29, 2024

Official implementation for LaCo (EMNLP 2024 Findings)

Jupyter Notebook 22 5 Updated Oct 3, 2024

For releasing code related to compression methods for transformers, accompanying our publications

Python 462 58 Updated Jan 16, 2025

[ICML 2024] Official Implementation of SLEB: Streamlining LLMs through Redundancy Verification and Elimination of Transformer Blocks

Python 43 6 Updated Feb 4, 2025

Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks

Python 4,335 743 Updated Aug 11, 2026

One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks

Python 4,360 638 Updated Aug 6, 2026

Official Repo of "MMBench: Is Your Multi-modal Model an All-around Player?"

310 19 Updated May 22, 2025

🚀 A curated list of awesome resources focusing on Context Compression techniques for Large Language Models(LLMs).

HTML 78 1 Updated Jan 17, 2026

CUDA Python: Performance meets Productivity

Cython 3,341 321 Updated Aug 13, 2026

[COLM 2025] Official PyTorch implementation of "Quantization Hurts Reasoning? An Empirical Study on Quantized Reasoning Models"

Python 77 8 Updated Jul 8, 2025

This is the official code for ZLST-Project, Generative Recommendation Benchmark

Python 68 4 Updated May 25, 2026

Official code implementation of Context Cascade Compression: Exploring the Upper Limits of Text Compression

Python 317 6 Updated Jan 27, 2026

Official repository of RARE: Retrieval-Augmented Reasoning Modeling [KDD 2026 Research Track]

Python 184 32 Updated May 20, 2026

A character-level language diffusion model trained on Tiny Shakespeare

Python 926 90 Updated Jan 16, 2026

Official PyTorch implementation for "Large Language Diffusion Models"

Python 3,933 272 Updated Jul 15, 2026

A quick rundown on each feature and its settings

866 53 Updated May 15, 2026

[ICML 2024] Quest: Query-Aware Sparsity for Efficient Long-Context LLM Inference

Cuda 400 50 Updated Jul 10, 2025

SeerAttention: Learning Intrinsic Sparse Attention in Your LLMs

Python 216 20 Updated Jul 10, 2026

[EMNLP'23, ACL'24] To speed up LLMs' inference and enhance LLM's perceive of key information, compress the prompt and KV-Cache, which achieves up to 20x compression with minimal performance loss.

Python 6,551 413 Updated Apr 8, 2026
Next