- United States
Stars
A comprehensive, enterprise-ready toolkit for deploying Agentic AI systems on Intel® Xeon processors and Intel accelerators. Built for organizations that need to move from prototype to production q…
LLM Inference analyzer for different hardware platforms
Codebase for layer wise N:M pruning pattern assignment for LLMs
[COLM 2025] SEAL: Steerable Reasoning Calibration of Large Language Models for Free
[COLM'25] CITER: Collaborative Inference for Efficient Large Language Model Decoding with Token-Level Routing
Official Implementation of LANTERN (ICLR'25) and LANTERN++(ICLRW-SCOPE'25)
[ECCV 2024] CLAMP-ViT: Contrastive Data-Free Learning for Adaptive Post-Training Quantization of ViTs
GEAR: An Efficient KV Cache Compression Recipefor Near-Lossless Generative Inference of LLM
ShiftAddLLM: Accelerating Pretrained LLMs via Post-Training Multiplication-Less Reparameterization
Refine high-quality datasets and visual AI models
This repository constains a Pytorch implementation of the paper, titled "Pre-defined Sparsity for Low-Complexity Convolutional Neural Networks"
Sparse learning library and sparse momentum resources.