Skip to content
View SimJeg's full-sized avatar

Block or report SimJeg

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Illico is a python library performing fast and lightweight wilcoxon rank-sum tests for single-cell RNASeq.

Python 21 4 Updated Jul 30, 2026

NVIDIA Object Oriented Agents: the Pythonic way to build AI Agents.

Python 1,518 200 Updated Aug 13, 2026

Framework for evaluating and improving agents

Python 4,168 1,547 Updated Aug 13, 2026

LLM KV cache compression made easy

Python 1,170 168 Updated Aug 10, 2026

Codebase leveraging computer vision for aerodynamics measurement on sailboats. Calibrated photogramometry based 3d sailshape reconstruction and video tracking of tell tales.

Python 45 1 Updated Jun 3, 2026

The open source coding agent.

TypeScript 196,789 25,291 Updated Aug 13, 2026

NVARC solution to ARC-AGI-2

Jupyter Notebook 137 31 Updated Jan 15, 2026

The evaluation framework for training-free sparse attention in LLMs

Python 129 12 Updated Jan 27, 2026

[NeurIPS 2025 D&B] 🚀 SWE-bench Goes Live!

Python 220 31 Updated Jun 11, 2026

The #1 open-source SWE-bench Verified implementation

Python 878 151 Updated Jun 9, 2025

Reference implementation of the Jupyter Notebook format

Python 311 174 Updated Aug 10, 2026

An extremely fast Python package and project manager, written in Rust.

Rust 88,699 3,483 Updated Aug 13, 2026

[MLSys'25] QServe: W4A8KV4 Quantization and System Co-design for Efficient LLM Serving; [MLSys'25] LServe: Efficient Long-sequence LLM Serving with Unified Sparse Attention

C++ 853 67 Updated Mar 6, 2025

The NVIDIA NeMo Agent toolkit is an open-source library for efficiently connecting and optimizing teams of AI agents.

Python 2,574 736 Updated Aug 12, 2026

[ICLR 2025] DuoAttention: Efficient Long-Context LLM Inference with Retrieval and Streaming Heads

Python 540 41 Updated Feb 10, 2025

NeMo Retriever Library is a scalable, performance-oriented document content and metadata extraction microservice. NeMo Retriever Library uses specialized NVIDIA NIM microservices to find, contextua…

Python 2,963 339 Updated Aug 13, 2026

Cold Compress is a hackable, lightweight, and open-source toolkit for creating and benchmarking cache compression methods built on top of GPT-Fast, a simple, PyTorch-native generation codebase.

Python 153 16 Updated Aug 9, 2024

The simplest implementation of recent Sparse Attention patterns for efficient LLM inference.

Jupyter Notebook 92 6 Updated Jul 17, 2025

A collection of LogitsProcessors to customize and enhance LLM behavior for specific tasks.

Python 388 27 Updated Jul 8, 2025

📰 Must-read papers on KV Cache Compression (constantly updating 🤗).

733 28 Updated Apr 15, 2026

Awesome LLM compression research papers and tools.

1,860 129 Updated Jun 30, 2026
Python 22 2 Updated Apr 17, 2025

♟️ Vectorized RL game environments in JAX

Python 636 54 Updated Mar 6, 2025

A framework for few-shot evaluation of language models.

Python 13,608 3,477 Updated Aug 11, 2026

This repo contains the source code for RULER: What’s the Real Context Size of Your Long-Context Language Models?

Python 1,603 135 Updated Jul 22, 2026

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Python 74,056 9,059 Updated Aug 10, 2026

Code for the paper "Visual Anagrams: Generating Multi-View Optical Illusions with Diffusion Models"

Jupyter Notebook 968 102 Updated Jul 17, 2026

Generative Representational Instruction Tuning

Jupyter Notebook 697 49 Updated Jun 25, 2025

Universal markup converter

Haskell 45,840 3,948 Updated Aug 13, 2026

Create and modify Word documents with Python

Python 5,693 1,298 Updated Aug 1, 2026
Next