Skip to content
View gigit0000's full-sized avatar
  • Kim Baksa's Lab, South Korea
  • 05:38 (UTC +09:00)

Block or report gigit0000

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Windows alt-tab on macOS

Swift 16,087 792 Updated Jul 9, 2026

Native space switching on macOS with no animation

Swift 1,888 63 Updated Jun 21, 2026

macOS system monitor in your menu bar

Swift 40,754 1,440 Updated Jul 22, 2026

DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm

C 19,168 1,687 Updated Jul 24, 2026

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,089 389 Updated Jul 24, 2026

📚A curated list of Awesome LLM/VLM Inference Papers with Codes: Flash-Attention, Paged-Attention, WINT8/4, Parallelism, etc.🎉

Python 5,413 428 Updated Jun 23, 2026

Easy, Fast, and Scalable Multimodal AI

Python 129 12 Updated Jun 2, 2026

Nvidia Instruction Set Specification Generator

Python 345 23 Updated Jul 9, 2024

Cataloging released Triton kernels.

310 18 Updated Sep 9, 2025

Learning Deep Representations of Data Distributions

TeX 1,017 110 Updated Jun 14, 2026

Small scale distributed training of sequential deep learning models, built on Numpy and MPI.

Python 165 9 Updated Oct 19, 2023

Python pdb for multiple processes

Python 82 9 Updated May 24, 2025

Large-scale LLM inference engine

C++ 1,810 206 Updated Jul 24, 2026

Memray is a memory profiler for Python

Python 15,178 459 Updated Jul 24, 2026
Python 7 Updated Jul 26, 2025

Triton Support in Compiler Explorer

TypeScript 5 Updated Aug 5, 2025

Run compilers interactively from your web browser and interact with the assembly

TypeScript 18,931 2,080 Updated Jul 24, 2026

This repo provides several classic attention variant implementation based on FlexAttention API.

Python 2 1 Updated May 18, 2025

A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM

Python 644 166 Updated Jul 24, 2026

OpenBLAS is an optimized BLAS library based on GotoBLAS2 1.13 BSD version.

C 7,528 1,701 Updated Jul 23, 2026

Hacker News

HTML 17 8 Updated Jul 24, 2026

Distribute and run LLMs with a single file.

C++ 1 Updated Jul 23, 2024

Distribute and run LLMs with a single file.

C++ 25,439 1,514 Updated Jul 24, 2026

CUDA on non-NVIDIA GPUs

Rust 14,642 920 Updated Jul 23, 2026

GPU & Accelerator process monitoring for AMD, Apple, Huawei, Intel, NVIDIA and Qualcomm

C 10,858 410 Updated May 6, 2026

A .NET MAUI app for displaying the top posts on Hacker News that demonstrates text sentiment analysis gathered using artificial intelligence

C# 280 40 Updated Jul 23, 2026

A curated list of awesome C frameworks, libraries, resources and other shiny things. Inspired by all the other awesome-... projects out there.

11,428 940 Updated Dec 27, 2025

Local AI voice assistant stack for Home Assistant (GPU-accelerated) with persistent memory, follow-up conversation, and Ollama model recommendations - settings designed for low VRAM systems.

246 23 Updated Jul 27, 2025

Debug Module for Embedded Systems

C 1 Updated Mar 20, 2026
Next