Skip to content
View gigit0000's full-sized avatar
  • Kim Baksa's Lab, South Korea
  • 17:31 (UTC +09:00)

Block or report gigit0000

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Windows alt-tab on macOS

Swift 16,144 823 Updated Aug 6, 2026

Native space switching on macOS with no animation

Swift 1,939 66 Updated Jun 21, 2026

macOS system monitor in your menu bar

Swift 41,043 1,462 Updated Aug 8, 2026

DeepSeek 4 Flash and PRO local inference engine for Metal, CUDA and ROCm

C 21,006 1,884 Updated Aug 5, 2026

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,138 401 Updated Aug 6, 2026

📚A curated list of Awesome LLM/VLM Inference Papers with Codes: Flash-Attention, Paged-Attention, WINT8/4, Parallelism, etc.🎉

Python 5,447 431 Updated Jul 26, 2026

Easy, Fast, and Scalable Multimodal AI

Python 130 12 Updated Jun 2, 2026

Nvidia Instruction Set Specification Generator

Python 348 23 Updated Jul 9, 2024

Cataloging released Triton kernels.

310 19 Updated Sep 9, 2025

Learning Deep Representations of Data Distributions

TeX 1,020 111 Updated Aug 7, 2026

Small scale distributed training of sequential deep learning models, built on Numpy and MPI.

Python 165 9 Updated Oct 19, 2023

Python pdb for multiple processes

Python 82 9 Updated May 24, 2025

Large-scale LLM inference engine

C++ 1,824 206 Updated Aug 8, 2026

Memray is a memory profiler for Python

Python 15,186 457 Updated Aug 7, 2026
Python 7 Updated Jul 26, 2025

Triton Support in Compiler Explorer

TypeScript 5 Updated Aug 5, 2025

Run compilers interactively from your web browser and interact with the assembly

TypeScript 18,961 2,089 Updated Aug 8, 2026

This repo provides several classic attention variant implementation based on FlexAttention API.

Python 2 1 Updated May 18, 2025

A unified library for building, evaluating, and storing speculative decoding algorithms for LLM inference in vLLM

Python 706 179 Updated Aug 7, 2026

OpenBLAS is an optimized BLAS library based on GotoBLAS2 1.13 BSD version.

C 7,551 1,704 Updated Aug 8, 2026

Hacker News

HTML 17 8 Updated Aug 9, 2026

Distribute and run LLMs with a single file.

C++ 1 Updated Jul 23, 2024

Distribute and run LLMs with a single file.

C++ 25,518 1,544 Updated Aug 3, 2026

CUDA on non-NVIDIA GPUs

Rust 14,737 932 Updated Aug 9, 2026

GPU & Accelerator process monitoring for AMD, Apple, Huawei, Intel, NVIDIA and Qualcomm

C 10,896 416 Updated May 6, 2026

A .NET MAUI app for displaying the top posts on Hacker News that demonstrates text sentiment analysis gathered using artificial intelligence

C# 279 40 Updated Aug 6, 2026

A curated list of awesome C frameworks, libraries, resources and other shiny things. Inspired by all the other awesome-... projects out there.

11,457 945 Updated Dec 27, 2025

Local AI voice assistant stack for Home Assistant (GPU-accelerated) with persistent memory, follow-up conversation, and Ollama model recommendations - settings designed for low VRAM systems.

246 23 Updated Jul 27, 2025

Debug Module for Embedded Systems

C 1 Updated Mar 20, 2026
Next