Skip to content
View EunjuYang's full-sized avatar
πŸ‘
:)
πŸ‘
:)

Organizations

@KAIST-NCL @nnstreamer @nntrainer @cloud-gtm

Block or report EunjuYang

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

A collection of sample agents built with Agent Development Kit (ADK)

Python 10,008 2,776 Updated Aug 1, 2026

This repo is for code that supports talks, blogs, and other activities the Google Cloud Developer Relations team engages in.

Jupyter Notebook 344 169 Updated Jul 31, 2026

A gallery that showcases on-device ML/GenAI use cases and allows people to try and use models locally.

Kotlin 24,320 2,593 Updated Jul 31, 2026

CausalLM

C++ 7 8 Updated May 7, 2026

An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex β€” developed and maintained with no human intervention.

Rust 194,956 109,379 Updated Jun 26, 2026

TurboQuant: Near-optimal KV cache quantization for LLM inference (3-bit keys, 2-bit values) with Triton kernels + vLLM integration

Python 1,707 189 Updated Mar 27, 2026

PyTorch building blocks for the OLMo ecosystem

Python 1,443 297 Updated Aug 1, 2026

Modeling, training, eval, and inference code for OLMo

Python 6,612 791 Updated Nov 24, 2025

A more memory-efficient rewrite of the HF transformers implementation of Llama for use with quantized weights.

Python 2,935 220 Updated Sep 30, 2023

Use Garry Tan's exact Claude Code setup: 23 opinionated tools that serve as CEO, Designer, Eng Manager, Release Manager, Doc Engineer, and QA

TypeScript 125,615 18,846 Updated Jul 15, 2026

250+ Fine-tuning & RL Notebooks for text, vision, audio, embedding, TTS models.

Jupyter Notebook 5,539 916 Updated Jul 30, 2026

"RAG-Anything: All-in-One RAG Framework"

Python 22,500 2,624 Updated Jul 20, 2026

The official implementation for [NeurIPS2025 Oral] Gated Attention for Large Language Models: Non-linearity, Sparsity, and Attention-Sink-Free

Jupyter Notebook 974 61 Updated Dec 20, 2025

omo/lazycodex: The coding agent for tokenmaxxers;the one and only agent harness for complex codebases. For your Codex, for your OpenCode

TypeScript 66,967 5,463 Updated Aug 1, 2026

Universal memory layer for AI Agents

Python 62,236 7,252 Updated Aug 1, 2026

This repository is a read-only mirror of https://gitlab.arm.com/kleidi/kleidiai

C++ 174 25 Updated Jul 31, 2026

Official PyTorch implementation for Hogwild! Inference: Parallel LLM Generation with a Concurrent Attention Cache

Python 142 10 Updated Aug 13, 2025

Accelerate local LLM inference and finetuning (LLaMA, Mistral, ChatGLM, Qwen, DeepSeek, Mixtral, Gemma, Phi, MiniCPM, Qwen-VL, MiniCPM-V, etc.) on Intel XPU (e.g., local PC with iGPU and NPU, discr…

Python 8,865 1,428 Updated Jan 28, 2026

MiniMax-M1, the world's first open-weight, large-scale hybrid-attention reasoning model.

Python 3,163 283 Updated Jul 7, 2025

The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.

Python 446 15 Updated Jul 11, 2025

[Fully open] [Encoder-free MLLM] Vision as LoRA

Python 389 31 Updated Jun 12, 2025

[NeurIPS 2025] TTRL: Test-Time Reinforcement Learning

Python 1,107 82 Updated Apr 15, 2026

NVIDIA Isaac GR00T N1.7 - A Foundation Model for Generalist Robots.

Python 7,718 1,376 Updated Jul 30, 2026

Code release for DynamicTanh (DyT)

Python 1,042 87 Updated Mar 30, 2025
Python 38 8 Updated May 30, 2025

SGLang is a high-performance serving framework for large language models and multimodal models.

Python 31,033 7,564 Updated Aug 1, 2026
C++ 324 96 Updated Jul 22, 2026

Paper list for Personal LLM Agents

433 27 Updated Jun 27, 2026
Next