Skip to content
View seyuboglu's full-sized avatar

Block or report seyuboglu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Stable Looped Models and their Scaling Laws

Python 174 12 Updated May 17, 2026

Algorithms for latent compaction

Python 263 28 Updated Apr 22, 2026

Archive a lifetime of email and chat. Offline search, analytics, and AI query over your full message history. Powered by SQLite and DuckDB

Go 2,004 140 Updated Aug 13, 2026

The context API to search, scrape, and interact with the web at scale. 🔥

TypeScript 166,725 9,366 Updated Aug 13, 2026

Processed / Cleaned Data for Paper Copilot

Python 954 47 Updated Jul 1, 2026

Storing long contexts in tiny caches with self-study

Python 312 42 Updated Mar 23, 2026

KV cache compression via sparse coding

Python 18 5 Updated Oct 26, 2025

A minimalistic framework for transparently training language models and storing comprehensive checkpoints for in-depth learning dynamics research.

Python 319 31 Updated Feb 19, 2026

Model Context Protocol Servers

TypeScript 89,528 11,437 Updated Aug 10, 2026

Big & Small LLMs working together

Python 1,346 150 Updated Mar 12, 2026

Democratizing Reinforcement Learning for LLMs

Python 5,784 602 Updated Aug 13, 2026

Discovering Interpretable Features in Protein Language Models via Sparse Autoencoders

Python 299 45 Updated Oct 31, 2025

Aioli: A unified optimization framework for language model data mixing

Jupyter Notebook 33 4 Updated Jan 17, 2025

Archon provides a modular framework for combining different inference-time techniques and LMs with just a JSON config file.

Python 209 31 Updated Mar 7, 2025

[ICLR2025] Breaking Throughput-Latency Trade-off for Long Sequences with Speculative Decoding

Python 156 12 Updated Dec 4, 2024

Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.

Python 6,245 573 Updated Aug 22, 2025

[NeurIPS 2024] Simple and Effective Masked Diffusion Language Model

Python 708 103 Updated Sep 29, 2025

(NeurIPS 2024) AvaTaR: Optimizing LLM Agents for Tool Usage via Contrastive Reasoning

Python 241 27 Updated Jun 10, 2025

TextGrad: Automatic ''Differentiation'' via Text -- using large language models to backpropagate textual gradients. Published in Nature.

Python 3,694 294 Updated Jul 25, 2025

The Fast Cross-Platform Package Manager

C++ 8,080 446 Updated Aug 12, 2026

Tile primitives for speedy kernels

Cuda 3,630 319 Updated Jul 13, 2026

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]

Python 20,055 2,194 Updated Aug 10, 2026

A curated reading list of research in Adaptive Computation, Inference-Time Computation & Mixture of Experts (MoE).

164 9 Updated Jan 1, 2025

Triton-based implementation of Sparse Mixture of Experts.

Python 281 29 Updated Oct 3, 2025

Code for exploring Based models from "Simple linear attention language models balance the recall-throughput tradeoff"

Python 256 19 Updated Jun 6, 2025

This repo contains data and code for the paper "Language Models Enable Simple Systems for Generating Structured Views of Heterogeneous Data Lakes"

Python 497 46 Updated Mar 26, 2024

🚀 Efficient implementations for emerging model architectures

Python 5,554 656 Updated Aug 13, 2026

Understand and test language model architectures on synthetic tasks.

Python 283 55 Updated Mar 22, 2026
Next