Skip to content
View kohjingyu's full-sized avatar
🫠
🫠

Highlights

  • Pro

Block or report kohjingyu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Swift 15 1 Updated Jun 9, 2026

Code for the multi-agent computer use project.

Python 21 3 Updated Jul 3, 2026
Python 11 1 Updated May 30, 2026

Playwright MCP server

TypeScript 36,093 3,019 Updated Aug 12, 2026

Code for the paper 🌳 Tree Search for Language Model Agents

Python 223 24 Updated Jul 25, 2024

VisualWebArena is a benchmark for multimodal agents.

Python 485 81 Updated Nov 9, 2024

800,000 step-level correctness labels on LLM solutions to MATH problems

Python 2,151 129 Updated Jun 1, 2023

🎨 Classic MS Paint, REVIVED + ✨Extras

JavaScript 7,849 667 Updated Jul 19, 2026

An open-source framework for training large multimodal models.

Python 4,118 319 Updated Aug 31, 2024

🐟 Code and models for the NeurIPS 2023 paper "Generating Images with Multimodal Language Models".

Jupyter Notebook 470 37 Updated Jan 19, 2024

MultimodalC4 is a multimodal extension of c4 that interleaves millions of images with text.

Python 954 38 Updated Mar 19, 2025

Easily compute clip embeddings and build a clip retrieval system with them

Jupyter Notebook 2,792 238 Updated Mar 28, 2026

Measuring Massive Multitask Language Understanding | ICLR 2021

Python 1,608 117 Updated May 28, 2023

Contrastive Fact Verification

Python 74 11 Updated Sep 17, 2022

by ex-googlers, for ex-googlers - a lookup table of similar tech & services

15,692 1,084 Updated Jan 12, 2026

🧀 Code and models for the ICML 2023 paper "Grounding Language Models to Images for Multimodal Inputs and Outputs".

Jupyter Notebook 484 37 Updated Oct 30, 2023

Cramming the training of a (BERT-type) language model into limited compute.

Python 1,368 105 Updated Jun 13, 2024

The simplest, fastest repository for training/finetuning medium-sized GPTs.

Python 62,090 10,698 Updated Nov 12, 2025

🤗 Diffusers: State-of-the-art diffusion models for image, video, and audio generation in PyTorch.

Python 34,306 7,236 Updated Aug 13, 2026

An open source implementation of CLIP.

Python 14,061 1,301 Updated Aug 10, 2026

MAGMA - a GPT-style multimodal model that can understand any combination of images and language. NOTE: The freely available model from this repo is only a demo. For the latest multimodal and multil…

Python 489 59 Updated Jul 22, 2025

Accessible large language models via k-bit quantization for PyTorch.

Python 8,413 903 Updated Aug 13, 2026

Reference BLEU implementation that auto-downloads test sets and reports a version string to facilitate cross-lab comparisons

Python 1,257 175 Updated Jul 17, 2026

COYO-700M: Large-scale Image-Text Pair Dataset

Python 1,255 38 Updated Nov 30, 2022

This repository hosts the code for our paper, "Simple and Effective Synthesis of Indoor 3D Scenes".

Jupyter Notebook 42 2 Updated Jul 1, 2022

The official implementation of Autoregressive Image Generation using Residual Quantization (CVPR '22)

Jupyter Notebook 1,029 114 Updated Jan 3, 2024

Structured state space sequence models

Jupyter Notebook 2,918 367 Updated Jul 17, 2024

Repo for external large-scale work

Python 6,547 714 Updated Apr 27, 2024

Restricted Boltzmann Machines in Python.

Python 972 374 Updated Apr 1, 2020
Next