Skip to content
View tomyoung903's full-sized avatar

Block or report tomyoung903

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

🤱🏻 Turn any webpage into a desktop app with one command.

Rust 60,573 12,369 Updated Aug 8, 2026

[ICLR 2025] <MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses>

Python 58 9 Updated Nov 12, 2025

AI powered local typing assistant built with Ollama

Python 328 39 Updated Jul 29, 2024

Use pytorch profile api to further analysis the training detailed information, like heaps and stacks, time consuming.

Python 2 1 Updated Jan 3, 2025

The official evaluation suite and dynamic data release for MixEval.

Python 254 40 Updated Nov 10, 2024

Custom CSS Plugin for Visual Studio Code. Based on vscode-icon

JavaScript 1,057 92 Updated Jul 29, 2026

Open-Sora: Democratizing Efficient Video Production for All

Python 29,261 3,004 Updated Apr 9, 2026

SpeeD: A Closer Look at Time Steps is Worthy of Triple Speed-Up for Diffusion Model Training

Python 188 7 Updated Jan 27, 2025

🙌 OpenHands: AI-Driven Development

TypeScript 83,695 10,824 Updated Aug 11, 2026

Byted PyTorch Distributed for Hyperscale Training of LLMs and RLs

Python 1,037 64 Updated Mar 3, 2026

Lossless Training Speed Up by Unbiased Dynamic Data Pruning

Python 347 19 Updated Sep 24, 2024

OpenCompass is an LLM evaluation platform, supporting a wide range of models (Llama3, Mistral, InternLM2,GPT-4,LLaMa2, Qwen,GLM, Claude, etc) over 100+ datasets.

Python 7,295 833 Updated Aug 4, 2026

This repository contains code to quantitatively evaluate instruction-tuned models such as Alpaca and Flan-T5 on held-out tasks.

Python 552 45 Updated Mar 10, 2024

A family of open-sourced Mixture-of-Experts (MoE) Large Language Models

Python 1,694 86 Updated Mar 8, 2024

[ICCV2023] Dataset Quantization

Python 261 19 Updated Jan 6, 2024

A framework for few-shot evaluation of language models.

Python 13,595 3,473 Updated Aug 11, 2026

Reference implementation for DPO (Direct Preference Optimization)

Python 2,904 237 Updated Aug 11, 2024

Efficient Dataset Distillation by Representative Matching

Python 114 9 Updated Feb 28, 2024

Inference code for Llama models

Python 59,553 9,794 Updated Jan 26, 2025

Inconsistencies in Masked Language Models

Python 7 1 Updated Mar 8, 2024

Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM

Python 7,864 674 Updated Jul 27, 2026

Making large AI models cheaper, faster and more accessible

Python 41,432 4,506 Updated Aug 10, 2026

Repository for ACL 2022 paper Mix and Match: Learning-free Controllable Text Generation using Energy Language Models

Python 46 5 Updated Mar 13, 2022

GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)

Python 7,653 600 Updated Jul 25, 2023

Thesis Latex Template for Nanyang Technological University (NTU)

TeX 183 61 Updated Oct 14, 2021

FusedChat is a dialogue dataset. It contains dialogue sessions fusing task-oriented dialogues and open-domain dialogues.

Python 29 2 Updated Jul 20, 2022

Unified MultiWOZ evaluation scripts for the context-to-response task.

Python 59 14 Updated Oct 11, 2023

DSTC8 Track 1 Task 1 End-to-End Multi-Domain Dialog Challenge Result:

Python 403 107 Updated Apr 11, 2023

PyTorch code for ICLR 2019 paper: Global-to-local Memory Pointer Networks for Task-Oriented Dialogue https://arxiv.org/pdf/1901.04713

Python 160 22 Updated Jul 11, 2019
Next