Skip to content
View xwinxu's full-sized avatar
💭
alignment and llms.
💭
alignment and llms.

Organizations

@Cohere-Labs-Community @VectorInstitute @UTMIST

Block or report xwinxu

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 591 61 Updated Jul 11, 2024

Home for "How To Scale Your Model", a short blog-style textbook about scaling LLMs on TPUs

HTML 1,325 189 Updated Jul 27, 2026

Official Implementation of ACL 2021 paper “GeoQA: A Geometric Question Answering Benchmark Towards Multimodal Numerical Reasoning”.

Python 77 8 Updated Jan 10, 2022
Jupyter Notebook 1,222 777 Updated Jul 24, 2026

The official evaluation suite and dynamic data release for MixEval.

Python 254 40 Updated Nov 10, 2024

RewardBench: the first evaluation tool for reward models.

Python 731 97 Updated Feb 16, 2026

What would you do with 1000 H100s...

Jupyter Notebook 1,188 72 Updated Jan 10, 2024

Puzzles for learning Triton

Jupyter Notebook 2,556 248 Updated Apr 1, 2026

18.06 course at MIT

Jupyter Notebook 3,267 786 Updated Apr 3, 2026

Robust recipes to align language models with human and AI preferences

Python 5,658 489 Updated May 26, 2026

Scalable training for dense retrieval models.

Python 298 32 Updated Jul 2, 2026

Python pdb for multiple processes

Python 82 9 Updated May 24, 2025

Machine Learning Engineering Open Book

Python 18,594 1,198 Updated Aug 12, 2026

Welcome to the Llama Cookbook! This is your go to guide for Building with Llama: Getting started with Inference, Fine-Tuning, RAG. We also show you how to solve end to end problems using Llama mode…

Jupyter Notebook 18,554 2,765 Updated May 19, 2026

Train transformer language models with reinforcement learning.

Python 19,054 2,903 Updated Aug 12, 2026

The simplest, fastest repository for training/finetuning medium-sized GPTs.

Python 62,049 10,691 Updated Nov 12, 2025

The TinyLlama project is an open endeavor to pretrain a 1.1B Llama model on 3 trillion tokens.

Python 9,019 628 Updated May 3, 2024

Expanding natural instructions

Python 1,045 197 Updated Dec 11, 2023

Ongoing research training transformer models at scale

Python 17,398 4,356 Updated Aug 12, 2026

Provides end-to-end model development pipelines for LLMs and Multimodal models that can be launched on-prem or cloud-native.

Python 521 150 Updated Apr 18, 2025

Reference implementation for DPO (Direct Preference Optimization)

Python 2,904 237 Updated Aug 11, 2024

Supercharge Your LLM Application Evaluations 🚀

Python 15,285 1,623 Updated Feb 24, 2026

Inference code for Llama models

Python 59,552 9,793 Updated Jan 26, 2025

[NeurIPS 2023] MeZO: Fine-Tuning Language Models with Just Forward Passes. https://arxiv.org/abs/2305.17333

Python 1,173 89 Updated Jan 11, 2024

Oh my tmux! My self-contained, pretty & versatile tmux configuration made with 💛🩷💙🖤❤️🤍

Shell 25,287 3,598 Updated Aug 8, 2026

This repository is to prepare for Machine Learning interviews.

1,656 398 Updated May 19, 2019

Instruct-tune LLaMA on consumer hardware

Jupyter Notebook 18,909 2,177 Updated Jul 29, 2024

An interactive exploration of Transformer programming.

Jupyter Notebook 277 21 Updated Nov 15, 2023
Next