Stars
A Unified Benchmark For Fidelity, Privacy and Utility of Synthetic Chest Radiographs
The official repo of the paper "MMLongBench Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly"
Inference API for many LLMs and other useful tools for empirical research
Code release for "Debating with More Persuasive LLMs Leads to More Truthful Answers"
HEmile / LLaDA
Forked from ML-GSAI/LLaDAMPS supported fork of "Large Language Diffusion Models"
Open source replication of Anthropic's Crosscoders for Model Diffing
Fully open reproduction of DeepSeek-R1
Agent Laboratory is an end-to-end autonomous research workflow meant to assist you as the human researcher toward implementing your research ideas
Large Concept Models: Language modeling in a sentence representation space
This repository contains the NarrativeQA dataset. It includes the list of documents with Wikipedia summaries, links to full stories, and questions and answers.
Stanford NLP Python library for Representation Finetuning (ReFT)
Learning Binding Affinities via Fine-tuning of Protein and Ligand Language Models
A package to generate summaries of long-form text and evaluate the coherence of these summaries. Official package for our ICLR 2024 paper, "BooookScore: A systematic exploration of book-length summ…
Official Implementation of "DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucination"
A framework for few-shot evaluation of language models.
A collection of awesome-prompt-datasets, awesome-instruction-dataset, to train ChatLLM such as chatgpt 收录各种各样的指令数据集, 用于训练 ChatLLM 模型。
Datasets for Instruction Tuning of Large Language Models
[NAACL'25 Oral] Steering Knowledge Selection Behaviours in LLMs via SAE-Based Representation Engineering
WildEval / ZeroEval
Forked from allenai/WildBenchA simple unified framework for evaluating LLMs
Code for the EMNLP 2024 paper "Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps"
This repository collects all relevant resources about interpretability in LLMs
open-source code for paper: Retrieval Head Mechanistically Explains Long-Context Factuality