-
Tsinghua University
Highlights
- Pro
Stars
Code for "Precise Localization of Memories: A Fine-grained Neuron-level Knowledge Editing Technique for LLMs" (ICLR 2025)
A comprehensive, unified and modular event extraction toolkit.
A toolkit for describing model features and intervening on those features to steer behavior.
Mechanistic Interpretability Visualizations using React
The nnsight package enables interpreting and manipulating the internals of deep learned models.
LLM Transparency Tool (LLM-TT), an open-source interactive toolkit for analyzing internal workings of Transformer-based language models. *Check out demo at* https://huggingface.co/spaces/facebook/l…
Stanford NLP Python library for understanding and improving PyTorch models via interventions
Large-scale, Informative, and Diverse Multi-round Chat Data (and Models)
A guidance language for controlling large language models.
[NeurIPS 2023] MeZO: Fine-Tuning Language Models with Just Forward Passes. https://arxiv.org/abs/2305.17333
GPT4All: Run Local LLMs on Any Device. Open-source and available for commercial use.
Code and documentation to train Stanford's Alpaca models, and generate the data.
arXiv LaTeX Cleaner: Easily clean the LaTeX code of your paper to submit to arXiv
GLM-130B: An Open Bilingual Pre-Trained Model (ICLR 2023)
MLNLP社区用来帮助大家避免论文投稿小错误的整理仓库。 Paper Writing Tips
Acceptance rates for the major AI conferences
Source code and dataset for EMNLP 2020 paper "MAVEN: A Massive General Domain Event Detection Dataset".
✨Fast Coreference Resolution in spaCy with Neural Networks
On Transferability of Prompt Tuning for Natural Language Processing
程序员延寿指南 | A programmer's guide to live longer
Efficient Training (including pre-training and fine-tuning) for Big Models
A plug-and-play library for parameter-efficient-tuning (Delta Tuning)
Code and datasets for the paper "Humor Detection: A Transformer Gets the Last Laugh"
Official repository for CMU Machine Learning Department's 10717: "The Art of the Paper".