Skip to content
View aryopg's full-sized avatar

Highlights

  • Pro

Organizations

@EdinburghClinicalNLP

Block or report aryopg

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Item response theory with Python

Python 15 2 Updated May 5, 2026

A Unified Benchmark For Fidelity, Privacy and Utility of Synthetic Chest Radiographs

Python 76 2 Updated Jun 12, 2026

The official repo of the paper "MMLongBench Benchmarking Long-Context Vision-Language Models Effectively and Thoroughly"

Python 175 10 Updated Jul 7, 2026

Inference API for many LLMs and other useful tools for empirical research

Python 134 42 Updated May 29, 2026

Code release for "Debating with More Persuasive LLMs Leads to More Truthful Answers"

Python 131 27 Updated Mar 22, 2024

MPS supported fork of "Large Language Diffusion Models"

Python 2 Updated Feb 27, 2025

Open source replication of Anthropic's Crosscoders for Model Diffing

Python 68 23 Updated Oct 27, 2024

s1: Simple test-time scaling

Python 6,664 757 Updated Jun 25, 2025

Fully open reproduction of DeepSeek-R1

Python 26,417 2,445 Updated Apr 2, 2026

Agent Laboratory is an end-to-end autonomous research workflow meant to assist you as the human researcher toward implementing your research ideas

Python 5,784 804 Updated Aug 20, 2025

A 100x faster SVD for PyTorch⚡️

C++ 492 39 Updated Oct 10, 2022

Large Concept Models: Language modeling in a sentence representation space

Python 2,368 212 Updated Jan 29, 2025

This repository contains the NarrativeQA dataset. It includes the list of documents with Wikipedia summaries, links to full stories, and questions and answers.

Shell 518 70 Updated Apr 15, 2020

Stanford NLP Python library for Representation Finetuning (ReFT)

Python 1,576 134 Updated Mar 5, 2026

Learning Binding Affinities via Fine-tuning of Protein and Ligand Language Models

Jupyter Notebook 40 6 Updated Mar 30, 2026

A package to generate summaries of long-form text and evaluate the coherence of these summaries. Official package for our ICLR 2024 paper, "BooookScore: A systematic exploration of book-length summ…

Python 130 11 Updated Oct 1, 2024

Official Implementation of "DeCoRe: Decoding by Contrasting Retrieval Heads to Mitigate Hallucination"

Jupyter Notebook 30 3 Updated Dec 18, 2024

A framework for few-shot evaluation of language models.

Python 13,504 3,456 Updated Jul 13, 2026
Python 10 1 Updated Dec 9, 2024

A collection of awesome-prompt-datasets, awesome-instruction-dataset, to train ChatLLM such as chatgpt 收录各种各样的指令数据集, 用于训练 ChatLLM 模型。

738 42 Updated Jun 17, 2026

Datasets for Instruction Tuning of Large Language Models

261 13 Updated Nov 30, 2023

A cat(1) clone with wings.

Rust 59,997 1,614 Updated Aug 1, 2026

[NAACL'25 Oral] Steering Knowledge Selection Behaviours in LLMs via SAE-Based Representation Engineering

Python 83 10 Updated Jun 20, 2026

A simple unified framework for evaluating LLMs

HTML 272 31 Updated Apr 14, 2025

Code for the EMNLP 2024 paper "Detecting and Mitigating Contextual Hallucinations in Large Language Models Using Only Attention Maps"

Python 152 12 Updated Oct 13, 2025
Python 498 71 Updated Jul 30, 2026

This repository collects all relevant resources about interpretability in LLMs

402 27 Updated Nov 1, 2024

open-source code for paper: Retrieval Head Mechanistically Explains Long-Context Factuality

Python 242 27 Updated Aug 2, 2024
Next