-
University of Edinburgh
- Edinburgh, Scotland
- alesuglia.github.io
- @ale_suglia
Highlights
- Pro
Stars
Can LLM agents play Keep Talking and Nobody Explodes in real-time?
A simple tool to update bib entries with their official information (e.g., DBLP or the ACL anthology).
[ICLR 2026] The offical Implementation of "Soft-Prompted Transformer as Scalable Cross-Embodiment Vision-Language-Action Model"
Official code repository for the paper "Learning Massively Multitask World Models for Continuous Control".
Code and implementations for the paper "AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning" by Zhiheng Xi et al.
All you need to get started with the LM Playpen Environment for Learning in Interaction.
Implementation of π₀, the robotic foundation model architecture proposed by Physical Intelligence
[ICLR 2025] Source code for paper "A Spark of Vision-Language Intelligence: 2-Dimensional Autoregressive Transformer for Efficient Finegrained Image Generation"
Repository for the EMNLP'24 paper "Repairs in a Block World: A New Benchmark for Handling User Corrections with Multi-Modal Language Models". Chiyah-Garcia et al. https://arxiv.org/abs/2409.14247
A dummy's guide to setting up (and using) HPC clusters on Ubuntu 22.04LTS using Slurm and Munge. Created by the Quant Club @ UIowa.
Educational framework exploring ergonomic, lightweight multi-agent orchestration. Managed by OpenAI Solution team.
a python framework to build, learn and reason about probabilistic circuits and tensor networks
SPOC: Imitating Shortest Paths in Simulation Enables Effective Navigation and Manipulation in the Real World
A comprehensive framework to explore whether embodied multimodal models are plausibly resilient
Manage scalable open LLM inference endpoints in Slurm clusters
Gemma open-weight LLM library, from Google DeepMind
Non-cooperative satellite operations challenge problems implemented in the Kerbal Space Program game engine
A playbook for systematically maximizing the performance of deep learning models.
Code for "Learning to Model the World with Language." ICML 2024 Oral.
🌸 Run LLMs at home, BitTorrent-style. Fine-tuning and inference up to 10x faster than offloading
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)