-
Mila & UdeM
- Montreal, Canada
- https://ez-hwh.github.io/
- @EZ_hwh
Highlights
- Pro
Lists (5)
Sort Name ascending (A-Z)
Stars
A curated list of research and projects on world models
Benchmark AI Agents on Enterprise Workflows
Open source code for ICLR 2026 Paper: Evaluating Memory in LLM Agents via Incremental Multi-Turn Interactions
Homepage for ProLong (Princeton long-context language models) and paper "How to Train Long-Context Language Models (Effectively)"
An agent-managed museum exhibit, built in Rust with Gajae-Code / LazyCodex — developed and maintained with no human intervention.
Memory Sparse Attention - A scalable, end-to-end trainable latent-memory framework for 100M-token contexts.
AI agents running research on single-GPU nanochat training automatically
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
The official implementation of the ICML 2024 paper "MemoryLLM: Towards Self-Updatable Large Language Models" and "M+: Extending MemoryLLM with Scalable Long-Term Memory"
H-Net: Hierarchical Network with Dynamic Chunking
PyTorch DeepSeek Sparse Attention (DSA) training & inference
Conditional Memory via Scalable Lookup: A New Axis of Sparsity for Large Language Models
[ICML 2026 Spotlight] Latent Collaboration in Multi-Agent Systems
Recursive Language Models for efficient long-context processing. Analyze 1M+ tokens by storing context in a Python REPL while reducing LLM token usage.
Unofficial implementation of Titans, SOTA memory for transformers, in Pytorch
Drop in a screenshot and convert it to clean code (HTML/Tailwind/React/Vue)
A Foundational Vision-Language-Action Model for Synergizing Cognition and Action in Robotic Manipulation
VLA-Adapter: An Effective Paradigm for Tiny-Scale Vision-Language-Action Model
[CVPR 2025] Open-source, End-to-end, Vision-Language-Action model for GUI Agent & Computer Use.
Code for MetaMorph Multimodal Understanding and Generation via Instruction Tuning
Repository for Meta Chameleon, a mixed-modal early-fusion foundation model from FAIR.
[ICLR 2026] Agentic Reinforced Policy Optimization (ARPO)
[ICLR 2026] OmniWorld: A Multi-Domain and Multi-Modal Dataset for 4D World Modeling
Open diffusion language model for code generation — releasing pretraining, evaluation, inference, and checkpoints.
The development and future prospects of large multimodal reasoning models.
gpt-oss-120b and gpt-oss-20b are two open-weight language models by OpenAI
Agent Reinforcement Trainer: train multi-step agents for real-world tasks using GRPO. Give your agents on-the-job training. Reinforcement learning for Qwen3.6, GPT-OSS, Llama, and more!
Tongyi Deep Research, the Leading Open-source Deep Research Agent