Skip to content
@horizon-llm

Horizon

Focus on what matters over a long time horizon.

Pinned Loading

  1. OpenKimi OpenKimi Public

    [ICML2026] Reproduce Kimi K1.5/K2 RL algorithm and rollout system

    Python 21 2

  2. Think-RM Think-RM Public

    [NeurIPS 2025] Think-RM: Enabling Long-Horizon Reasoning in Generative Reward Models

    Python 17 1

  3. uncertainty-router uncertainty-router Public

    [NeurIPS 2025] Ask a Strong LLM Judge when Your Reward Model is Uncertain

    Python 11

  4. HeaPA HeaPA Public

    [COLM2026] Difficulty-Aware Heap Sampling and On-Policy Query Augmentation for LLM Reinforcement Learning

    Python 7

  5. RESD RESD Public

    [arXiv 2026] Learning from Rare Success and Rich Feedback via Reflection-Enhanced Self-Distillation

    Python 19 1

  6. CTPO CTPO Public

    Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

    Python 4

Repositories

Showing 10 of 13 repositories
  • IOI2026 Public
    horizon-llm/IOI2026's past year of commit activity
    C++ 2 CC0-1.0 0 0 0 Updated Aug 25, 2026
  • horizon-llm/horizon-llm.github.io's past year of commit activity
    0 0 0 0 Updated Aug 17, 2026
  • CoMem Public

    [ICML 2026] CoMem: Context Management with A Decoupled Long-Context Model

    horizon-llm/CoMem's past year of commit activity
    Python 4 Apache-2.0 0 0 0 Updated Aug 15, 2026
  • AlphaQuanter Public

    [ACL2026] AlphaQuanter: An End-to-End Tool-Orchestrated Agentic Reinforcement Learning Framework for Stock Trading.

    horizon-llm/AlphaQuanter's past year of commit activity
    Python 75 12 2 0 Updated Jul 3, 2026
  • CTPO Public

    Rethinking Importance Sampling in LLM Policy Optimization: A Cumulative Token Perspective

    horizon-llm/CTPO's past year of commit activity
    Python 4 Apache-2.0 0 1 0 Updated Jun 24, 2026
  • RESD Public

    [arXiv 2026] Learning from Rare Success and Rich Feedback via Reflection-Enhanced Self-Distillation

    Python 19 Apache-2.0 1 0 5 Updated May 21, 2026
  • adaptive-rollout Public

    Implementation for paper Improving Sampling Efficiency in RLVR through Adaptive Rollout and Response Reuse

    horizon-llm/adaptive-rollout's past year of commit activity
    Python 0 Apache-2.0 0 0 0 Updated Apr 17, 2026
  • horizon-llm/automated-w2s-research's past year of commit activity
    Python 0 48 0 0 Updated Apr 13, 2026
  • OpenKimi Public

    [ICML2026] Reproduce Kimi K1.5/K2 RL algorithm and rollout system

    horizon-llm/OpenKimi's past year of commit activity
    Python 21 Apache-2.0 2 1 0 Updated Apr 9, 2026
  • ToolOrchestrationReward Public

    Multi-Step Tool Orchestration with Constrained Data Synthesis and Graduated Rewards

    horizon-llm/ToolOrchestrationReward's past year of commit activity
    Python 1 Apache-2.0 0 1 0 Updated Apr 7, 2026

People

This organization has no public members. You must be a member to see who’s a part of this organization.

Top languages

Loading…

Most used topics

Loading…