Skip to content
View hank0316's full-sized avatar

Block or report hank0316

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. AdaSearch AdaSearch Public

    This includes the original implementation of "AdaSearch: Balancing Parametric Knowledge and Search in Large Language Models via Reinforcement Learning".

    10

  2. MiuLab/DogeRM MiuLab/DogeRM Public

    The code used in the paper "DogeRM: Equipping Reward Models with Domain Knowledge through Model Merging"

    Python 6

  3. s3prl/s3prl s3prl/s3prl Public

    Self-Supervised Speech Pre-training and Representation Learning Toolkit

    Python 2.6k 534

  4. virginiakm1988/ML2022-Spring virginiakm1988/ML2022-Spring Public

    **Official** 李宏毅 (Hung-yi Lee) 機器學習 Machine Learning 2022 Spring

    Jupyter Notebook 2.6k 540

  5. allenai/reward-bench allenai/reward-bench Public

    RewardBench: the first evaluation tool for reward models.

    Python 732 98

  6. NovaSky-AI/SkyRL NovaSky-AI/SkyRL Public

    SkyRL: A Modular Full-stack RL Library for LLMs

    Python 2.2k 403