Skip to content
View wangclnlp's full-sized avatar
  • Northeastern University
  • Shengyang

Block or report wangclnlp

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Pinned Loading

  1. NiuTrans/Vision-LLM-Alignment NiuTrans/Vision-LLM-Alignment Public

    This repository contains the code for SFT, RLHF, and DPO, designed for vision-based LLMs, including the LLaVA models and the LLaMA-3.2-vision models.

    Python 122 9

  2. NiuTrans/GRAM NiuTrans/GRAM Public

    Code for ICML 2025 paper "GRAM: A Generative Foundation Reward Model for Reward Generalization"

    Python 22 2

  3. MRO MRO Public

    Code for NeurIPS 2025 paper "MRO: Enhancing Reasoning in Diffusion Language Models via Multi-Reward Optimization"

    Python 6

  4. MRMBench MRMBench Public

    Code for AAAI 2026 paper "Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models"

    Python 2

  5. MSRL MSRL Public

    Code for CVPR 2026 paper "MSRL: Scaling Generative Multimodal Reward Modeling via Multi-Stage Reinforcement Learning"

    Python 12 1

  6. RRC RRC Public

    Code for paper "RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction"

    Python 3