-
Northeastern University
- Shengyang
Lists (1)
Sort Name ascending (A-Z)
Stars
Local-first desktop renderer for copied Markdown, LaTeX, Mermaid, and terminal output.
Code for paper "RRC: Unlocking Generative Reward Models in LLM Reinforcement Learning via Ranking-Based Reward Construction"
Self-hosted AI assistant with tool use, multi-agent orchestration, coding copilot and a lightweight Flask + vanilla JS stack.
The offical repo for "LLMs Improving LLMs: Agentic Discovery for Test-Time Scaling"
基于小牛翻译(NiuTrans)API 的 MCP Provider,提供文字翻译工具和语种目录资源,方便在 Cursor/mcp-cli 等客户端中引用。
[ICLR'26] R-HORIZON: How Far Can Your Large Reasoning Model Really Go in Breadth and Depth?
Code for CVPR 2026 paper "MSRL: Scaling Generative Multimodal Reward Modeling via Multi-Stage Reinforcement Learning"
An introduction to ODEs and their applications in vision and language
Advancing Block Diffusion Language Models for Test-Time Scaling
A Diagnostic Guardrail Framework for AI Agent Safety and Security
The official GitHub repo for the survey paper "A Survey on Diffusion Language Models".
Code for AAAI 2026 paper "Probing Preference Representations: A Multi-Dimensional Evaluation and Analysis Method for Reward Models"
Building a inclusive, scalable, and high-performance multilingual translation model
Co-Reinforcement Learning for Unified Multimodal Understanding and Generation
Code for NeurIPS 2025 paper "MRO: Enhancing Reasoning in Diffusion Language Models via Multi-Reward Optimization"
A tool for translating the content of LaTeX documents into various other natural languages (e.g., translating an arXiv paper from English to Chinese).
Multilingual Translations of "Foundations of Large Language Models" and NLPBook.
wangclnlp / GRAM
Forked from NiuTrans/GRAMIn this repository, we provide generative foundation reward models to produce more accurate reward scores, thereby enabling better alignment of LLMs.
BARTScore: Evaluating Generated Text as Text Generation
Code for ICML 2025 paper "GRAM: A Generative Foundation Reward Model for Reward Generalization"
Code for TASLP 2025 paper "Learning Evaluation Models from Large Language Models for Sequence Generation"
A comprehensive book on neural networks and large language models in NLP
Fine-tuning Large Language Diffusion Models
[COLM'25] DPE: Effective Length Extrapolation via Dimension-Wise Positional Embeddings Manipulation
Official Implementation for the paper "d1: Scaling Reasoning in Diffusion Large Language Models via Reinforcement Learning"