-
Nanyang Technological Univeristy
- Singapore
- https://zongliny.github.io
Stars
Qwen3-Coder is the code version of Qwen3, the large language model series developed by Qwen team.
ParaVT: Taming the Tool Prior Paradox for Parallel Tool Use in Agentic Video Reinforcement Learning
[AAAI 2026] Benchmarking Language Model Agents in Algorithm Search for Combinatorial Optimization
Now, Stronger AI Pushes Frontiers, Stronger Our Shared Future.
A clean implementation based on AlphaZero for any game in any framework + tutorial + Othello/Gobang/TicTacToe/Connect4 and more
Your own personal AI assistant. Any OS. Any Platform. The lobster way. 🦞
[ICML 2026] <MOOSE-Star: Unlocking Tractable Training for Scientific Discovery by Breaking the Complexity Barrier>
DeepResearchEval: An Automated Framework for Deep Research Task Construction and Agentic Evaluation.
This repository contains a summary of knowledge cut-off dates for various large language models (LLMs), such as GPT, Claude, Gemini, Llama, and more.
Monitor Google Scholar author citation counts and track changes automatically without opening tabs.
MLGym A New Framework and Benchmark for Advancing AI Research Agents
The official github repo for "Diffusion Language Models are Super Data Learners".
Thesis Latex Template for Nanyang Technological University (NTU)
MiroMind-M1 is a fully open-source series of reasoning language models built on Qwen-2.5, focused on advancing mathematical reasoning.
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
PyTorch code and models for VJEPA2 self-supervised learning from video.
[NeurIPS 2025] <MOOSE-Chem2: Exploring LLM Limits in Fine-Grained Scientific Hypothesis Discovery via Hierarchical Search>
[ACL 2026] <ResearchBench: Benchmarking LLMs in Scientific Discovery via Inspiration-Based Task Decomposition>
The RL Bridge for LLM-based Agent Applications. Made Simple & Flexible.
EasyR1: An Efficient, Scalable, Multi-Modality RL Training Framework based on veRL
[ICLR 2025] <MOOSE-Chem: Large Language Models for Rediscovering Unseen Chemistry Scientific Hypotheses>
Model Context Protocol Servers
DeepSeekMath: Pushing the Limits of Mathematical Reasoning in Open Language Models
Recent research papers about Foundation Models for Combinatorial Optimization
A high-throughput and memory-efficient inference and serving engine for LLMs
Must-read Papers on Large Language Model (LLM) as Optimizers and Automatic Optimization for Prompting LLMs.