Stars
verl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
MM-Eureka V0 also called R1-Multimodal-Journey, Latest version is in MM-Eureka
MM-EUREKA: Exploring the Frontiers of Multimodal Reasoning with Rule-based Reinforcement Learning
Awesome things about generative recommendation models.
Search-R1: An Efficient, Scalable RL Training Framework for Reasoning & Search Engine Calling interleaved LLM based on veRL
A unified, comprehensive and efficient recommendation library
[NIPS2023]Implementation of Foundation Model is Efficient Multimodal Multitask Model Selector
Codes for "Template-free Prompt Tuning for Few-shot NER".