🫠
OK, large language models.
-
Tencent
- Beijing, China
-
08:23
(UTC +08:00)
Pinned Loading
-
Towards-A-Deep-Understanding-of-Multilingual-E2E-ST
Towards-A-Deep-Understanding-of-Multilingual-E2E-ST PublicTowards a Deep Understanding of Multilingual End-to-End Speech Translation, Findings of EMNLP 2023
Python 5
-
verl-project/verl
verl-project/verl Publicverl/HybridFlow: A Flexible and Efficient RL Post-Training Framework
-
CD-RLHF
CD-RLHF PublicForked from ernie-research/CD-RLHF
[ACL'25] Official code of curiosity-driven RLHF
Python
-
MA-RLHF
MA-RLHF PublicForked from ernie-research/MA-RLHF
[ICLR'25] MA-RLHF: Reinforcement Learning from Human Feedback with Macro Actions
Python
-
tjunlp-lab/FuxiTranyu
tjunlp-lab/FuxiTranyu PublicFuxiTranyu-8B is an open-source multilingual large language model trained from scratch with 606B tokens.
Something went wrong, please refresh the page to try again.
If the problem persists, check the GitHub status page or contact support.
If the problem persists, check the GitHub status page or contact support.