Popular repositories Loading
-
robust-loss-mlml
robust-loss-mlml PublicForked from xinyu1205/robust-loss-mlml
Code for paper: Simple and Robust Loss Design for Multi-Label Learning with Missing Labels
Python
-
trlx
trlx PublicForked from CarperAI/trlx
A repo for distributed training of language models with Reinforcement Learning via Human Feedback (RLHF)
Python
-
ColossalAI
ColossalAI PublicForked from hpcaitech/ColossalAI
Making large AI models cheaper, faster and more accessible
Python
-
PaLM-rlhf-pytorch
PaLM-rlhf-pytorch PublicForked from lucidrains/PaLM-rlhf-pytorch
Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the PaLM architecture. Basically ChatGPT but with PaLM
Python
-
DeepSpeedExamples
DeepSpeedExamples PublicForked from deepspeedai/DeepSpeedExamples
Example models using DeepSpeed
Python
-
awesome-RLHF
awesome-RLHF PublicForked from opendilab/awesome-RLHF
A curated list of reinforcement learning with human feedback resources (continually updated)
If the problem persists, check the GitHub status page or contact support.