Stars
energycomp56 / peft
Forked from huggingface/peft🤗 PEFT: State-of-the-art Parameter-Efficient Fine-Tuning.
This repo contains the syllabus of the Hugging Face Deep Reinforcement Learning Course.
energycomp56 / RL4LMs
Forked from allenai/RL4LMsA modular RL library to fine-tune language models to human preferences
Code and documentation to train Stanford's Alpaca models, and generate the data.
energycomp56 / alpaca.cpp
Forked from antimatter15/alpaca.cppLocally run an Instruction-Tuned Chat-Style LLM
antimatter15 / alpaca.cpp
Forked from ggml-org/llama.cppLocally run an Instruction-Tuned Chat-Style LLM
energycomp56 / transformers
Forked from huggingface/transformers🤗 Transformers: State-of-the-art Machine Learning for Pytorch, TensorFlow, and JAX.