Home Projects LLM-RLHF-Tuning
LLM-RLHF-Tuning

LLM-RLHF-Tuning

by Joyce94 · GitHub

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

# fine-tuning# language-model# llama# llm
View on GitHub
⭐ Stars
452
🍴 Forks
24
📜 License
License not specified
📅 Created
2023
🔄 Last commit
2 yr ago
🏷️ Category
fine-tuning
💻 Language
🖥️ Self-hostable
Likely
You maintain this project?

Claim its page: indexed whatever its rank, translated into six languages, and enriched with what you write yourself.

Claim this page →
LLM-RLHF-Tuning — GitHub preview card
📈 Star history
453452
2026-07-202026-07-21
📄 About

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

Frequently asked questions

What is LLM-RLHF-Tuning?

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

Is LLM-RLHF-Tuning open source?

LLM-RLHF-Tuning is an open-source project.

Is LLM-RLHF-Tuning free?

Yes. LLM-RLHF-Tuning is free and open source — you can use, modify and self-host it.

🏅 Maintainer of this project?
OpenSourceAI badge — LLM-RLHF-Tuning

Add this live badge to your README — your GitHub stars and directory rank, refreshed daily.

[![OpenSourceAI](https://opensourceai.tech/badge.php?tool=joyce94-llm-rlhf-tuning)](https://opensourceai.tech/project/joyce94-llm-rlhf-tuning.html)
More badge options →
🧬 Related projects🧬 View the DNA map →