Rlhf Open Source Projects

Browse 41 Rlhf open source projects, ranked by GitHub stars. Find the most popular Rlhf tools and libraries.

Share your experience:✍️ Write a Post❓ Ask a Question
73,376 stars

hiyouga/LLaMA-Factory

Unified Efficient Fine-Tuning of 100+ LLMs & VLMs (ACL 2024)

Metrics details
Stars73,376
37,378 stars

LAION-AI/Open-Assistant

OpenAssistant is a chat-based assistant that understands tasks, can interact with third-party systems, and retrieve information dynamically to do so.

Metrics details
Stars37,378
12,190 stars

RUCAIBox/LLMSurvey

The official GitHub page for the survey paper "A Survey of Large Language Models".

Metrics details
Stars12,190
9,788 stars

OpenLLMAI/OpenRLHF

A Ray-based High-performance RLHF framework (Support 70B+ full tuning & LoRA & Mixtral)

Metrics details
Stars9,788
7,244 stars

InternLM/InternLM

Official release of InternLM series (InternLM, InternLM2, InternLM2.5, InternLM3).

Metrics details
Stars7,244
7,132 stars

ymcui/Chinese-LLaMA-Alpaca-2

中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)

Metrics details
Stars7,132
5,638 stars

huggingface/alignment-handbook

Robust recipes to align language models with human and AI preferences

Metrics details
Stars5,638
5,039 stars

argilla-io/argilla

Argilla is a collaboration tool for AI engineers and domain experts to build high-quality datasets

Metrics details
Stars5,039
4,413 stars

opendilab/awesome-RLHF

A curated list of reinforcement learning with human feedback resources (continually updated)

Metrics details
Stars4,413
3,719 stars

hiyouga/ChatGLM-Efficient-Tuning

Fine-tuning ChatGLM-6B with PEFT | 基于 PEFT 的高效 ChatGLM 微调

Metrics details
Stars3,719
3,481 stars

Docta-ai/docta

A Doctor for your data

Metrics details
Stars3,481
3,334 stars

argilla-io/distilabel

Distilabel is a framework for synthetic data and AI feedback for engineers who need fast, reliable and scalable pipelines based on verified research papers.

Metrics details
Stars3,334
2,004 stars

tatsu-lab/alpaca_eval

An automatic evaluator for instruction-following language models. Human-validated, high-quality, cheap, and fast.

Metrics details
Stars2,004
1,693 stars

THUDM/ImageReward

[NeurIPS 2023] ImageReward: Learning and Evaluating Human Preferences for Text-to-image Generation

Metrics details
Stars1,693
1,611 stars

PKU-Alignment/safe-rlhf

Safe RLHF: Constrained Value Alignment via Safe Reinforcement Learning from Human Feedback

Metrics details
Stars1,611
1,602 stars

THUDM/WebGLM

WebGLM: An Efficient Web-enhanced Question Answering System (KDD 2023)

Metrics details
Stars1,602
1,222 stars

xtreme1-io/xtreme1

Xtreme1 is an all-in-one data labeling and annotation platform for multimodal data training and supports 3D LiDAR point cloud, image, and LLM.

Metrics details
Stars1,222
908 stars

ContextualAI/HALOs

A library with extensible implementations of DPO, KTO, PPO, ORPO, and other human-aware loss functions (HALOs).

Metrics details
Stars908
742 stars

GaryYufei/AlignLLMHumanSurvey

Aligning Large Language Models with Human: A Survey

Metrics details
Stars742
659 stars

jerry1993-tech/Cornucopia-LLaMA-Fin-Chinese

聚宝盆(Cornucopia): 中文金融系列开源可商用大模型,并提供一套高效轻量化的垂直领域LLM训练框架(Pretraining、SFT、RLHF、Quantize等)

Metrics details
Stars659
564 stars

voidful/TextRL

Implementation of ChatGPT RLHF (Reinforcement Learning with Human Feedback) on any generation model in huggingface's transformer (blommz-176B/bloom/gpt/bart/T5/MetaICL)

Metrics details
Stars564
481 stars

mindspore-courses/step_into_llm

MindSpore online courses: Step into LLM

Metrics details
Stars481
452 stars

Joyce94/LLM-RLHF-Tuning

LLM Tuning with PEFT (SFT+RM+PPO+DPO with LoRA)

Metrics details
Stars452
412 stars

CambioML/pykoi

pykoi: Active learning in one unified interface

Metrics details
Stars412
391 stars

glgh/awesome-llm-human-preference-datasets

A curated list of Human Preference Datasets for LLM fine-tuning, RLHF, and eval.

Metrics details
Stars391
339 stars

WangRongsheng/MedQA-ChatGLM

🛰️ 基于真实医疗对话数据在ChatGLM上进行LoRA、P-Tuning V2、Freeze、RLHF等微调,我们的眼光不止于医疗问答

Metrics details
Stars339
305 stars

HMUNACHI/jax-models

Explore implementations of deep learning concepts like Transformers, Attention, Llama, GPT, InstructGPT, RLHF, Gaussian Processes, Bayesian Inference, Newton Raphson, Distributed Trainers and more!

Metrics details
Stars305
276 stars

jianzhnie/open-chatgpt

The open source implementation of ChatGPT, Alpaca, Vicuna and RLHF Pipeline. 从0开始实现一个ChatGPT.

Metrics details
Stars276
240 stars

jasonvanf/llama-trl

LLaMA-TRL: Fine-tuning LLaMA with PPO and LoRA

Metrics details
Stars240
228 stars

lhao499/chain-of-hindsight

Chain-of-Hindsight, a simpler and more effective alternative to RLHF

Metrics details
Stars228
220 stars

jackaduma/Vicuna-LoRA-RLHF-PyTorch

A full pipeline to finetune Vicuna LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the Vicuna architecture. Basically ChatGPT but with Vicuna

Metrics details
Stars220
205 stars

mengdi-li/awesome-RLAIF

A continually updated list of literature on Reinforcement Learning from AI Feedback (RLAIF)

Metrics details
Stars205
202 stars

liziniu/ReMax

Code for Paper (ReMax: A Simple, Efficient and Effective Reinforcement Learning Method for Aligning Large Language Models)

Metrics details
Stars202
196 stars

Miraclemarvel55/ChatGLM-RLHF

对ChatGLM直接使用RLHF提升或降低目标输出概率|Modify ChatGLM output with only RLHF

Metrics details
Stars196
182 stars

tomekkorbak/pretraining-with-human-feedback

Code accompanying the paper Pretraining Language Models with Human Preferences

Metrics details
Stars182
182 stars

PKU-Alignment/beavertails

BeaverTails is a collection of datasets designed to facilitate research on safety alignment in large language models (LLMs).

Metrics details
Stars182
172 stars

xrsrke/instructGOOSE

Implementation of Reinforcement Learning from Human Feedback (RLHF)

Metrics details
Stars172
168 stars

csmile-1006/PreferenceTransformer

Preference Transformer: Modeling Human Preferences using Transformers for RL (ICLR2023 Accepted)

Metrics details
Stars168
138 stars

jackaduma/ChatGLM-LoRA-RLHF-PyTorch

A full pipeline to finetune ChatGLM LLM with LoRA and RLHF on consumer hardware. Implementation of RLHF (Reinforcement Learning with Human Feedback) on top of the ChatGLM architecture. Basically ChatGPT but with ChatGLM

Metrics details
Stars138
119 stars

opening-up-chatgpt/opening-up-chatgpt.github.io

Tracking instruction-tuned LLM openness. Paper: Liesenfeld, Andreas, Alianda Lopez, and Mark Dingemanse. 2023. “Opening up ChatGPT: Tracking Openness, Transparency, and Accountability in Instruction-Tuned Text Generators.” In Proceedings of the 5th International Conference on Conversational User Interfaces. doi:10.1145/3571884.3604316.

Metrics details
Stars119
118 stars

l294265421/alpaca-rlhf

Finetuning LLaMA with RLHF (Reinforcement Learning with Human Feedback) based on DeepSpeed Chat

Metrics details
Stars118
Get A Weekly Email With Trending Rlhf Projects
Stay updated on Rlhf plus related topics you pick below.

Copyright 2018-2026 Awesome Open Source.  All rights reserved.