Reinforcement Learning from Human Feedback (RLHF)
Mina Parham
AI Engineer



prompt_data = load_dataset("center-for-humans-and-machines/rlhf-hackathon-prompts",
split="train")
prompt_data['prompt'][0]
'Tầm quan trọng của biến đổi khí hậu là gì?'
Input=, {{Text}}:, ###Human:from datasets import load_dataset
preference_data = load_dataset("trl-internal-testing/hh-rlhf-helpful-base-trl-style",
split="train")

def extract_prompt(text):
# Extract the prompt as the first element in the list
prompt = text[0]["content"]
return prompt
# Apply the extraction function to the dataset
preference_data_with_prompt = preference_data.map(
lambda example: {**example, 'prompt': extract_prompt(example['chosen'])}
)
sample = preference_data_with_prompt.select(range(1))
sample['prompt']
'Những vitamin nào thiết yếu cho cơ thể hoạt động?'
sample['chosen']
[ { "content": "Những vitamin nào thiết yếu cho cơ thể hoạt động?", "role":
"user" }, { "content": "Có một số vitamin rất quan trọng đảm bảo cơ thể hoạt
động bình thường, gồm vitamin A, C, D, E và K cùng ...}]
Reinforcement Learning from Human Feedback (RLHF)