test on HH-RLHF #3

LoverLost · 2024-03-18T17:06:58Z

I see the code and find that in the HH-RLHF dataset you use the red-team data for test. I want to know how the test scores are calculated? I didnt find ground-truth in the red-team dataset. How are the scores for harmless and helpful calculated in the paper?

hongyanz · 2024-03-18T18:39:19Z

We use GPT-4's evaluation as the ground-truth. We also show that GPT-4 and human share similar evaluation results in the paper.

shanpoyang654 · 2024-04-08T13:08:24Z

We use GPT-4's evaluation as the ground-truth. We also show that GPT-4 and human share similar evaluation results in the paper.

I got an output file named res_0.json which contains outputs of LLM. Do I need to put the outputs into GPT4 API to get the evaluation as the groundtruth? It means that there isn't an evaluation process in the code now, right？
Thank you for your code and effort and hope for your reply!

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

test on HH-RLHF #3

test on HH-RLHF #3

LoverLost commented Mar 18, 2024

hongyanz commented Mar 18, 2024

shanpoyang654 commented Apr 8, 2024

test on HH-RLHF #3

test on HH-RLHF #3

Comments

LoverLost commented Mar 18, 2024

hongyanz commented Mar 18, 2024

shanpoyang654 commented Apr 8, 2024