开始使用免费开始使用

AI Safety Question 3

You want to embed the concept of safety into a pre-trained LLM by fine-tuning it. You have a dataset of prompts and pairs of possible answers, along with labels created by human safety evaluators indicating their preference for one answer over the other. Which technique is the most suitable for fine-tuning the LLM in this scenario?

本练习是课程的一部分

Responsible AI for Developers: Privacy & Safety

查看课程

动手互动练习

通过我们的互动练习之一,将理论转化为实践

开始练习