Model Evaluation for Generative AI Question 3
During the evaluation of an LLM, you find that the model often produces responses that sound fluent and grammatical but are factually incorrect. Which of the following evaluation challenges does this example illustrate?
이 연습은 강의의 일부입니다
