การประเมินด้วย METEOR
METEOR มีจุดเด่นในการประเมินคุณลักษณะเชิงความหมายของข้อความ โดยทำงานคล้ายกับ ROUGE ด้วยการเปรียบเทียบผลลัพธ์ที่โมเดลสร้างขึ้นกับผลลัพธ์อ้างอิง ข้อความทั้งสองถูกเตรียมไว้ให้แล้วในตัวแปร generated และ reference — ลองประเมินคะแนนได้เลย
ไลบรารี evaluate ถูกโหลดไว้ให้แล้ว
แบบฝึกหัดนี้เป็นส่วนหนึ่งของหลักสูตร
Python เบื้องต้นสำหรับ LLMs
คำแนะนำการฝึกหัด
- คำนวณและพิมพ์คะแนน METEOR
แบบฝึกหัดเชิงโต้ตอบแบบลงมือทำ
ลองทำแบบฝึกหัดนี้โดยเติมโค้ดตัวอย่างนี้ให้สมบูรณ์
meteor = evaluate.load("meteor")
generated = ["The burrow stretched forward like a narrow corridor for a while, then plunged abruptly downward, so quickly that Alice had no chance to stop herself before she was tumbling into an extremely deep shaft."]
reference = ["The rabbit-hole went straight on like a tunnel for some way, and then dipped suddenly down, so suddenly that Alice had not a moment to think about stopping herself before she found herself falling down a very deep well."]
# Compute and print the METEOR score
results = ____
print("Meteor: ", ____)