开始使用免费开始使用

Quiz 4 - Question 1

Imagine that you have built a tiny language model with a vocabulary of 5 tokens. This model predicted the following probability distribution over the next token:

the: 0.05
chased: 0.01
a: 0.04
lion: 0.44
zebra: 0.46

If you apply top-k sampling with k=3, what is the modified probability distribution from which the model samples the next token?

本练习是课程的一部分

Google DeepMind: Discover The Transformer Architecture

查看课程

动手互动练习

通过我们的互动练习之一,将理论转化为实践

开始练习