0 XP

Output logits · What comes out

The other end

A model doesn’t output a word

Text goes in as tokens — the chunks a model reads. What comes back isn’t a word, or even a token. It’s one raw score for every single token in the vocabulary, all two hundred thousand of them, saying how well each would fit next. Those scores are called logits.

The starting scores are illustrative, but the softmax, sampling, loss and perplexity are the real formulas, computed live. softmax