Generating output one token at a time, each conditioned on all previous tokens, the basis of LLM decoding.
Related terms: LLM
← All terms