The decoder modifies logits for previously seen tokens before choosing the next one. Strong penalties can reduce loops but may also damage legitimate repetition and factual wording.
A repetition penalty adjusts token scores to discourage the model from reusing tokens or phrases it has already generated.
The decoder modifies logits for previously seen tokens before choosing the next one. Strong penalties can reduce loops but may also damage legitimate repetition and factual wording.