Get ready for the GARP Risk and AI Exam with flashcards and multiple choice questions. Each question comes with hints and explanations. Prepare for success!

Multiple Choice

What is the maximum number of tokens the LLM can process in a single sequence?

This question tests understanding of the input size limit for an LLM—the context length. Context length is the maximum number of tokens the model can attend to in a single sequence, including both the prompt you provide and any tokens the model generates in response. The decoding methods shown in the options, like Top-K sampling and Top-P (Nucleus) sampling, determine which next token to choose but do not set how many tokens can be processed at once. Statelessness refers to whether the model retains information across separate interactions, which is unrelated to the per-sequence token limit. If your prompt plus the generated content would exceed the context length, you must shorten the prompt or work within a sliding window so everything fits within that fixed limit.

This question tests understanding of the input size limit for an LLM—the context length. Context length is the maximum number of tokens the model can attend to in a single sequence, including both the prompt you provide and any tokens the model generates in response. The decoding methods shown in the options, like Top-K sampling and Top-P (Nucleus) sampling, determine which next token to choose but do not set how many tokens can be processed at once. Statelessness refers to whether the model retains information across separate interactions, which is unrelated to the per-sequence token limit. If your prompt plus the generated content would exceed the context length, you must shorten the prompt or work within a sliding window so everything fits within that fixed limit.