Context window / context length

Appears in 1 tutorial

The maximum number of tokens (prompt + output) a model can handle at once, e.g.

As used in LLM Infrastructure →

The maximum number of tokens (prompt + output) a model can handle at once, e.g. "128K context."