It appears the context window setting is not being respected; based on my observations, responses to prompts containing longer text snippets do not seem to consider the entirety of that text content regardless of the size of the context window configured in the settings. Only the Gemma 3 models exhibit behavior suggesting they are aware of a limited context window ("Completion failed: Context is full"), implying other models are effectively ignoring the specified parameter, which significantly impacts usability for tasks requiring understanding of longer documents.
It appears the context window setting is not being respected; based on my observations, responses to prompts containing longer text snippets do not seem to consider the entirety of that text content regardless of the size of the context window configured in the settings. Only the Gemma 3 models exhibit behavior suggesting they are aware of a limited context window ("Completion failed: Context is full"), implying other models are effectively ignoring the specified parameter, which significantly impacts usability for tasks requiring understanding of longer documents.