Hi
I'm encountering a "token limit exceeded" error when using GPT-4.1 via Copilot, even though — based on my estimate — the input size is nowhere near the 1 million token limit.
I tested the same input with Gemini 2.5 Pro, which also supports 1 million tokens, and it processed the request without any problems. In both cases, the actual input is far below the stated token limit.
Since I currently have to estimate token usage manually, it would be really helpful if Copilot or the API provided an input identifier or token count feedback so that users can better understand and debug such issues. As a workaround, I’ve been forced to start new chats and reinsert the full context, which is quite inefficient.
Could you please check whether this is a bug or an undocumented limit in how token usage is measured in GPT-4.1? I'm happy to provide sample inputs or more context if helpful.
Thanks in advance!
Best regards,
L.
Hi
I'm encountering a "token limit exceeded" error when using GPT-4.1 via Copilot, even though — based on my estimate — the input size is nowhere near the 1 million token limit.
I tested the same input with Gemini 2.5 Pro, which also supports 1 million tokens, and it processed the request without any problems. In both cases, the actual input is far below the stated token limit.
Since I currently have to estimate token usage manually, it would be really helpful if Copilot or the API provided an input identifier or token count feedback so that users can better understand and debug such issues. As a workaround, I’ve been forced to start new chats and reinsert the full context, which is quite inefficient.
Could you please check whether this is a bug or an undocumented limit in how token usage is measured in GPT-4.1? I'm happy to provide sample inputs or more context if helpful.
Thanks in advance!
Best regards,
L.