Hackcat

What is Context and How Does It Differ from Account Quotas?

Understanding the model's capacity to process material at once versus cumulative usage.

Context is the material to be processed in the current round

Context includes the current question, relevant content from the conversation, as well as text from files and tool results input into the model. Its capacity is limited, so even a short question may approach the limit due to a long prior conversation or large output.

A token is the system's unit for calculating text length and does not equal the number of Chinese characters. The same number of characters in code, English, and Chinese may have different token counts.

Account quota is a different kind of limit

Daily request quotas or monthly usage limits focus on the account's usage over a period of time. Shortening the material can help reduce context pressure but will not immediately restore a used-up monthly quota; waiting for quota recovery also will not automatically shorten an excessively long input.

For example, a brief follow-up question appended with a large amount of historical logs may have its capacity mainly occupied by the logs. Retaining records relevant to the current time period first, then indicating that other parts are omitted, can help the model focus on the problem and also make it easier for you to check whether the analysis covers necessary conditions.

  1. First confirm whether the error message states content is too long or quota is insufficient.
  2. If content is too long, remove irrelevant material or segment by question.
  3. If quota is insufficient, check the account's corresponding recovery prompts.

How to organize materials

Provide the task description first, then the essential fragments needed to complete the task. For long logs, first specify the time range and error keywords; for code, provide related files and call relationships first. Do not repeatedly paste the entire history for “insurance”; this will occupy space for processing new issues and make key points harder to identify.