LanguageModel: measureContextUsage() method
Limited availability
This feature is not Baseline because it does not work in some of the most widely-used browsers.
Secure context: This feature is available only in secure contexts (HTTPS), in some or all supporting browsers.
The measureContextUsage() method of the LanguageModel interface estimates how many context window tokens the given input would consume without sending it to the model or modifying the session's state.
This allows you to check how much of the context window a given input requires before deciding whether to send it. The result can be compared against LanguageModel.contextWindow and LanguageModel.contextUsage to determine whether the input can fit into the context window limit.
This is particularly useful for long-context applications such as document summarization, where you may need to split or truncate content to stay within the context window limit.
Syntax
measureContextUsage(input)
measureContextUsage(input, options)
Parameters
input-
The content to append to the context window. This is either:
- A string — Shorthand for a single textual message.
- An array of objects, each representing a single message in a conversation with a language model.
Objects may have the following properties:
role-
A string indicating the point of view the message is phrased from. Must be one of:
system-
A system-level instruction that guides the model's overall behavior. This must be the first instruction passed to the model.
user-
A message from the user, which the API should respond to.
assistant-
An input that provides context for the AI assistant, such as its persona or the format of its responses. Such messages mainly serve to provide context/history, and further shape how the model responds.
content-
A string representing a textual prompt, or an array of objects. Each object includes the following properties:
type-
An enumerated value representing the type of content. This can be one of:
audio-
Audio content.
image-
Image content.
text-
Textual content.
tool-call-
A tool invocation issued by the model.
tool-response-
The result of a tool invocation.
value-
The content of the message. If the
typeistext, this is always a string. If thetypeisaudioorimage, thevaluecan be one of several different object types; see What data types are accepted?.
prefixOptional-
A boolean, defaulting to
false. Whentrue, the message is treated as a prefix for the model's next generated response rather than a complete turn.
optionsOptional-
Options for measuring context usage. Properties include:
responseConstraint-
An object following the structure defined by JSON Schema defining the precise format the model's output should be delivered in. When provided and
omitResponseConstraintInputisfalse, any implementation-defined constraint-description message is included in the measurement. omitResponseConstraintInput-
A boolean; when
true, the automatic constraint-description message is excluded from the measurement. signal-
An
AbortSignalto cancel the operation.
Return value
A Promise that resolves with a Number representing the number of context window tokens the input would consume.
Exceptions
AbortErrorDOMException-
Thrown if the operation was cancelled via the
signaloption. NotAllowedErrorDOMException-
Thrown if usage of the method is blocked by a
language-modelPermissions-Policy. NotSupportedErrorDOMException-
Thrown if:
- A message's
roleisassistantand itstypeis anything other thantext. - A message's
typeistextand itsvalueis not a string. - The input or output text is in a language the user agent doesn't support for prompting.
- A message's
typeisimageoraudiobut the type was not listed inexpectedInputs, or thevalueis not an accepted data type.
- A message's
SyntaxErrorDOMException-
Thrown if:
- No messages are included in the messages array.
- A message's
prefixproperty is set totrueand:- The message's
roleis notassistant. - The message is not the last item in the messages array.
- The message's
TypeError-
Thrown if:
omitResponseConstraintInputistruebutresponseConstraintis not provided.- A message's
roleissystembut it was not the first message passed to the context.
Examples
>Warning when the context is nearly full
The following example uses a function to verify that context is available before calling LanguageModel.prompt(). It first calculates the remaining context and passes that value to measureContextUsage(). If needed is less than or equal to remaining, it returns true and the session continues.
const promptText = "Let me ask you an interesting question...";
async function contextAvailable(promptText) {
const remaining = session.contextWindow - session.contextUsage;
const needed = await session.measureContextUsage(promptText);
return needed <= remaining;
}
const session = await LanguageModel.create();
if (await contextAvailable(promptText)) {
const response = await session.prompt(promptText);
console.log(response);
} else {
console.warn("Prompt skipped: Not enough context window remaining.");
}
Specifications
| Specification |
|---|
| Prompt API> # dom-languagemodel-measurecontextusage> |