Skip to content
For the complete documentation index, see llms.txt. Markdown versions of documentation pages are available by appending .md to the page URL.
Primary navigation

Get input token counts

InputTokenCountResponse beta().responses().inputTokens().count(InputTokenCountParamsparams = InputTokenCountParams.none(), RequestOptionsrequestOptions = RequestOptions.none())
POST/responses/input_tokens

Returns input token counts of the request.

Returns an object with object set to response.input_tokens and an input_tokens count.

ParametersExpand Collapse
InputTokenCountParams params
Optional<List<Beta>> betas
Optional<Conversation> conversation

The conversation that this response belongs to. Items from this conversation are prepended to input_items for this response request. Input items and output items from this response are automatically added to this conversation after this response completes.

Optional<Input> input

Text, image, or file inputs to the model, used to generate a response

Optional<String> instructions

A system (or developer) message inserted into the model’s context. When used along with previous_response_id, the instructions from a previous response will not be carried over to the next response. This makes it simple to swap out system (or developer) messages in new responses.

Optional<String> model

Model ID used to generate the response, like gpt-4o or o3. OpenAI offers a wide range of models with different capabilities, performance characteristics, and price points. Refer to the model guide to browse and compare available models.

Optional<Boolean> parallelToolCalls

Whether to allow the model to run tool calls in parallel.

Optional<Personality> personality

A model-owned style preset to apply to this request. Omit this parameter to use the model’s default style. Supported values may expand over time. Values must be at most 64 characters.

Optional<String> previousResponseId

The unique ID of the previous response to the model. Use this to create multi-turn conversations. Learn more about conversation state. Cannot be used in conjunction with conversation.

Optional<Reasoning> reasoning

gpt-5 and o-series models only Configuration options for reasoning models.

Optional<Text> text

Configuration options for a text response from the model. Can be plain text or structured JSON data. Learn more:

Optional<ToolChoice> toolChoice

Controls which tool the model should use, if any.

Optional<List<BetaTool>> tools

An array of tools the model may call while generating a response. You can specify which tool to use by setting the tool_choice parameter.

DeprecatedOptional<Truncation> truncation

The truncation strategy to use for the model response. - auto: If the input to this Response exceeds the model’s context window size, the model will truncate the response to fit the context window by dropping items from the beginning of the conversation. - disabled (default): If the input size will exceed the context window size for a model, the request will fail with a 400 error.

ReturnsExpand Collapse
class InputTokenCountResponse:
long inputTokens
JsonValue; object_ "response.input_tokens"constant"response.input_tokens"constant

Get input token counts

package com.openai.example;

import com.openai.client.OpenAIClient;
import com.openai.client.okhttp.OpenAIOkHttpClient;
import com.openai.models.responses.inputtokens.InputTokenCountParams;
import com.openai.models.responses.inputtokens.InputTokenCountResponse;

public final class Main {
    private Main() {}

    public static void main(String[] args) {
        OpenAIClient client = OpenAIOkHttpClient.fromEnv();

        InputTokenCountParams params = InputTokenCountParams.builder()
            .model("gpt-5")
            .input("Tell me a joke.")
            .build();

        InputTokenCountResponse response = client.responses().inputTokens().count(params);
    }
}
{
  "object": "response.input_tokens",
  "input_tokens": 11
}
Returns Examples
{
  "object": "response.input_tokens",
  "input_tokens": 11
}