GPT-4.1 Nano (batch) chat
For tasks that demand low latency, GPT‑4.1 nano is the fastest and cheapest model in the GPT-4.1 series. It delivers exceptional performance at a small size with its 1 million...
Capabilities
Context Window 1M tokens
Max Output 32k tokens
Inputs
Outputs
Pricing (per 1M tokens)
Input $0.05
Output $0.20
Cache Read $0.01
Cache Write -
Supported Parameters
max_tokensresponse_formatseedstructured_outputstemperaturetool_choicetoolstop_p