Skip to main content
The --max-tokens flag controls the maximum number of output tokens for agent responses.

Quick Start

Usage

Options

Examples

Short Response

Long-form Content

With Research

Token Limits by Model

Setting max-tokens higher than the model’s limit will be capped automatically.