Skip to main content
The --image flag enables image processing with vision-capable AI models.

Quick Start

Usage

Basic Image Analysis

Expected Output:

Specify Vision Model

Combine with Other Features

Supported Image Formats

Use Cases

Document Analysis

Expected Output:

Chart/Graph Analysis

Code Screenshot Analysis

UI/UX Review

Object Detection

Expected Output:

Image Path Options

Best Practices

For best results, use high-resolution images with clear content. Blurry or low-quality images may produce less accurate descriptions.
Image processing uses more tokens than text-only prompts. Use --metrics to monitor costs.

Image Quality

Use clear, well-lit images for best results

Specific Prompts

Be specific about what you want to analyze in the image

File Size

Large images are automatically resized; originals under 20MB recommended

Model Selection

Use GPT-4o or Claude 3 for complex image analysis