GPT-4o ("o" for "omni") is OpenAI's versatile, high-intelligence flagship multimodal model. It processes text and image inputs to generate text outputs and serves as the primary choice for most tasks outside the o-series reasoning models. The model supports a 128,000-token input context with 16,384 output tokens, and exposes streaming, function calling, structured outputs, fine-tuning, and predicted outputs across Chat Completions, Responses, Realtime, and Assistants APIs.
Share cards5 images




