OpenAIOpenAI·text

GPT-5.4 Mini

VisionReasoningWeb searchFunction calling
Advertised as gpt-5.4-minigpt-54-miniopenai-gpt-54-mini
Quick reference
GPT-5.4 Mini — TLDR
  • - 🆕 Compact GPT-5.4-class model for high-throughput, latency-sensitive workloads.
  • - 📏 400,000-token context window with up to 128K output tokens.
  • - 👁️ Accepts text and image inputs.
  • - 🔧 Supports function calling, tool use, file search, and computer use.
  • - 🌐 Built-in web search for grounded responses.
  • - ⚡ Optimized for speed in coding assistants and parallel subagents.
  • - 🎯 Recommended for classification, extraction, ranking, and coding subtasks.
💰 Best price on AntSeed
$0.0067 / $0.027
per 1M · cheapest in / out
📏 Context
400K tokens
🐜 Sellers
18
advertising on AntSeed
Provider

OpenAI is an American artificial intelligence research organization headquartered in San Francisco, structured as both a for-profit public benefit corporation and a nonprofit foundation. The lab developed the GPT family of large language models, the DALL-E image generation…

Explore 29 more models by OpenAI →
About this model

GPT-5.4 Mini, released in 2026 by OpenAI, is a smaller, faster sibling within the GPT-5.4 generation, bringing many of the strengths of GPT-5.4 to a model designed for high-volume, cost-sensitive deployments. It supports text and image inputs, tool use, function calling, web search, file search, computer use, and skills, alongside a 400,000-token context window. OpenAI positions it for workloads where latency directly shapes the product experience, such as responsive coding assistants and computer-using systems that interpret screenshots.

Within the catalog's mini lineage, it follows GPT-4o Mini. According to OpenAI, GPT-5.4 Mini and its companion nano are the company's most capable small models yet, and OpenAI now recommends starting with GPT-5.4 mini for most new low-latency, high-volume workloads in place of the earlier GPT-5 mini.

In practice, OpenAI describes a delegation pattern in Codex where a larger model like GPT-5.4 handles planning, coordination, and final judgment, while GPT-5.4 Mini subagents tackle narrower subtasks in parallel—searching a codebase, reviewing a large file, or processing supporting documents.

It sits alongside other GPT-5.4 tier models, including GPT-5.4 Pro, and the later GPT-5.5 and GPT-5.5 Pro releases. OpenAI recommends it as a default starting point for new low-latency, high-volume agent and chat workloads.

Sources
developers.openai.comGPT-5.4 mini Model | OpenAI API· developers.openai.comopenai.comIntroducing GPT-5.4 mini and nano | OpenAI· openai.com

This About section is AI-generated from public sources via VeniceStats + Venice inference, with no human editing. It may contain inaccuracies.

Usage on AntSeed
Tokens served
129.84M
input + output
Requests
26,127
settled calls
Buyers
24
distinct, on this model
Sellers used
14
of 18 advertising
Settled
$59.31
gross USDC, this model
Sellers serving GPT-5.4 Mini (18)compare on the network explorer →
SellerReputation↓RoutingInput $/MCached $/MOutput $/MCategoriesAPI

"Best price" and the seller table are live AntSeed catalog data (advertised $/1M tokens — or $ per generated image for unit-billed image models — not settled amounts). Reputation = buyer trust score (0-100, the AntSeed SDK's own formula). "Routing" = the SDK's default buyer routing (what the VPR desktop app ships with): a trust ≥ 60 gate on the effective reputation, then cheapest-first among routable sellers; live failover state (per-peer cooldowns) is buyer-side runtime and not included. Model knowledge (TLDR, provider, About) via the VeniceStats enrichment layer. Advertised catalog, not the model used in any specific purchase. "Usage on AntSeed" counts only settlements whose buyers share the per-model split on-chain (metadata v2/v3, opt-in), so every usage figure is a lower bound.