OpenAI launch snapshot
GPT-4o (2024-05-13)
The original GPT-4o API snapshot released in May 2024, useful today as a historical text-and-vision baseline for legacy compatibility checks, regression analysis, and migration planning.
- Context window
- 128K
- tokens
- Max output
- 4K
- tokens
- Input
- $5.00
- per 1M tokens
- Output
- $15.00
- per 1M tokens
- Shutdown
- Oct 23
- 2026
01 / Overview
What GPT-4o (2024-05-13) Is
GPT-4o (2024-05-13) is the first dated GPT-4o API snapshot, released when OpenAI introduced GPT-4o on May 13, 2024.
The Original API Version of GPT-4o
The model ID gpt-4o-2024-05-13 captures the launch-era behavior of GPT-4o rather than a later revision of the family. For developers maintaining older applications, that distinction can matter because later GPT-4o snapshots changed pricing, output limits, supported response formats, and model behavior.
OpenAI initially exposed GPT-4o to developers as a text-and-vision model. The API snapshot accepts text and image input and produces text output. Although the broader GPT-4o research announcement demonstrated native audio and video capabilities, those modalities were not exposed through this standard May 2024 API snapshot.
The original snapshot supports a 128,000-token context window and up to 4,096 output tokens. It launched at $5 per million input tokens and $15 per million output tokens.
- API model ID:
gpt-4o-2024-05-13. - Released with GPT-4o on May 13, 2024.
- Text and image input with text output.
- Useful primarily for legacy compatibility and historical regression testing today.
- Provider
- OpenAI
- Family
- GPT-4o
- Snapshot
- 2024-05-13
- Model ID
- gpt-4o-2024-05-13
- Knowledge cutoff
- Oct 1, 2023
- Input modalities
- Text, Image
- Output modality
- Text
- Lifecycle
- Deprecated
02 / Launch baseline
Why the May 2024 GPT-4o Snapshot Still Matters
This snapshot is most valuable as a historical baseline: it preserves the first GPT-4o API behavior that many production applications originally integrated against.
Measure How Far Your Application Has Moved Since Launch
An application created in 2024 may contain prompts, retry rules, JSON parsing, image preprocessing, tool definitions, and acceptance thresholds tuned around the first GPT-4o snapshot.
That legacy behavior can make migration more complicated than simply changing a model string. A later model may be cheaper and more capable but respond with different phrasing, produce longer outputs, choose tools differently, or interpret screenshots in another way.
Using the May snapshot in an evaluation set lets a team reconstruct its original baseline. You can then compare a newer GPT-4o snapshot or a current replacement model against the behavior the application was actually built around.
This is a different use case from choosing a current general-purpose model. The question is not whether the May snapshot is the best model today; it is whether your replacement preserves or improves the production behaviors that matter.
- 01
Recreate the original baseline
Run representative legacy prompts against the exact model version used when the application was first built.
- 02
Identify compatibility assumptions
Find parsers, prompt patterns, image-processing choices, and tool workflows that depend on launch-era behavior.
- 03
Compare a replacement
Run the same cases against a newer model and measure quality, cost, latency, output shape, and failure rate.
- 04
Retire the dependency
Remove reliance on the deprecated snapshot before the OpenAI shutdown date.
03 / Vision
GPT-4o (2024-05-13) as a Vision Baseline
Vision was one of the major reasons GPT-4o mattered at launch: developers could send text and images to one general-purpose model and receive a text response.
Preserve Old Visual Test Cases Before Migrating
Applications that adopted GPT-4o early often used it for screenshots, photographed documents, charts, UI states, diagrams, and other image-assisted workflows.
Those workloads should be part of any migration evaluation. A replacement model may have stronger vision overall, but production safety depends on the specific images your system processes: small interface text, dense tables, unusual crops, low-resolution scans, or visual layouts that historically caused failures.
The May snapshot gives you a launch-era reference point. Keep the image, prompt, and expected result fixed, then compare the same input against the replacement model.
The API version of this snapshot does not expose native audio or video input. The May 2024 product announcement demonstrated broader GPT-4o multimodality, but developers initially received GPT-4o in the API as a text-and-vision model.
- 01
UI screenshots
Re-test interface understanding, visible errors, controls, labels, and support screenshots against a newer model.
- 02
Scanned documents
Compare extraction quality on forms, receipts, photographed pages, invoices, and other visual documents.
- 03
Charts and diagrams
Measure whether a replacement preserves or improves interpretation of graphical evidence.
- 04
Image preprocessing
Keep crops, resolution, and prompt structure constant so the model comparison remains meaningful.
04 / Pricing
GPT-4o (2024-05-13) Pricing
The first GPT-4o snapshot launched at $5.00 per million input tokens and $15.00 per million output tokens.
Later GPT-4o Snapshots Became Significantly Cheaper
Pricing is one of the clearest historical differences between the May snapshot and later GPT-4o versions. When OpenAI released gpt-4o-2024-08-06, it described the newer model as 50% cheaper on input and 33% cheaper on output than gpt-4o-2024-05-13.
That makes the May snapshot useful for cost migration analysis. If an application still uses this model, compare the full cost per accepted task against the replacement rather than treating its historical token rates as representative of the GPT-4o family today.
OpenAI did not document a separate cached-input rate for this launch snapshot in the same way it later did for prompt-caching-compatible GPT-4o versions.
- $5.00 per 1M input tokens.
- $15.00 per 1M output tokens.
- No separate cached-input price is listed for this snapshot in the launch-era pricing.
- Later GPT-4o snapshots reduced both input and output pricing.
1M tokens · USD
- Input
- $5.00
- Output
- $15.00
Example: 8K input + 2K output
- Input cost
- $0.0400
- Output cost
- $0.0300
- Estimated total
- $0.0700
05 / Context
128K Context With a 4K Output Limit
GPT-4o (2024-05-13) supports a 128,000-token context window but only up to 4,096 output tokens, making its output ceiling materially smaller than later GPT-4o snapshots.
Output Length Is a Useful Migration Test
The large input window made the launch snapshot practical for substantial conversations, documents, and image-assisted prompts. But the 4,096-token output limit can constrain long reports, code generation, or verbose structured responses.
Later GPT-4o snapshots increased the output ceiling. That means a migration can change not only quality and price, but also how much content the model can return in a single call.
When evaluating a replacement, include cases that historically hit output truncation. A newer model may simplify application logic by reducing continuation calls, partial-response handling, or manual output splitting.
- Context window: 128,000 tokens.
- Maximum output: 4,096 tokens.
- Large input capacity for a 2024 multimodal model.
- Much smaller output ceiling than later GPT-4o revisions.
Context window
128,000
Max output
4,096
The May 2024 snapshot combines a large 128K context window with a comparatively small 4K output cap.
06 / Capabilities
API Capabilities of the First GPT-4o Snapshot
The May 2024 snapshot supports the core GPT-4o launch workflow—text, vision, streaming, function calling, and JSON-mode integrations—but it predates several features added to later GPT-4o versions.
Do Not Assume Every Modern GPT-4o Feature Exists Here
The model supports function calling, which allows applications to connect model decisions to predefined tools. JSON mode can be used when valid JSON is required.
However, response_format Structured Outputs with strict JSON Schema support was introduced for gpt-4o-2024-08-06 and later snapshots, not the May 2024 snapshot. GPT-4o fine-tuning also launched on the August 2024 base model rather than this original version.
This difference is important for migration testing. If an older application relies on manual JSON validation or repair logic, moving to a model with native schema-constrained output can be an architectural improvement rather than merely a model swap.
- Supported
Text
Accept text input and generate text output.
- Supported
Image input
Analyze screenshots, photos, scanned pages, charts, diagrams, and other supported images.
- Supported
Streaming
Receive generated text progressively.
- Supported
Function calling
Generate application-defined function calls for tool-connected workflows.
- Supported
JSON mode
Generate valid JSON when using the compatible JSON response mode.
- Not listed
Structured Outputs response format
Strict json_schema response formatting was introduced with gpt-4o-2024-08-06 and later.
- Not listed
Fine-tuning
GPT-4o fine-tuning launched on the later gpt-4o-2024-08-06 base model.
07 / Deprecation
GPT-4o (2024-05-13) Is Deprecated
OpenAI has deprecated gpt-4o-2024-05-13 and schedules API access to shut down on October 23, 2026.
This Snapshot Should Now Be Treated as a Migration Dependency
As of October 4, 2026, the shutdown date is close. OpenAI lists GPT-5.6 Sol as the recommended replacement for this snapshot.
That changes the purpose of evaluating the model. New applications should not select the May 2024 snapshot as a fresh production dependency. Existing users should instead document the behavior they need to preserve, build an evaluation set, test the replacement, and remove the deprecated model from production configuration.
A useful migration suite should include normal traffic and failure cases: image inputs, long prompts, tool calls, JSON parsing, output-length edge cases, and any requests where the old model required special handling.
- 01
Snapshot released
GPT-4o and gpt-4o-2024-05-13 launched on May 13, 2024.
- 02
Deprecation announced
OpenAI deprecated the legacy snapshot as part of its older-model retirement program.
- 03
Recommended replacement
OpenAI lists GPT-5.6 Sol as the replacement path for gpt-4o-2024-05-13.
- 04
API shutdown
OpenAI schedules the snapshot to become unavailable on October 23, 2026.
08 / Evaluation
GPT-4o (2024-05-13) Strengths and Limitations
The first GPT-4o snapshot is now most useful as a historical compatibility baseline. Its value comes from reproducing old behavior long enough to migrate safely, not from competing with current models for new workloads.
Strengths
Historical production baseline
Teams can reproduce the first GPT-4o API behavior when validating changes to older applications.
Text-and-vision integration
The snapshot provides the original GPT-4o text-and-image workflow used by many 2024 integrations.
Function calling
Legacy tool-connected applications can test their existing function definitions against the exact old model version.
Clear migration reference
A fixed snapshot makes improvements and regressions easier to measure when moving to a modern replacement.
What to consider
Imminent shutdown
OpenAI schedules API removal for October 23, 2026, so continued production dependency has a hard deadline.
Higher historical pricing
The $5 input / $15 output rates are substantially higher than later GPT-4o pricing.
4K output limit
Maximum output is only 4,096 tokens, far below later GPT-4o snapshots and current-generation models.
Missing newer GPT-4o features
The snapshot predates response-format Structured Outputs and GPT-4o fine-tuning support.
Migrate with a measurable baseline
Compare your legacy GPT-4o workload before shutdown
Run the prompts, screenshots, JSON workflows, and tool calls your application still sends to gpt-4o-2024-05-13, then compare the same cases against a replacement model before changing production traffic.
Start FreeConnect your own OpenAI API key and use the launch snapshot only while it remains available. OpenAI schedules API shutdown for October 23, 2026.
Common Questions
What is GPT-4o-2024-05-13?
GPT-4o-2024-05-13 is the original dated GPT-4o API snapshot released on May 13, 2024. It preserves launch-era GPT-4o behavior for text-and-image applications.
How much does GPT-4o (2024-05-13) cost?
The snapshot uses the original GPT-4o pricing of $5.00 per 1M input tokens and $15.00 per 1M output tokens.
What is the context window of GPT-4o-2024-05-13?
The model supports a 128,000-token context window and up to 4,096 output tokens.
Does GPT-4o-2024-05-13 support images?
Yes. The API snapshot accepts text and image input and generates text output.
Does GPT-4o-2024-05-13 support native audio or video?
No. Although the broader GPT-4o launch demonstrated audio and video capabilities, developers initially received this API snapshot as a text-and-vision model.
Does GPT-4o-2024-05-13 support function calling?
Yes. Function calling is supported and can be used to connect the model to application-defined tools.
Does GPT-4o-2024-05-13 support Structured Outputs?
It does not support the strict json_schema response format introduced with gpt-4o-2024-08-06 and later. JSON mode and function calling remain available for legacy structured workflows.
Can GPT-4o-2024-05-13 be fine-tuned?
No. OpenAI's GPT-4o fine-tuning release used the later gpt-4o-2024-08-06 base model.
Is GPT-4o-2024-05-13 deprecated?
Yes. OpenAI schedules the snapshot for API shutdown on October 23, 2026.
What should replace GPT-4o-2024-05-13?
OpenAI currently lists GPT-5.6 Sol as the recommended replacement. Existing applications should compare the replacement against their real GPT-4o prompts, image inputs, tool calls, JSON workflows, and output-length edge cases before switching production traffic.
Model information
Last updated
This page uses OpenAI documentation for the GPT-4o family, the May 13, 2024 GPT-4o launch, the August 2024 Structured Outputs release, and the current OpenAI deprecation schedule. The May 2024 snapshot launched at $5 per 1M input tokens and $15 per 1M output tokens, predates response-format Structured Outputs and GPT-4o fine-tuning, and is scheduled for API shutdown on October 23, 2026.