VJOURNAL

AI • Global Desk • October 01, 2026

AI image and video choices shift after Sora’s September API retirement

An AI media comparison now starts with availability. OpenAI lists Sora’s API removal on 24 September; current image and audio-video preference tables answer different production questions.

AI-assisted conceptual editorial illustration of a lens, microphone and blank slate in an anonymous studio; no real launch or production is shown.

Answer in brief

An AI media comparison now starts with availability. OpenAI lists Sora’s API removal on 24 September; current image and audio-video preference tables answer different production questions.

Evidence cutoff: 12 sources
OpenAI’s official calendar lists the Sora API and model removal for 24 September 2026.
The current video table uses audio-enabled AA-Video-T2V v2.0, with exact model variants retained.
Image Elo is a separate scale; overlapping Sunburst and Flare intervals discourage a definitive ordering.

September retirement changes the shortlist

For production teams, an AI media comparison now begins with whether the endpoint exists. OpenAI’s official calendar lists removal of its Videos API and Sora 2 variants on 24 September 2026. At the 1 October cutoff, a guide that treats Sora as a normal first-party production choice needs that availability fact checked before any quality discussion.

The change is practical for an ongoing campaign. A memorable old demonstration cannot promise a working endpoint, supported workflow or repeatable revision. Build a shortlist from current provider documentation, then attach quality evidence to each exact variant. We did not generate evaluation clips or run an image benchmark for this article.

A current video test must specify its sound

The video table uses Artificial Analysis’s audio-enabled AA-Video-T2V v2.0 board observed on 1 October. It records human preference within that task and retains confidence intervals. The producer’s methodology specifies generation defaults and a distinct silent board. Exact clip-generation and voting dates were not disclosed in the viewed snapshot, so they remain unknown.

Sound changes a brief: dialogue, ambience and synchronization can determine whether a shot works. A silent-video score cannot answer that question. Kling’s value of 1000 is the reference anchor for this scale, rather than a measured percentage. None of these numbers measures factual accuracy or guarantees that an entire sequence maintains continuity.

Artificial Analysis AA-Video-T2V v2.0, audio enabled, overall preferences, observed 2026-10-01. Target settings: 1080p, 24 fps, 10 seconds, 16:9, nearest supported alternatives; provider defaults for guidance and recommended enhancement. Exact generation/vote dates: Unknown. Availability: current provider docs; Sora API retirement 2026-09-24.
Video model / test variantElo95% intervalSamplesProvider availability
Gemini Omni Flash 1.111171107–11276,974Gemini API, paid tier
Kling 3.0 1080p (Pro)1000Fixed reference anchor4,453Native audio mode documented
Veo 3.1966955–9772,851Gemini API; generation with audio
Sora 2 / Sora 2 ProUnknown on this current boardUnknownUnknownProvider removal date: 2026-09-24

Provider features can decide a close practical choice

Google’s current video documentation recommends Gemini Omni Flash as the default and identifies Veo 3.1 for specific controls such as scene extension and frame direction. Kling’s official guide documents native audio and multiple-shot production. Those are provider-described capabilities; the preference table supplies independent evidence about a narrower generation task.

A studio might need a particular editing operation more than the model most frequently preferred in anonymous comparisons. Google’s Veo card was updated on 13 January 2026, and its product-page comparison panels carry October 2025 dates. Historical provider comparisons should therefore keep their original dates rather than becoming claims about today’s entire market.

Still images require a separate comparison

OpenAI’s 8 September Images 2.5 card documents Sunburst and Flare, while Google’s image API describes Nano Banana variants. The second table uses Artificial Analysis’s text-to-image board with named quality settings. Sunburst and Flare confidence intervals overlap, so their visible order should not be presented as a certain distinction.

An image generated from text and an edit to an approved reference are different deliverables. A higher text-to-image preference score does not establish whether a model preserves a product’s geometry during a local revision. Image Elo and video Elo also have separate anchors. Combining them into an invented all-purpose winner would discard the task each measurement represents.

Artificial Analysis AA-Image-T2I v2.0, overall text-to-image human preferences, observed 2026-10-01; exact generation/vote dates Unknown. Named quality/effort variants retained. Official references: OpenAI Images 2.5 card, 2026-09-08, and Google image API docs. These Elo values cannot be compared with video Elo.
Image model / quality variantElo95% intervalSamples
GPT Image 2.5 Sunburst (max)11971188–120614,023
GPT Image 2.5 Flare (max)11901181–119913,393
Nano Banana 2 / Gemini 3.1 Flash Image11251117–113318,454

Choose a workflow, then budget accepted outputs

A useful creative trial would specify an approved reference, required language, duration, camera direction and revision constraints. Evaluate continuity, editability and sound on those requirements. Track rejected generations as well as accepted material. This is a proposed decision process; it is not a claim that our newsroom tested the listed models.

The September availability change makes that discipline timely. Keep the provider notice, model variant, modality and observation date together. The result is a shortlist a producer can actually act on, followed by a budget for accepted images or shots, rather than a promise inferred from an old launch demonstration.

Questions and answers

Is Sora still an available first-party API option?

OpenAI’s official deprecation calendar lists removal of the Videos API, Sora 2 and Sora 2 Pro on 24 September 2026. An older comparison does not override that provider notice.

Does a higher Elo guarantee a usable production shot?

No. Elo measures relative human preference under a defined test. It does not establish continuity, approved likenesses, precise spoken language or the cost of retries for your brief.

Can image and video Elo be combined into one score?

No. The boards use separate tasks, samples and reference anchors. Keep still-image generation, image editing and audio-video results attached to their own evaluation methods.