GPT-5.2, Runway 4.5, and Image AI: A Release Roundup
Three releases landed this week. OpenAI shipped GPT-5.2, Runway deployed Gen-4.5, and the industry formed a standards body for AI agents. OpenAI also announced a $1 billion investment from Disney. The announcements are below, with the numbers as reported.
GPT-5.2: Specs and First Benchmark Results
Section titled “GPT-5.2: Specs and First Benchmark Results”OpenAI launched GPT-5.2 after a short delay. The release follows complaints that GPT-5.1 was unreliable on accuracy. The model ships with a 400,000-token context window, about 300,000 words, and a 128,000-token output limit. API pricing is $1.75 per million input tokens and $14 per million output tokens.
On SWE-bench Pro, GPT-5.2 scores 55.6%. That is up from 50.8% for GPT-5.1. Claude Opus 4.5 sits at 52%, and Gemini 3 Pro at 43.3%. These are vendor-reported figures on one benchmark. Independent comparisons are still thin, and accuracy tests in production settings are pending.
Disney Invests $1 Billion in OpenAI
Section titled “Disney Invests $1 Billion in OpenAI”OpenAI announced a $1 billion investment from Disney. The deal gives OpenAI access to Disney’s IP library for Sora video generation and the native image tools. Possible products include personalized Disney+ shorts, such as AI-generated clips of Disney characters.
Image Generation: What Ships with GPT-5.2
Section titled “Image Generation: What Ships with GPT-5.2”GPT-5.2 ships with native image generation. In testing, the model renders photoreal portraits, readable text, and code overlays. Examples include whiteboard slogans and JSON overlays on product shots. It shows fewer proportion errors than earlier GPT image models. Subtle artifacts remain in eyes and skin, and results vary on recognizable faces.
Agentic AI Foundation: A Standards Body for Agents
Section titled “Agentic AI Foundation: A Standards Body for Agents”OpenAI, Anthropic, and Block launched the Agentic AI Foundation under the Linux Foundation. Google, Microsoft, Amazon, Bloomberg, and Cloudflare back the group. The goal is a common standard so agents from different vendors operate across apps under the same safety rules. Without such a standard, agents that handle email, bookings, and troubleshooting risk locking users into one vendor.
Runway Gen-4.5: Deployed and Tested
Section titled “Runway Gen-4.5: Deployed and Tested”Runway started deploying Gen-4.5 this week. Runway calls the results state-of-the-art for motion, physics, and prompt adherence, and the model leads its internal text-to-video charts. It simulates weight, fluid dynamics, and consistent faces. It does not generate audio.
Hands-on tests of the deployed model:
- Glass sphere on marble stairs: realistic bounces, water splashes, and refractions. The prompt match is close.
- Rainy street walker: umbrella physics, a subtle smile, and handheld camera jitter read correctly.
- Anime explorer: foreground consistency holds. The background is unstable.
- Barista latte pour: swirling milk, steam, and blurred patrons look correct.
- Neon alley chase: reflections are accurate. Minor physics and camera errors appear in the 5-second clip.
Prompt fidelity is the model’s main advantage. Veo 3.1 still leads on realism and sound integration.
In Brief
Section titled “In Brief”- Mistral released Devstral 2, a coding model with public weights. It scores 72.2% on internal benchmarks, close to DeepSeek v3.2.
- Zhipu AI released GLM-4.6V, a vision model for tool calling. Qwen updated Omni Flash with more lifelike voices.
- OpenAI paused shopping suggestions that looked like ads and added user controls.
- ChatGPT gained Adobe connectors for Acrobat, Express, and Photoshop. Early tests show actual limits.
- Meta took over the Limitless pendant, an always-on audio recorder. Privacy questions remain unanswered.
- Alibaba released Image2LoRA, which builds style and character LoRAs from a single image.
Rivian: Autonomy Plans and Silicon
Section titled “Rivian: Autonomy Plans and Silicon”At Rivian’s AI and Autonomy Day, the company showed custom silicon built with Nvidia and integrated LiDAR. Its roadmap targets hands-free driving and unsupervised Level 4 operation by 2027-28. A voice assistant handles calendar, messages, and car controls.
McDonald’s AI Ad Draws Criticism
Section titled “McDonald’s AI Ad Draws Criticism”McDonald’s released a fully AI-generated holiday ad. It drew criticism for looking low-budget beside the company’s production spend. Commenters asked for work by people, with AI used in limited roles.
The Takeaway
Section titled “The Takeaway”The week’s releases show a maturing market: specialized models, a standards body, and clearer pricing. The figures above come from the vendors. Independent testing will decide which claims hold.