
Three video models — h3-max-turbo, veo3.1-lite and wan-3.0 — each got the same 1672×941 chinchilla image and the same prompt. All three failed the same three of five checks.
Summaries like this, in your inbox every morning.
The test started from a practical question: which animal scenes earn views. Counting 1,279 YouTube animal videos, the author found median views varied by an order of magnitude — 53.1 million for drinking scenes, 21.25 million for sneezing, but 5.16 million for grooming. Grooming looked safe and its motion easy to see, so it was chosen. That choice is what later mattered.
The first round changed only the model, holding image, prompt, length and judging consistent, with Codex scoring five items. All three models passed camera lock and centering and failed the other three, and their failure locations matched. wan-3.0, the most expensive at $0.10 per second, was the only one to fail centering and changed the face mid-clip. The two cheaper models matched item by item, so price did not predict quality.
The author then separated model limits from subject difficulty by hiding the front paws and shrinking the motion to whisker twitches and blinks, dropping wan-3.0. Fusion, warping and disappearance did not reappear. What remained across both remaining models were thin lines — whisker count and length, and forehead and cheek stripes. These point to a shared weakness in current models, not a gap between them.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
An engineer ran a single ffmpeg command with fps=1/5 and tile=8x16 on an 8-minute 24-second 1080p clip; it fin…

Anthropic added Claude Dashboards, which turns company data into live dashboards, and Claude Motion, which cre…

HyperFrames, HeyGen's open-source (Apache 2.0) video tool, turns an HTML file into MP4 using headless Chrome a…

Anthropic launched Dashboards and Motion in beta for Claude

Pollo AI launched Pollo Agent, using OpenAI's GPT‑6 Astra and GPT‑Image‑2.5 to guide video creation, and says…

Google LLC launched the web-based SynthID Detector, which flags AI-made images, video and audio from Google, O…