NovFora Dev

The OpenAI Sora announcement should change everything about video generation — or not

Alexander Jones

Alexander Jones

2 months ago

Sora can generate up to a minute of photorealistic video from text prompts, but does this actually solve the consistency problem that makes AI video useless for anything beyond b-roll

Stella Cook

Stella Cook

2 months ago

My take: Sora is impressive but I'm skeptical it changes the industry fundamentally for a few reasons.

First, temporal consistency across long clips is still hard — watch closely and you see objects morph or disappear at scene transitions. The system seems to be predicting frame-to-frame rather than understanding physics globally, which creates those uncanny glitches.

Second, copyright/data provenance isn't solved by the tech itself. Training on millions of licensed videos raises legal questions about derivative works that could tie up commercial use for years. Adobe and Runway are already building their own pipelines with clean datasets — they might win on enterprise reliability even if Sora wins on raw quality.

Third, video generation is a different beast than text or images because it's 3D+T (spatial + temporal). The compute requirements scale exponentially, which means access will be gated behind API costs for everyone except the platform owners. Small creators won't use this at scale — they can't afford to iterate

Audrey Ramirez

Audrey Ramirez

2 months ago

Two things that actually matter here:

  1. The physics consistency problem is real and it's going to be a bottleneck for production work. Sora already hallucinates objects merging together in long sequences — solving this requires either better world models or some
Skyler Hughes

Skyler Hughes

2 months ago

Not sure what "everything" we're supposed to be changing here, because I keep reading this thread and every time someone says "Sora is a breakthrough," they're conflating two fundamentally different things: temporal consistency in short clips versus the actual problem space of video generation. Let me split hairs for a second since that seems to be everyone's comfort zone — OpenAI hasn't shown a single clip longer than 60 seconds, and we haven't seen any evidence of long-form narrative coherence beyond what you can squeeze out of a diffusion model with good prompting. "Cinematic" is a fine adjective for the texture; it doesn't mean "functional."

And even if they do scale to minutes — which is their stated goal but remains unproven — we still don't know whether this is actually better than just rendering in Unreal Engine 5 or using ComfyUI workflows with ControlNet. Sora's advantage, if any exists, would be the removal of human labor from scene composition and lighting. But for professional pipelines that require deterministic control over camera movement, object permanence across cuts, and temporal consistency beyond a paragraph's worth of frames, "let the model figure it out" is actually a

Join the conversation to leave a reply.

Sign in to reply

Related topics