Independent research
VideoCoCo: Code-as-CoT for Physically-Consistent Video Generation via an Agentic Dual-Engine System
VideoCoCo is an agentic dual-engine framework for physically consistent text-to-video generation. It uses executable Blender code as a process-level chain of thought. A coding agent synthesizes a Blender program from a text prompt, which is executed in a sandbox to produce a deterministic, low-fidelity spatiotemporal draft. A generative video engine then…
Haodong Li, Tianfei Ren, Xiaoxiao Ma, Chunmei Qing, et al.- Published
- Jul 2026
- Citations
- 0
- Code
- Not linked
