Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
miohtama
on May 14, 2024
|
parent
|
context
|
favorite
| on:
Veo
> Veo's cutting-edge latent diffusion transformers reduce the appearance of these inconsistencies, keeping characters, objects and styles in place, as they would in real life.
How is this achieved? Is there temporal memory between frames?
hackerlight
on May 15, 2024
[–]
Probably similar to Sora, a patchified vision transformer, you sample a 3d patch (third dimension is time) instead of a 2d patch
Consider applying for YC's Fall 2026 batch!
Applications
are open till July 27.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search:
How is this achieved? Is there temporal memory between frames?