Skip to content

Show HN: 4K lip-sync from a 12-word prompt–V2A beats my 2023 render farm

veo-3.app
3 pointsprovinescoch293discuss
On HN

I used to queue 6 hrs on a 4090 just to get dialogue that still floated half a second late. Last week I swapped the entire pipeline for Google’s Veo 3 running behind a 6-line FastAPI wrapper; now the same 30-second spot drops in 18 s with 99 % mouth-match accuracy. No timeline, no keyframes—just text → mp4 with synced sfx baked in. Benchmark repo and the wrapper are MIT on GitHub, plug your key and roast my latency numbers. Live demo at https://veo-3.app/

Comments

No comments yet.

AboutSource Built by g1lg1l

Hackerly is an independent reader for Hacker News, built on the public HN API. Not affiliated with Y Combinator.