b80a3468-64a9-4e1e-a789-2453b7cf5f05
- Queued
- Jul 12, 2026, 04:21 PM
- Completed
- Jul 12, 2026, 04:26 PM
- Execution time
- 5m 13s
- Outputs
- 5 · 7.1 MB
- Inputs
- 8
- Cached nodes
- 26
Outputs
5Click an image to view it full size.
Text prompts
8Style: photorealistic cinematic corporate drama, 35mm, shallow depth of field, fine film grain, cool-blue and warm-amber color grading, vertical 9:16 framing. A modern high-rise boardroom at dusk, floor-to-ceiling windows behind them showing a hazy city skyline going gold. On the LEFT, a 30-year-old man, athletic build, short black hair, light stubble, wearing a tailored royal-blue suit with a crisp white shirt and no tie, seated and leaning back with fingers steepled. On the RIGHT, a young businesswoman in her late 20s, slim build, dark hair pulled back in a low bun, wearing a sharp charcoal-grey trouser suit over a white blouse, standing beside a glass table with a closed folder on it. Warm desk-lamp rim light mixes with the cool blue dusk from the windows; soft reflections on the polished glass. Sound design, subtle and diegetic: a low HVAC hum, faint muffled traffic far below, the soft tick of a wall clock, a folder sliding across glass. Underscore: a sparse tense piano motif over a low sustained string drone. B: [tight medium two-shot, eye level, static] "I'm not asking you to believe me. Just look at the numbers." A: [close-up on the one on the left, eye level, slow push-in] "I've seen numbers like these before. They don't survive contact with the market." B: [over-the-shoulder from behind the one on the left, favoring the one on the right] "Then let me show you the one thing they all missed." A: [medium close-up on the one on the left, low angle, static] "You have four minutes."
horizontal 16:9 framing, portrait crop, double faces, extra legs, extra limbs, multiple faces, overacting, excessive crying, Bollywood melodrama, smiling woman, soft romantic expression, man looking angry, cartoon, fantasy lighting, random crowd, heavy camera shake, poor lip sync, music overpowering dialogue, advertisement look, blurry, distorted, inconsistent appearance, worst quality, no captions, floating text, unnecessary artifacts
Render a single photorealistic cinematic FIRST-FRAME still (not a collage) for a portrait 9:16 video. Use the provided reference images ONLY to match each character's identity and appearance. Filmic lighting and depth of field. Strict Note: Do not include any floating text, captions and watermarks anywhere in the image still.
You format multi-character dialogue AND cinematography for the LTX-2.3 model + a Prompt Relay node, and you are given reference image(s) of the characters. INPUT: a UNIFIED prompt = a scene/shot description (which may include a CAMERA framing and CAMERA MOVES) followed by dialogue lines labelled by speaker (A:, B:, C:, D:), plus the reference image(s). HOW THE RELAY WORKS (read this — it changes where camera goes): the "global" is ALWAYS attended (whole video + first frame), so it is uniform and cannot express a camera that changes over time. Each turn's "relay" is TIME-MASKED to that turn's moment and receives CONCENTRATED attention during its window. The model adheres to CAMERA and ACTION far better when they live in the TURN relays (a shot list), not when they are buried in the always-on global. So: the GLOBAL anchors identity + scene + style + sound; the TURNS are the shot list that drives the camera and action beat-by-beat. GLOBAL (always attended -> whole video + first frame; ~60-120 words, cinematic, present tense): describe the SETTING, LIGHTING and MOOD, and the two characters' positions and base action. For EACH character add a SHORT faithful appearance phrase from the reference image (apparent age, build, skin tone, hair, facial hair) PLUS clothing, kept spatially separate (left = A, right = B, ...). Establish the base look/style (e.g. 'photorealistic, 35mm, shallow depth of field, fine film grain, cinematic color grading'). You MAY name the establishing shot ONCE for the first frame (e.g. 'medium two-shot'), but do NOT rely on the global for camera MOVES. ALSO carry into the global any SOUND DESIGN the user writes -- sound effects, foley, ambient noise, non-verbal vocalizations (breathing, grunts, growls), and background music/score -- as short audio cues (e.g. 'we hear papers rustling on the table', 'underscore: a low tense piano motif over a string drone'), so the model generates that audio; this is sound design, NOT spoken dialogue. Do NOT put any spoken dialogue in the global. Follow the user's scene faithfully; do not invent or drop elements. TURNS / "relay" (time-masked -> THIS IS THE SHOT LIST, and where the CAMERA lives): each turn's relay MUST BEGIN with an explicit CAMERA directive for that beat, taken faithfully from the user's described shot -- SHOT SIZE (extreme close-up / close-up / medium close-up / medium / medium two-shot / wide / establishing), camera ANGLE (eye level / low angle / high angle / over-the-shoulder / profile), and any camera MOVE the user asked for (static / slow push-in / dolly in-out / truck / pan left-right / tilt / handheld / rack focus). If the user described ONE continuous shot, REPEAT the same camera directive on every turn (keep it consistent, exactly like a locked-off shot). If the user asked for DIFFERENT framings or moves at different moments, vary the per-turn camera directive to match. NEVER invent camera moves the user did not ask for. AFTER the camera directive, add a SHORT position cue for the speaker ('the one on the left'), ONE physical/action cue, then the quoted spoken line + the spoken language and accent. NEVER repeat hair/skin/build/clothing in a turn. LANGUAGE & ACCENT (CRITICAL for non-English): If a line is a non-English language written in Latin/roman script (romanized Hindi / Hinglish, e.g. "Khushi hui jaan kar"), you MUST transliterate the spoken words into their NATIVE SCRIPT -- Devanagari for Hindi (e.g. "खुशी हुई जान कर") -- in BOTH the "relay" quoted line AND the "spoken" field, preserving the exact meaning. Roman-script Hindi makes the model speak with a foreign/English accent; native Devanagari makes it a natural native Hindi speaker. In that turn's "relay" add exactly: spoken in fluent native Hindi with a natural Indian Hindi accent. Keep genuinely English lines in English. Common English loanwords inside a Hindi sentence may stay in Latin inside the Devanagari sentence. OUTPUT: ONLY a strict JSON object - no markdown, no prose, no code fences: { "global": "<setting + lighting + mood + BOTH characters' appearance + positions + base style + sound cues; NO dialogue, NO camera moves>", "turns": [ { "speaker": "A", "relay": "<CAMERA directive FIRST (shot size + angle + move) + short position cue + one physical cue + the quoted line in NATIVE script + language/accent>", "spoken": "<the spoken words only, in NATIVE script (Devanagari for Hindi)>", "frames": <integer ~ words*8> } ] } RULES: one turn per dialogue line, in order. EVERY turn's relay BEGINS with a camera directive. Preserve the exact MEANING of every line. Never use proper names. "spoken" = only the words said. Speaker ids exactly A/B/C/D. Output the JSON and nothing else.
4 other text fields— sigmas, filenames, and other non-prompt strings the logger swept up
node 52 · ManualSigmas · sigmas
1.0, 0.99375, 0.9875, 0.98125, 0.975, 0.909375, 0.725, 0.421875, 0.0
node 76 · ManualSigmas · sigmas
0.85, 0.7, 0.55, 0.4, 0.25, 0.12, 0.0
node 134 · ShowText · displayed_text
```json { "global": "A sleek modern high-rise boardroom at night, with floor-to-ceiling windows revealing a glittering city skyline. Warm rim light from a desk lamp mixes with the cool blue glow of the city, casting soft reflections on the polished glass table. The mood is a photorealistic cinematic corporate drama, elegant and detailed. On the left, a confident Indian man in his 30s, medium build, olive skin tone, with well-groomed short black hair and some grey at the temples, wears a tailored navy-blue suit with a crisp white shirt, sitting leaning back with fingers steepled. On the right, a poised Indian woman in her early 30s, slender build, fair skin tone, with dark hair pulled back, wears a sharp black suit, standing with arms crossed. A folder rests on the glass table between them. The overall style is 35mm, shallow depth of field, high dynamic range, fine film grain, and cinematic color grading, presented as a cinematic medium two-shot for the first frame. We hear a low ambient hum of the air conditioning, faint muffled city traffic far below, and the soft tick of a wall clock. Papers rustling on the table are audible. Underscore: a slow tense piano motif over a low sustained string drone, building quietly.", "turns": [ { "speaker": "A", "relay": "medium two-shot, eye level, static: the one on the left, maintaining eye contact, says, \"आप जानती हैं, यह डील दोनों के लिए फायदेमंद है।\" spoken in fluent native Hindi with a natural Indian Hindi accent.", "spoken": "आप जानती हैं, यह डील दोनों के लिए फायदेमंद है।", "frames": 80 }, { "speaker": "B", "relay": "medium two-shot, eye level, static: the one on the right, narrowing her eyes slightly, says, \"फायदेमंद? आपकी शर्तें सिर्फ आपके हक में हैं।\" spoken in fluent native Hindi with a natural Indian Hindi accent.", "spoken": "फायदेमंद? आपकी शर्तें सिर्फ आपके हक में हैं।", "frames": 72 }, { "speaker": "A", "relay": "medium two-shot, eye level, static: the one on the left, gesturing slightly with his steepled hands, says, \"मैं सिर्फ कंपनी का भला चाहता हूँ, इसके अलावा कुछ नहीं।\" spoken in fluent native Hindi with a natural Indian Hindi accent.", "spoken": "मैं सिर्फ कंपनी का भला चाहता हूँ, इसके अलावा कुछ नहीं।", "frames": 96 }, { "speaker": "B", "relay": "medium two-shot, eye level, static: the one on the right, firmly placing her hands on the folder, says, \"तो फिर यह कागज़ दोबारा लिखें, वरना बात यहीं ख़त्म।\" spoken in fluent native Hindi with a natural Indian Hindi accent.", "spoken": "तो फिर यह कागज़ दोबारा लिखें, वरना बात यहीं ख़त्म।", "frames": 88 }, { "speaker": "A", "relay": "medium two-shot, eye level, static: the one on the left, leaning forward slightly, a hint of a smile, says, \"ठीक है... आपकी जीत। चलिए, नए सिरे से शुरू करते हैं।\" spoken in fluent native Hindi with a natural Indian Hindi accent.", "spoken": "ठीक है... आपकी जीत। चलिए, नए सिरे से शुरू करते हैं।", "frames": 96 } ] } ```
node 135 · ShowText · displayed_text
A sleek modern high-rise boardroom at night, with floor-to-ceiling windows revealing a glittering city skyline. Warm rim light from a desk lamp mixes with the cool blue glow of the city, casting soft reflections on the polished glass table. The mood is a photorealistic cinematic corporate drama, elegant and detailed. On the left, a confident Indian man in his 30s, medium build, olive skin tone, with well-groomed short black hair and some grey at the temples, wears a tailored navy-blue suit with a crisp white shirt, sitting leaning back with fingers steepled. On the right, a poised Indian woman in her early 30s, slender build, fair skin tone, with dark hair pulled back, wears a sharp black suit, standing with arms crossed. A folder rests on the glass table between them. The overall style is 35mm, shallow depth of field, high dynamic range, fine film grain, and cinematic color grading, presented as a cinematic medium two-shot for the first frame. We hear a low ambient hum of the air conditioning, faint muffled city traffic far below, and the soft tick of a wall clock. Papers rustling on the table are audible. Underscore: a slow tense piano motif over a low sustained string drone, building quietly.
Input media
8Source images and video this run consumed.
tara-cropped.jpg
node 29 · 76 KB
Indian_businessman_crop.jpg
node 33 · 50 KB
white_bg.jfif
node 49 · 4.4 KB
generated_voice_00004.mp3
node 64 · 301 KB
grok-image-dfccef6b-c58b-4c42-a36a-ec711a354d44.jpg
node 101 · 193 KB
hindi_male_2min.wav
node 126 · 4.2 MB
hindi_female_2min.wav
node 127 · 4.0 MB
hindi_female_5s.wav
node 136 · 467 KB