Workflow operators
Shape the text before it reaches the Generator with the Merge, Caption, Style and Pose operators.
What this video covers
Operators work on the text before it reaches the Generator: Merge, Caption, Style and Pose.
Merge joins incoming prompts into one text, in letter order; you set the number of inputs in its menu, from two to twelve.
Caption turns an image into text. We wire the taco truck photo into Caption, and its output into Merge's second input.
Caption is billed per image; its estimated five credits are added to the total in the run bar.
Double-click a Style node to open Prompt Lab: one thousand and sixty presets across style, camera, light, lens, stock and palette.
Under Light, we search for "neon" and pick Neon Wash. The preset's description is added to the end of the incoming text, and the node takes on its artwork.
You can chain Style nodes: in the second one, we pick the 500T negative from Stock.
A Pose node opens the Pose Atlas: three hundred and fifty poses in four classes. The pose you pick is added to the prompt as a body instruction.
From Fashion & Advertising, under Commercial & Product, we pick Two-Hand Display.
When you run it, the operators build the text in order: the caption cost three credits and the image one hundred.
Caption's menu shows the description it wrote, and in the result the scene from the photo, the cook and the neon light come together in one image.
Click the image once and it opens full size in a new tab.
In short: Merge gathers texts, Caption turns an image into text, and Style and Pose add to it. In the last part, we'll cover managing your runs.
More from Going further
All tutorials
1:43Tutorial 14Editing your images
Edit an image with an editing model: put it in the reference slot, say what should change and what should stay, keep the aspect ratio and combine several images.
1:31Tutorial 15aYour first workflow
Build your first chain in Workflow: add nodes, pick models, wire an image into a video, check the price and run it.
1:50Tutorial 15bBatch generation in Workflow
Run one line with many inputs using Batch, Dice and Stack, and see how each one affects the price.
1:49Tutorial 15dManaging workflow runs
The Preflight Check, running part of a graph, groups, cancelling, saving workflows and the panels on the right of the canvas.
1:24Tutorial 16aAI voice-over
Turn text into speech in LUVI: pick a voice-over model and a voice, set the stability and language, and generate.
1:26Tutorial 16bAI music
Make background music, or a full song with your own lyrics, in LUVI with MiniMax Music 2.6.
1:06Tutorial 16cTranscribing audio
Turn a recording into text with xAI STT v1, with speaker labels and filler words when you need them.
2:06Tutorial 17Talking avatars and lip sync
Make a portrait talk from an image and an audio file, or give an existing video a new voice with lip-sync models.
1:23Tutorial 18aGenerating 3D models
Turn an image or a description into a 3D model with Hunyuan 3D or HI3D, and see how extra views and parameters shape the result.
1:17Tutorial 18bThe 3D viewer and downloads
Inspect your 3D model in the viewer with lenses, lighting, wireframe and texture views, then download it as a GLB.