Prompt flow
- PROBLEM
- ANALYSIS
- ROUTING
- STRUCTURED PROMPT
SYSTEM 03
MULTIMODAL AI PROMPT ORCHESTRATION
Local multimodal tool for analysing image and video references and composing structured prompts for generative AI workflows.
SYSTEM IDENTITY
SYSTEM IDPROMPT-DIRECTOR-03
STATUSACTIVE R&D
TYPEMULTIMODAL AI / VISION / LOCAL MODELS
Reference analysis and prompt composition pipeline for image and video generation workflows.
01 / OVERVIEW
Generative AI workflows often use multiple references, and each reference can cover a different aspect of the desired output.
Prompt Director groups this context from analysis to routing and turns it into a structured prompt shape before the target generation stage.
Prompt flow
02 / USE CASES
UC-01
Turn local visual references into structured descriptions for later prompt composition.
UC-02
Prepare motion and scene context from local video material when a workflow needs it.
UC-03
Map separate references to characteristics such as identity, wardrobe, action or lighting.
UC-04
Combine analysis, routed context and user intent into a generation-ready prompt shape.
03 / MULTIMODAL INPUT
The tool handles local image and video references as input materials, with optional grouped references for richer prompt context.
Input family
Single image or pair of images prepared locally.
Local capture or local file as a motion source.
Optional grouped inputs routed independently by target characteristic.
04 / ANALYSIS PIPELINE
Input sources first pass through local routing. Video streams may use frame preparation with FFMPEG when needed, while OCR is applied only when the image content includes text to enrich descriptors.
Analysis flow
Optional branch map
VIDEO STREAM
FFMPEG for frame extraction before Qwen3-VL visual context can be used.
IMAGE STREAM
OCR can run as an optional step when text is detected in visual references.
05 / REFERENCE ROUTING
Each source is mapped to outputs that are meaningful for later prompt composition.
Reference routing topology
REFERENCE A
PORTRAIT
REFERENCE B
BODY
REFERENCE C
VIDEO
06 / PROMPT COMPOSITION
Composition combines structured analysis, routing output and explicit intent. The output format is designed for R&D exploration with MiniMax H3 / reference generation prompts.
Prompt composition map
07 / TECH STACK
08 / STATUS
SYSTEM ID PROMPT-DIRECTOR-03
STATE ACTIVE R&D
DEPTH MULTIMODAL GENERATIVE AI TOOLING