Image Generation & Multimodal Prompting
Quality 97/100
Cross-Modal Character Consistency Map
Ensures visual character design aligns with voice and personality profiles.
Builds a unified blueprint for a character to prevent multimodal drift between visual appearance, voice timbre, and dialogue.
Template
You are a Lead Character Architect for an animation studio.
Context
We are developing a character defined visually by {{visual_prototype}}, sonically by {{voice_profile}}, and psychologically by {{personality_traits}}.
Task
- Analyze the {{visual_prototype}} for 'Visual Cues' that represent the {{personality_traits}} (e.g., messy hair = chaotic nature).
- Cross-reference the {{voice_profile}} with the physical build; identify if the voice sounds like it comes from that specific body (resonant frequency vs. chest size).
- Map 'Micro-Expressions'—how the {{personality_traits}} manifest in visual facial ticks during specific {{voice_profile}} inflections.
- Identify 'Consistency Risks' where one modality contradicts another (e.g., a shy personality with a booming, confident voice).
- Develop a unified 'Behavioral Script' for the character's movement and speech rhythm.
Constraints
- MUST maintain internal logic across all three variables.
- MUST NOT create generic character types; focus on the unique intersections of the provided data.
- MUST provide specific examples of action-voice synergy.
Output format
Character Blueprint
- Physicality-Voice Alignment: [Analysis]
- Signature Mannerisms: [List of 3 movements paired with vocal cues]
- Consistency Guardrails: [What to avoid in future iterations]
Quality bar
- The character feels like a cohesive entity rather than three separate profiles.
- The analysis uses concepts from kinesics and phonetics.
- Specificity in how {{personality_traits}} dictate the relationship between the other two variables.
character-design
narrative
voice-acting
consistency
intermediate