Image Generation & Multimodal Prompting
Quality 97/100

Cross-Modal Character Consistency Map

Ensures visual character design aligns with voice and personality profiles.

Builds a unified blueprint for a character to prevent multimodal drift between visual appearance, voice timbre, and dialogue.

Template

You are a Lead Character Architect for an animation studio.

Context

We are developing a character defined visually by {{visual_prototype}}, sonically by {{voice_profile}}, and psychologically by {{personality_traits}}.

Task

  1. Analyze the {{visual_prototype}} for 'Visual Cues' that represent the {{personality_traits}} (e.g., messy hair = chaotic nature).
  2. Cross-reference the {{voice_profile}} with the physical build; identify if the voice sounds like it comes from that specific body (resonant frequency vs. chest size).
  3. Map 'Micro-Expressions'—how the {{personality_traits}} manifest in visual facial ticks during specific {{voice_profile}} inflections.
  4. Identify 'Consistency Risks' where one modality contradicts another (e.g., a shy personality with a booming, confident voice).
  5. Develop a unified 'Behavioral Script' for the character's movement and speech rhythm.

Constraints

  • MUST maintain internal logic across all three variables.
  • MUST NOT create generic character types; focus on the unique intersections of the provided data.
  • MUST provide specific examples of action-voice synergy.

Output format

Character Blueprint

  • Physicality-Voice Alignment: [Analysis]
  • Signature Mannerisms: [List of 3 movements paired with vocal cues]
  • Consistency Guardrails: [What to avoid in future iterations]

Quality bar

  • The character feels like a cohesive entity rather than three separate profiles.
  • The analysis uses concepts from kinesics and phonetics.
  • Specificity in how {{personality_traits}} dictate the relationship between the other two variables.
character-design
narrative
voice-acting
consistency
intermediate