{ jsonpromptstudio }

Action: Gesture

Facepalm Frustration

Hand covering face in frustration or disappointment. Best for: Frustration content, disappointment reactions, exasperation moments.

facepalmfrustrationdisappointmentgestureexasperation

  1. Pick your model's tab — the JSON re-shapes itself to that model's block order and limits.
  2. Click any green value to edit it in place — swap the subject, setting, or style for your own.
  3. Copy and paste the JSON as your prompt in the tool — links below.
Google · duration 4s, 6s, or 8s (8s required for 1080p, 4K, or reference images) · 720p, 1080p, or 4K · audio native · verified 2026-07-11 - click any value to edit
{
  "scene": "Person with hand covering face in clear frustration gesture",
  "action": {
    "description": "facepalming",
    "dialogue": "I can't believe this happened again."
  },
  "camera": {
    "angle": "close-up angle",
    "distance": "close-up shot",
    "movement": "frustrated facepalm gesture"
  }
}
  • Veo 3.1 follows structured prompts unusually well: lock camera, lighting and audio as separate JSON fields.
  • Native audio: describe dialogue, ambient sound and music directly in the audio field.
Kuaishou · duration expanded limits; exact caps vary by mode · mode-dependent · audio native audio-visual output · verified 2026-07-11 - click any value to edit
{
  "action": {
    "description": "facepalming",
    "dialogue": "I can't believe this happened again."
  },
  "scene": "Person with hand covering face in clear frustration gesture",
  "camera": {
    "angle": "close-up angle",
    "distance": "close-up shot",
    "movement": "frustrated facepalm gesture"
  }
}
  • Kling prompt guidance emphasizes Subject, Movement, Scene, Camera Language and Lighting.
  • Write the action field as a shot note, not a keyword list.
ByteDance · duration 4-15s · 480p, 720p, 1080p, 4K ratios in Runway API · audio native synchronized audio on supported hosts · verified 2026-07-11 - click any value to edit
{
  "camera": {
    "angle": "close-up angle",
    "distance": "close-up shot",
    "movement": "frustrated facepalm gesture"
  },
  "action": {
    "description": "facepalming",
    "dialogue": "I can't believe this happened again."
  },
  "scene": "Person with hand covering face in clear frustration gesture"
}
  • Use a strict block order: CAMERA -> SUBJECT -> ACTION -> ENVIRONMENT -> LIGHTING -> STYLE.
  • Keep the camera block to a shot type plus one movement; stacking moves degrades output.
Alibaba · duration workflow-dependent · 720p open-weights TI2V · audio none in base text/image-to-video pipeline · verified 2026-07-11 - click any value to edit
{
  "action": {
    "description": "facepalming",
    "dialogue": "I can't believe this happened again."
  },
  "scene": "Person with hand covering face in clear frustration gesture",
  "camera": {
    "angle": "close-up angle",
    "distance": "close-up shot",
    "movement": "frustrated facepalm gesture"
  }
}
  • Wan 2.2 is the latest version with genuinely open, downloadable weights.
  • Open-weights Wan 2.2 has no native audio track in the base text/image-to-video pipeline.
Runway · duration 2-10s · 720p; output dimensions vary by aspect ratio · audio none in video generation · verified 2026-07-11 - click any value to edit
{
  "action": {
    "description": "facepalming",
    "dialogue": "I can't believe this happened again."
  },
  "camera": {
    "angle": "close-up angle",
    "distance": "close-up shot",
    "movement": "frustrated facepalm gesture"
  },
  "scene": "Person with hand covering face in clear frustration gesture"
}
  • Runway recommends clear, direct language; text-to-video should describe both visual elements and motion.
  • Image-to-video prompts should focus on describing the motion of the scene.

Run this prompt

Copy the JSON above, then paste it as your prompt in any of these tools. Veo 3.1 is free to try inside Gemini (daily limit applies).

What each block does

BlockPurpose
sceneWhere it happens: location, time of day, weather, mood of the space.
actionWhat visibly happens during the clip. Describe motion the camera can see, not internal states.
cameraShot framing and movement. One movement per clip reads best on every current model.

More action: gesture prompts

all →