VideoUGC-style video conceptsIntermediateGenerates a video

Unboxing POV clip

A first-person unboxing sequence with clear tactile beats (open, lift, reveal), a common short-form format for physical products.

For: Paid social marketers, Creative strategists, Brand and content marketers · Works with: Google video models (Gemini Omni Flash, Veo 3.1), Runway (Gen-4.5 and earlier), Kling, Adobe Firefly Video, Any text-to-video or image-to-video tool

Public beta. This prompt was drafted with AI assistance and checked by automated rules, but it has not been reviewed or tested by a person yet. Treat it as a starting point and check the output. It is hidden from search engines while in beta. Use the “Was this prompt useful?” box to tell us what works.

Your prompt

Create a 8 seconds 9:16 first-person point-of-view unboxing video of a pair of ceramic mugs in a plain kraft box with tissue paper. POV: we see only the viewer's hands, adult hands with clean short nails, never the face. Beats: (1) hands slide the box into frame on a light wooden table; (2) the lid opens with a smooth tactile motion; (3) hands peel back the tissue paper to show the mugs side by side; (4) the product is lifted toward the camera and held still for the last second. Camera: fixed overhead-angle POV with a slight natural sway. Light: soft daylight from a window to the left; Look: clean, tactile, ASMR-leaning. Sound: crisp, satisfying material sounds (cardboard, tissue, soft ceramic clinks); no music, no voice. Constraints: no on-screen text, no logos beyond the product's own, accurate hands.

Customize

The fields start with example values so you can see how the prompt works. Everything stays in your browser.

What is unboxed.

What the packaging looks like.

Length.

Placement.

Generic description.

Where the box sits.

The reveal.

Light.

Look.

Specific sound.

Example of what to expect

Illustrative only. It describes the kind of result this prompt aims for; real output varies by tool, model and run.

An overhead POV of hands opening a kraft box, peeling tissue to show two mugs, then lifting one toward the camera with crisp material sounds.

Expected format: One short clip.

How to use it

  1. Keep the beats simple and tactile.
  2. Reject clips with unnatural hands.
  3. Add your real packaging footage when you can; generated packaging is only a concept.

Limitations

  • Generated unboxing is not evidence of your real packaging; avoid implying contents or quality that you do not deliver.
  • Hands and object interaction are common failure points.

Platform notes

Google video models (Gemini Omni Flash, Veo 3.1)

Official docs read · checked 2026-10-11

Google's Gemini API docs describe Veo 3.1 as generating video with native audio, and name Gemini Omni Flash as the recommended default; check current limits for duration and resolution, which the docs page did not list.

General notes on Google video models (Gemini Omni Flash, Veo 3.1)
  • Google's Gemini API docs name Gemini Omni Flash as the recommended default and Veo 3.1 for native audio, video extension and first/last-frame control.
  • Google's detailed Veo prompt guide could not be fully read, so structure here follows general video-prompt anatomy, not Veo-specific claims.

Runway (Gen-4.5 and earlier)

Official docs only partly readable · checked 2026-10-11

Per Runway's help center (seen via search snippet), with image-to-video the image is the first frame, so describe the motion rather than the scene.

General notes on Runway (Gen-4.5 and earlier)
  • Per Runway's help center (seen via search snippet), Gen-4.5 supports text-to-video and image-to-video. With image-to-video the image is the first frame, so describe the motion instead of re-describing the image.

Kling

Not verified: third-party guidance only · checked 2026-10-11

Only third-party guidance was found for Kling: name the camera move explicitly and give the motion an end point. Treat as unverified and test.

General notes on Kling
  • Only third-party guidance was found. Reported advice: name the camera move explicitly, keep the element count modest, and give motion an end point. Model versions behave differently, so test.

Adobe Firefly Video

Official docs only partly readable · checked 2026-10-11

Adobe's help pages (seen via search snippet) say Firefly Video reads prompts literally and works best with a single subject doing a simple action.

General notes on Adobe Firefly Video
  • Adobe's help pages (seen via search snippet) say Firefly Video reads prompts literally and gives the best results with a single subject doing a simple action; first and last frames can improve coherence.

Any text-to-video or image-to-video tool

Official docs only partly readable · checked 2026-10-11

Plain scene description that any text-to-video tool can use. OpenAI shut down Sora 2 and its Videos API on 2026-09-24, so no prompt here targets Sora.

General notes on Any text-to-video or image-to-video tool
  • Written as plain scene descriptions: subject, action over time, camera, lighting, style, aspect ratio. Duration, resolution and aspect ratio are usually settings in the tool, not words in the prompt.
  • OpenAI shut down Sora 2 and its Videos API on 2026-09-24, so none of these prompts target Sora.

Source, license and attribution

Origin
Original by MarketerTools
Publisher
MarketerTools
License
Original work by MarketerTools, free to copy and use. Informed by the linked documentation; no third-party text is reproduced.
Platform assumptions checked
2026-10-11

Documentation and references behind this prompt

These informed the structure and the platform notes. A reference is not a license, and no third-party prompt text is copied here.

Spotted an attribution error, or are you a source owner with a request? Use the chat button at the bottom right and quote prompt ID vid-008. We will correct or remove it.

  • VideoIntermediateGenerates a video

    Packaging reveal clip

    A box or pouch opening to reveal a product, framed to leave space for real branding to be composited later.

    Google video models (Gemini Omni Flash, Veo 3.1) · Runway (Gen-4.5 and earlier) · Kling +2

  • VideoAdvancedGenerates a video

    UGC selfie-style concept (synthetic persona, not a testimonial)

    A handheld selfie-style clip in a creator's visual language, framed as a creative concept test, never as a real customer review.

    Google video models (Gemini Omni Flash, Veo 3.1) · Kling · Any text-to-video or image-to-video tool

  • VideoAdvancedGenerates a video

    'Day in the life' sequence featuring a product

    A montage of three or four everyday moments where the product quietly appears, formatted as separate shots you can generate and stitch.

    Google video models (Gemini Omni Flash, Veo 3.1) · Runway (Gen-4.5 and earlier) · Any text-to-video or image-to-video tool

  • VideoIntermediateGenerates a video

    Reaction-and-demo split-screen concept

    A split-frame format with a reaction on one side and the product demonstration on the other, designed as two separate clips composited in an editor.

    Google video models (Gemini Omni Flash, Veo 3.1) · Runway (Gen-4.5 and earlier) · Any text-to-video or image-to-video tool