ALACRITY · THE SNATCH — FACEPLANT FINISH — 15-Second First-Person POV Combat Video | Neta Studio

ALACRITY · THE SNATCH — FACEPLANT FINISH — 15-Second First-Person POV Combat Video | Neta Studio

Final Video: ALACRITY · THE SNATCH — FACEPLANT FINISH

ALACRITY · THE SNATCH — FACEPLANT FINISH is a 15-second, 9:16 vertical first-person combat short shot entirely from the challenger’s own eyes on the Killing Floor — a black steel combat grid suspended over a dark industrial void. You are the opponent. You never see your own body. You only see your rust-red sleeves, your huge gloved hands, and the petite assassin who is about to end you.

Seven timed beats, one continuous POV: the face-off, the superhuman slip as she vanishes outside the weapon line, the search, the sideways return, the kill beat where her gear and the electric-cyan belt fill the entire frame, the struggle where her abs flex and ripple harder the more you fight — and one decisive twist. Then the camera pitches forward and faceplants onto the steel mesh, while she stays hanging above.

No dialogue, no narration, no third-person cut. Just arena ambience, chain rattle, boot skid, one dry snap, the heavy hit of the mesh — and a restrained industrial-metal score that cuts off with the impact.

Watch it now: ALACRITY · THE SNATCH — FACEPLANT FINISH →

The Video

15 seconds · 9:16 vertical (720×1280) · 24 fps · no dialogue — rendered with Seedance 2.0 Fast in a single image-to-video pass from this world’s own reference frames, with the seven-beat timeline, the absolute first-person camera rule and the faceplant finish all authored directly in the prompt.

Download the MP4 →

How It Was Made

Production Process Breakdown

Stage 1: World Building from One Lineclaude-opus-5
The world starts from a single sentence — a first-person POV video request built around Alacrity, “The Killing Floor Queen”, and her signature technique: The Snatch. Neta Studio expands that one line into a full POV combat world: the Killing Floor arena (black steel grid over a dark industrial void, narrow beams, red hazard stripe, hanging chains), the assassin herself (explicitly adult, exactly 5 feet tall, dirty-blonde, barefoot, torn black combat gear, electric-cyan belt), and the spectator-turned-opponent camera who never appears on screen.
Stage 2: Character Lock & Key Artgpt-image-2
Alacrity’s identity is locked before any motion exists: a clean face reference plus character profile and design sheets, then the visual bible — a director-markup pitch board of the whole 15-second short, and a dedicated kill-moment key frame that fixes the clamp composition as a hard rule. That key frame is the reference the whole video had to match at the 10.3 second mark.
Stage 3: Iterating the Cutseedance-2-0-fast
The world went through 27 video renders to get here — 20 delivered takes and 7 failed attempts that never made it out of the queue. Successive passes tightened the camera rule, the scale separation between the 5-foot assassin and her much larger opponent, the hand-anatomy constraints and the timing of the vanish-and-return swing. Every reroll taught the prompt something; the survivors became the world’s key-frame catalogue.
Stage 4: The Final 15-Second Passseedance-2-0-fast
The finished cut adds the two notes that define it: a stronger abs-flexing struggle — her core visibly rippling and clenching harder against the opponent’s clawing hands — and a faceplant ending, the camera pitching forward and slamming face-first onto the steel mesh instead of falling backwards. Both are written as hard rules into the same 15-second prompt, so the whole sequence renders in one continuous pass, 9:16 vertical at 720×1280.
Stage 5: Cutouts, Frames & Publishbirefnet-generalseedance-2-0-fast
The final MP4 is published behind a shareable link, while the world’s own frames are pulled straight out of the render for the gallery — the face-off, the search and the faceplant drop. Character and key-frame stills are cut to transparent PNG assets with the BiRefNet model, so the same material can be remixed straight into the next shot or sequence.

Generation Prompt

“Reference 1: the pitch board. Take from it ONLY the photoreal live-action texture, the adult petite woman's identity, her costume, and the arena location. Do NOT reproduce its annotations, arrows, text, labels, or layout. Reference 2: a clean key-frame still defining the EXACT kill-moment composition — a pure closeup of the woman's torso: her torn black combat top, electric-cyan belt, and extremely defined glistening midsection fill the ENTIRE frame edge to edge; the cyan belt crosses the bottom of the frame; the top edge cuts her lower chest; NO legs, NO feet, NO hands, NO face, NO background. The kill beat MUST reproduce this composition exactly. Reference 3: the woman's face — her exact facial identity. In every shot where she is visible, especially the opening face-off, her face must match reference 3: dirty-blonde hair, warm golden-brown eyes, confident smirk, petite adult features, pale skin. References 2 and 3 contain no text, no watermark, no logo. No handwritten text, no arrows, no panel borders, no subtitles, no on-screen text, no watermark anywhere in the output. ABSOLUTE CAMERA RULE: the entire video remains first-person POV from the larger adult opponent's body-mounted camera. The viewer IS the opponent and has NO visible body: no viewer legs, no viewer torso; his gloved hands enter frame only to claw at her during the struggle. Do not show the camera wearer's face, back, torso, or a third-person view. HAND ANATOMY IS A HARD RULE: every human in the frame has exactly two hands, one left and one right, with correct thumbs; no duplicated hands, no two right hands, no mirrored hands, no extra limbs, no fused fingers. Visual separation: the petite woman is an explicitly adult martial artist, exactly 5 feet tall, slender and compact, with dirty-blonde hair, golden-brown eyes, an extremely defined midriff with sculpted abdominal muscles and obliques visible through her torn black combat gear, a bright electric-cyan belt at the waist, black gloves, and bare feet; the larger opponent wears bulky rust-red and ochre combat gear, dark boots, and a long weighted steel staff with a red-orange grip. She is visibly much smaller than him, never average-height, tall, broad, or childlike. The petite woman moves faster than a human can track, but every movement is continuous and anatomically coherent, no teleportation, no morphing, no limb duplication. Her first swing away is slightly slower, so she visibly arcs out of sight and seems to vanish; her return swing is faster and comes from the side, off-axis. COMBAT RULES: she never grips the opponent with her hands and never grabs the staff; her hands only grip a beam or hanging chain to anchor her own position. She redirects and fights with her legs, hips, and core. Every visible bare foot has exactly five separate anatomically correct toes, no extra digits, no fused toes, no shoes. THE KILL BEAT COMPOSITION IS A HARD RULE AND MUST MATCH REFERENCE 2: when she executes the head-scissors, the camera is INSIDE the grip with no gap between the lens and her body: her torso is pressed directly against the lens so her torn black combat gear and electric-cyan belt fill the ENTIRE frame edge to edge, exactly like the key-frame still, her defined glistening midsection filling the frame, the cyan belt crossing the bottom, the top edge at her lower chest; NO legs, NO feet, NO face, NO background are visible; his gloved hands may enter from the bottom and sides as he claws at her; the arena, the grid, the chains, and everything else completely disappear from view. ABS-FLEXING DURING THE STRUGGLE IS A HARD RULE: as the opponent claws and fights back, her abdominal muscles are visibly rippling, bulging, and flexing hard against his resistance, her defined abs and obliques contracting and releasing in the closeup, the flexing itself is the fight; the more he struggles, the harder her core clenches, until the single decisive twist that ends it. She never touches the ground and never falls: she stays suspended, anchored by one hand on a hanging chain or beam above. Only the opponent drops: when the neck breaks, his knees buckle and the camera pitches FORWARD and drops face-first, FACEPLANTING onto the steel mesh; the final view is the mesh rushing up and the impact, while she remains hanging above. The fight stays on the platform; no one falls into the void. STRUGGLE AND SNAP TIMING ARE A HARD RULE: the opponent does not stand still: his large gloved hands enter the frame from the bottom and sides, clawing and gripping at her body, and the camera shakes violently while her gear and belt fill the frame and her abs flex and ripple. The finish is fast and decisive: her defined midsection visibly twists sharply once, the lock wrenches, one sharp non-graphic neck-snap sound lands, and his hands go limp and drop away. From lock to break is at most one second. Scene: black steel grid over a dark industrial void, narrow beams and gaps, red hazard stripe, several hanging chains, one support column, cold arena air, scuffed mesh, dust, breath on the lens. Beat timeline: 0.0-2.5 seconds, FACE-OFF POV, opponent's rust-red sleeves, large gloved hands, dark boot tips, and staff frame the foreground; the petite woman stands ahead, much smaller, black-and-cyan gear, bare feet planted on a narrow beam, facing the camera, her face matching reference 3 exactly; low industrial drone, no speech; 2.5-5.0 seconds, SUPERHUMAN SLIP POV, the opponent swings the staff, she slips outside the weapon line faster than the eye, then arcs away on a chain slightly slower, swinging out of sight behind a chain and support column as if she vanished; staff whistle, chain rattle; 5.0-8.0 seconds, SEARCH POV, the opponent stays planted and cautious, staff raised, scanning; chains sway; tense low percussion, no chase; 8.0-10.3 seconds, SIDE SURPRISE POV, she returns faster than before from the peripheral side on a hanging chain at impossible speed, cyan belt cutting across the lens; he flinches and backpedals in place, boots skidding on the mesh, never reaching the edge; 10.3-11.0 seconds, KILL BEAT POV, she springs up and locks her legs around his head, and the camera is suddenly inside the grip, matching reference 2 exactly: her torn black combat gear and electric-cyan belt fill the ENTIRE frame edge to edge, her defined glistening midsection filling the frame, the cyan belt crossing the bottom of the frame, no legs, no feet, no face, no background, his gloved hands entering from the bottom and sides, no gap between the lens and her body; 11.0-12.2 seconds, STRUGGLE AND FLEX POV, his large gloved hands claw at her body, the camera shakes violently, and her abdominal muscles visibly ripple, bulge, and flex hard in the closeup, contracting against his resistance, the flexing itself carrying the fight; 12.2-12.8 seconds, FAST SNAP POV, her defined midsection visibly twists sharply once, the lock wrenches the lens, one crisp dry non-graphic neck-snap sound lands, his hands go limp and drop out of frame; 12.8-15.0 seconds, FACEPLANT DROP POV, her body pulls away from the lens, the view pitches FORWARD and drops face-first toward the steel mesh, the mesh rushing up, the camera FACEPLANTING onto the grid with the heavy impact, while she remains suspended above; immediate cut to black after the final hit. AUDIO RULE: no voiceover, no dialogue, no speech, no narration. Use only native arena ambience, breath, staff whistle, boot skid, chain rattle, steel impact, struggle sounds, one crisp dry non-graphic neck-snap sound synchronized with the visible recoil, the heavy impact of the faceplant hitting the mesh, and a restrained industrial-metal score of low drones, sub-bass pulses, sparse mechanical percussion, and a final cut-off hit. Hard constraints: face-to-face opening and color separation remain; she is visibly present before the first attack with her face matching reference 3; the away-swing is slightly slower to read as a vanish, the return-swing faster and from the side; the kill happens entirely on the platform; the kill beat reproduces reference 2's composition with her torso, gear, and belt filling the entire frame and nothing else visible, no legs, no feet, no face, no background during the clamp; her abs visibly flex and ripple hard during the struggle; the opponent's fall ends in a FACEPLANT onto the mesh, pitched forward, not backward; the snap is fast and decisive, at most one second from lock to break; every human has exactly two hands, one left and one right; she never touches the ground, only the opponent's body drops face-first to the mesh; she remains exactly 5 feet tall and visibly much smaller; hands never grab or hold the opponent or the weapon, only beam or chain; every visible foot has exactly five toes; no gore, no blood, no broken bones shown, no exposed anatomy, no sexualized framing, no childlike depiction, no duplicate character, no extra limbs, no hand clipping, no ghost model, no third-person angle, no camera-wearer face or torso, no text, no watermark, no subtitles.”

Model: Seedance 2.0 Fast — image-to-video, single pass, 15s @ 720×1280 (9:16), with three locked references: the pitch board, the kill-moment key frame and Alacrity’s face.

Asset Prompts

Kill-Moment Key Frame — Reference 2gpt-image-2
“Use the attached references only for the identity of an adult petite woman martial artist: dirty-blonde hair, extremely defined abdominal muscles, torn black combat gear, bright electric-cyan belt at the waist, black gloves, bare feet. Create ONE photoreal POV key frame, 9:16, from INSIDE a grappling leg lock. THE CAMERA IS BETWEEN HER LEGS, LOOKING UP AT HER BODY, AND HER FEET ARE BEHIND THE CAMERA — they cannot appear in frame, there is no lower body in frame at all. The frame is filled one hundred percent by her torso: her torn black combat top and bright cyan belt fill the entire frame edge to edge; her extremely defined abdominal muscles are visible through the torn opening in the black fabric, occupying the center of the frame; only the very tops of her bare thighs press the extreme left and right frame edges; NO feet, NO shins, NO crossed legs, NO forearms, NO hands, NO face, NO background arena — the black fabric, the cyan belt, and her defined midriff are the only things in the frame, from just below the belt line up to just below her chest. Dark industrial mood, red edge light only as a faint sliver at the extreme frame edges, slight handheld motion blur, breath fog on the lens. Adult woman only, non-graphic, no gore, no blood, no exposed anatomy beyond the athletic midriff opening, no sexualized framing. Photoreal live-action, dark red-and-black cinematic grade. Absolutely NO text, NO watermark, NO logo, NO signature, NO stamp, NO caption anywhere in the image.”
Alacrity Face Reference — Reference 3gpt-image-2
“Alacrity face reference: adult petite woman, dirty-blonde hair, golden-brown eyes, confident smirk, black choker”
Director-Markup Pitch Board — Reference 1gpt-image-2
“A polished director-markup pitch board for a 15-second vertical photoreal live-action found-footage POV combat short titled exactly: ALACRITY · THE SNATCH. Dark charcoal production sheet with faint grid texture, clean editorial layout, no scrapbook elements, no coffee stains, no tape. Dominant HERO FRAME is a real first-person camera frame from the viewpoint of an overbalanced combatant at the edge of a black steel combat grid called the Killing Floor: only the camera wearer's gloved hands and forearms are visible in foreground, one hand reaching toward the hazard-striped rim; Alacrity is visible only in partial action fragments permitted by strict POV, her black tactical gloves crossing into frame, a flash of petite athletic lower-body positioning and pale blonde hair at the top edge, no face, no torso, no full body, no third-person view. Red-and-black arena lighting, scuffed steel grid, red hazard stripe, vertical drop into darkness below, breath fog and sweat on the lens, slight handheld motion blur. The charged instant is BEFORE the decisive break; no blood, no gore, no exposed anatomy. Two smaller supporting stills: a close look-down at the hazard-striped edge with boots losing purchase; a canted view of Alacrity's gloved hand catching near the camera and the grid slipping away. Add restrained red director arrows and concise technical notes pointing to visible elements: POV”

Reference Frames

Why Choose Neta Studio

Professional — The cloud harness carries the whole workflow from a one-line world seed to a finished 9:16 cinematic cut, all in the cloud with no local setup. Stable access to the newest models: Claude Opus 5 GPT Image 2 Seedance 2.0 Fast BiRefNet General Kling Image-to-Video.

One-click generation — Officially tuned skills drive every stage: world building, character lock, key art, reference frames, motion, transparent asset cutouts, and instant publishing to a shareable link.

Community & IP-first — Adapt any world or start from a single sentence. Keep your frames, prompts and sprites, then remix them into the next shot, sequence or story.

Studio Interior

Start Creating

This video was generated by Neta Studio from a single sentence. The world took minutes. The 15-second cut took minutes. You can do it too — adapt this world or start fresh from any idea. No coding required, free to start.

×