scripting-and-storyboarding
Compare original and translation side by side
🇺🇸
Original
English🇨🇳
Translation
Chinesescripting-and-storyboarding
脚本撰写与分镜设计
The pre-production system — start from the spine, write both columns, envision every beat, number the
shots, and edit on paper first. The agent writes the plan; the human shoots and judges; WoopSocial
publishes the finished video. (Craft skill — no tool file.)
前期制作系统——从核心框架入手,撰写双栏内容,预想每个节拍,为镜头编号,先在纸上完成剪辑。由Agent撰写方案;人类负责拍摄与判断;WoopSocial发布成品视频。(制作技能——无工具文件。)
The POV: videos are won in pre-production — the cheapest edit is on paper
核心观点:视频的胜负在前期制作阶段——最廉价的剪辑是纸上剪辑
Chaotic shoots and endless edits are a pre-production deficit, not a talent deficit: the plan is "where you
catch problems before they become expensive fixes on set or in post." Four truths carry this skill. (1) A
words-only script plans half the video — the visual half gets improvised, and improvisation defaults to a
static talking face; the two-column AV script (audio | visual) writes what's seen, beat by beat, so the
change-every-few-seconds rhythm is planned, not hoped for. (2) The storyboard is a decision document, not
art — its only question is "what's on screen at this beat?", answered by a shot table, stick figures, or AI
previz frames (2026 pipelines make character-consistent boards in minutes — the tested benchmark: 12 frames
from ~3 hours to 40–60 minutes, with explicit camera intent and a character bible as the craft rules).
(3) The shot list regrouped BY SETUP is the batching trick — two weeks of content in one afternoon comes
from sorting shots by location/framing/outfit, never by video. (4) Runtime math doesn't negotiate —
~130–150 wpm means a 300-word script is a 2-minute video; the stopwatch pass and the paper cut happen before
the shoot. And the 2026 twist: AI video sequences need more pre-production, not less — every panel becomes
a generation's shot brief, and previz frames cost cents where generations cost credits.
混乱的拍摄和无休止的剪辑源于前期制作的不足,而非能力欠缺:方案是“在问题演变为片场或后期昂贵的修正之前,提前发现并解决问题的环节”。该技能遵循四个核心原则。(1) 纯文字脚本仅规划视频的一半内容——视觉部分会被即兴发挥,而即兴发挥往往会变成单调的访谈镜头;双栏AV脚本(音频 | 视觉)逐节拍写明画面内容,从而提前规划好每几秒切换一次的节奏,而非寄希望于临场发挥。(2) 分镜脚本是决策文档,而非艺术作品——它只需要回答“这个节拍的屏幕上是什么内容?”,可以用镜头表、简笔画或AI预演帧来呈现(2026年的流程可在几分钟内生成角色一致的分镜——经测试的基准:12帧分镜从约3小时缩短至40–60分钟,明确镜头意图并以角色设定手册作为创作规则)。(3) 按拍摄场景重新分组的镜头清单是批量拍摄的关键技巧——一下午完成两周的内容制作,源于按地点/取景框/服装对镜头进行分类,而非按视频分类。(4) 时长计算没有商量余地——每分钟约130–150词的语速意味着300词的脚本对应2分钟的视频;在拍摄前必须完成计时审核和纸上剪辑。2026年的新趋势:AI视频序列需要更多前期制作,而非更少——每个分镜格都成为生成镜头的brief,预演帧的成本仅为几分钱,而生成视频则需要消耗算力点数。
Read these first
需先阅读以下内容
- idea-generation-and-ideation — the idea being productionized.
- short-form-video-script (or the long-form skill) — the retention spine this turns into production.
- brand-profile + design-and-templates — visual language, end cards.
- idea-generation-and-ideation——即将投入制作的创意。
- short-form-video-script(或长内容制作技能)——该前期制作系统基于此核心框架展开。
- brand-profile + design-and-templates——视觉语言、片尾卡片。
The framework: SCENE
SCENE框架
(Depth: .)
references/the-scene-framework.md- S — Start from the spine: one job, audience, format target, beat outline — before any words.
- C — Columns: audio + visual: the two-column AV script with [VO]/[SFX]/[Action] tags; if the visual isn't written, it will be improvised; no "me talking" twice in a row.
- E — Envision every beat: the cheapest board that answers "what's on screen" — shot table → stick figures → AI previz frames (explicit camera intent, character bible + ref image, micro-beats, directional not final); skip boards honestly where a shot list suffices; for AI video, each panel = the shot brief.
- N — Number the shots: shot list with setups → regroup by setup for batching (sequence setups, buffer
takes on hooks, label footage ); runtime math + feasibility pass.
vid#-shot# - E — Edit on paper first: table read with a stopwatch, cut the middle not the spine, the honesty pass (no staged-as-candid, no lifted scripts) — then hand off to the shoot, the edit, and WoopSocial.
(详细内容:。)
references/the-scene-framework.md- S — 从核心框架入手:先确定核心目标、受众、格式要求、节拍大纲——再撰写具体内容。
- C — 双栏:音频 + 视觉:带有[VO]/[SFX]/[Action]标签的双栏AV脚本;如果视觉内容未写明,就会被即兴发挥;禁止连续出现两次“我在说话”的内容。
- E — 预想每个节拍:用最简便的方式回答“屏幕上是什么内容”——镜头表 → 简笔画 → AI预演帧(明确镜头意图、角色设定手册+参考图、微节拍、仅作方向指引而非最终画面);如果镜头清单足够说明问题,可如实跳过分镜;对于AI视频,每个分镜格 = 镜头brief。
- N — 为镜头编号:带有拍摄场景的镜头清单 → 按拍摄场景重新分组以实现批量拍摄(按场景排序、为钩子镜头预留备用素材、将素材标记为);完成时长计算与可行性审核。
vid#-shot# - E — 先在纸上完成剪辑:用秒表进行桌读,删减冗余内容而非核心框架,进行真实性审核(不制作摆拍伪装真实的内容、不抄袭脚本)——然后将方案移交至拍摄、剪辑环节,最终由WoopSocial发布。
The reality (verify-quarterly)
实际应用情况(每季度验证)
Stable craft: the two-column AV script (broadcast standard), ~130–150 wpm runtime math, the
storyboard-as-decision-document, the by-setup batch regroup. The 2026 AI layer (attribute): script→board
pipelines matured (Boords — free tier, sign-off layer; Storyboarder.ai — 250K+ creators, animatics; LTX,
mStudio, Studiovity; free Wonderunit); character consistency is now baseline; the practitioner benchmark cut
12-frame boards from ~3 hrs to 40–60 min using LLM-beats → visual directives → seeded frames — with tags
([VO]/[SFX]/[Action]), explicit camera intent ("the AI will guess — don't make it"), locked palette, and a
character bible; boards stay directional; traditional illustration runs ~$50–300/frame (the economics behind
the shift). AI-video sequences: panel = shot brief (luma's DREAM); board first, generate second. Attribute
all; verify-quarterly. Full detail: ; the templates
and two worked examples: .
references/scripting-and-storyboarding-2026-reality.mdreferences/templates-and-examples.md成熟的制作方法:双栏AV脚本(广播级标准)、每分钟约130–150词的时长计算、作为决策文档的分镜脚本、按拍摄场景分组的批量拍摄方式。2026年的AI层(特性):脚本转分镜的流程已成熟(Boords——免费版,含审批层;Storyboarder.ai——25万+创作者使用,可制作动态分镜;LTX、mStudio、Studiovity;免费工具Wonderunit);角色一致性现已成为基础要求;从业者使用LLM生成节拍→视觉指令→种子帧的流程,将12帧分镜的制作时间从约3小时缩短至40–60分钟——带有[VO]/[SFX]/[Action]标签、明确镜头意图(“AI会自行猜测——不要让它猜”)、固定调色板和角色设定手册;分镜仅作方向指引;传统插画的成本约为每帧50–300美元(这也是转向AI分镜的经济原因)。AI视频序列:分镜格 = 镜头brief(luma的DREAM功能);先做分镜,再生成视频。**注明所有来源;每季度验证。**详细内容:;模板及两个实例:。
references/scripting-and-storyboarding-2026-reality.mdreferences/templates-and-examples.mdHonest scope (never violate)
明确范围(绝不违反)
- The agent writes every planning artifact (beats, AV script, board/briefs, shot list, batch plan, timing pass, footage map); the human shoots/generates, judges every take, and approves — the agent cannot see footage or operate a camera and never fabricates "that take works." AI previz follows the image rules (no unpermitted likeness; original/consented characters; disclosure if frames publish).
- Production integrity: no staged-as-candid content (actors posing as unaffiliated strangers = deceptive
endorsement; skits are fine disclosed); no lifted scripts (structure study yes, verbatim no); honest
runtime/feasibility. WoopSocial publishes the finished video; it does not script, storyboard, shoot, or
edit. (Full scope: .)
references/scope-and-connections.md
- Agent撰写所有规划文件(节拍、AV脚本、分镜/brief、镜头清单、批量拍摄计划、计时审核、素材映射);人类负责拍摄/生成内容、判断每一条素材并进行审批——Agent无法查看素材或操作相机,也绝不会编造“这条素材可用”的说法。AI预演帧需遵循图像规则(不得使用未经授权的肖像;使用原创/经同意的角色;若发布预演帧需披露)。
- 制作诚信:不制作伪装成真实场景的摆拍内容(演员假扮无关陌生人属于误导性宣传;短剧需明确披露则没问题);不抄袭脚本(可以学习结构,但不得逐字照搬);如实标注时长/可行性。WoopSocial仅发布成品视频;不负责脚本撰写、分镜设计、拍摄或剪辑。(完整范围:。)
references/scope-and-connections.md
Distinct from its siblings (route correctly)
与同类技能/工具的区别(正确区分)
scripting-and-storyboarding (this) = the pre-production system · short-form-video-script = the
retention-words craft (WATCH writes the spine; SCENE productionizes it) · talking-head-and-piece-to-camera
= the on-camera delivery · capcut / descript = the edit executing the AV script's visual plan · luma /
ai-video = the generations whose shot briefs the panels become · flux / image-prompt = previz frames +
reference stills · storytelling-and-narrative = the narrative angle/WHAT this schedules into shots ·
idea-generation-and-ideation = supplies the idea.
scripting-and-storyboarding(本技能) = 前期制作系统 · short-form-video-script = 内容留存率相关的文字创作技能(WATCH撰写核心框架;SCENE将其转化为制作方案) · talking-head-and-piece-to-camera = 镜头前的表演技巧 · capcut / descript = 执行AV脚本视觉方案的剪辑工具 · luma / ai-video = 以分镜格作为镜头brief的内容生成工具 · flux / image-prompt = 预演帧+参考静帧 · storytelling-and-narrative = 将叙事角度/内容转化为镜头安排的技能 · idea-generation-and-ideation = 提供创意来源。
Where this connects
关联技能/工具
Reads first: idea-generation-and-ideation + short-form-video-script/long-form + brand-profile +
design-and-templates. Feeds: the shoot day (the human), talking-head-and-piece-to-camera, luma
(shot briefs), capcut/descript (footage map + AV script), content-calendar (the batch). Publishes via:
the finished video → scheduling-and-queue → WoopSocial. Measure with: shoot efficiency + edit speed +
published retention via analytics-and-reporting — never fabricated.
需先参考:idea-generation-and-ideation + short-form-video-script/长内容制作技能 + brand-profile + design-and-templates。输出至:拍摄环节(人类)、talking-head-and-piece-to-camera、luma(镜头brief)、capcut/descript(素材映射+AV脚本)、content-calendar(批量内容)。发布路径:成品视频 → scheduling-and-queue → WoopSocial。衡量指标:拍摄效率 + 剪辑速度 + 发布内容的留存率,通过analytics-and-reporting统计——绝不编造数据。
Definition of done
完成标准
A shootable, editable plan: the spine confirmed (one job, beat outline from the script skill), the two-column
AV script written with every visual beat specified ([VO]/[SFX]/[Action] tags; no "me talking" twice in a row),
the storyboard produced at the cheapest sufficient fidelity (shot table, stick figures, or AI previz frames
with explicit camera intent + a character bible — directional, never final art; skipped honestly where a shot
list suffices), a numbered shot list regrouped by setup with the batch plan, buffer takes, and a labeled
footage map, the runtime verified by stopwatch math (~130–150 wpm) and cut on paper before the shoot, and the
honesty pass held (no staged-as-candid, no lifted scripts, feasibility stated straight); AI-video panels
written as generation shot briefs with previz-before-credits economics; the human shooting and judging,
the edit receiving a clean handoff, and the finished video publishing via WoopSocial; nothing staged as
real, nothing plagiarized, no fabricated production claims; and correctly distinguished from
short-form-video-script, talking-head-and-piece-to-camera, capcut/descript, and luma.
产出可拍摄、可编辑的方案:核心框架已确认(核心目标、来自脚本技能的节拍大纲)、撰写完成双栏AV脚本并明确每个视觉节拍(带有[VO]/[SFX]/[Action]标签;禁止连续出现两次“我在说话”的内容)、以最低必要精度制作分镜脚本(镜头表、简笔画或带有明确镜头意图+角色设定手册的AI预演帧——仅作方向指引,绝非最终艺术作品;若镜头清单足够说明问题则如实跳过)、生成按拍摄场景重新分组的编号镜头清单,包含批量拍摄计划、备用素材和标记好的素材映射、通过秒表计算验证时长(每分钟约130–150词)并在拍摄前完成纸上剪辑、完成真实性审核(不制作摆拍伪装真实的内容、不抄袭脚本、如实说明可行性);AI视频的分镜格作为生成镜头的brief,遵循预演优先于算力消耗的成本原则;人类负责拍摄与判断,剪辑环节获得清晰的移交文档,成品视频通过WoopSocial发布;无摆拍伪装真实内容、无抄袭、无虚假制作声明;并与short-form-video-script、talking-head-and-piece-to-camera、capcut/descript、luma等正确区分。