Install any skill in seconds. Free to start, no credit card required.
Get Started Free →Create, edit, caption, voice, assemble, validate, and export editable video timelines with Timeline Studio. Use for automatic video editing, AI voiceover videos, subtitle generation, image-to-video assembly, short-form video production, deterministic local video rendering, .timeline project automation, or end-to-end editor evaluation in Codex, Claude Code, Copilot, and Gemini CLI.
.claude/skills/martindelophy-edit-timeline-studio/SKILL.md| Test case | Without → With | Effect | Δ tokens | Δ turns |
|---|---|---|---|---|
| case-04 | ✗→✓ | ▲ Improved | 202% | 0% |
| case-12 | ✗→✓ | ▲ Improved | 196% | 0% |
| case-07 | ✗→✓ | ▲ Improved | 393% | 0% |
| case-08 | ✗→✓ | ▲ Improved | 287% | 0% |
| case-09 | ✗→✓ | ▲ Improved | 285% | 0% |
Turn the user's exact editorial request and media into reversible Timeline Studio edits. Keep the editable timeline as the source of truth; never replace it with an opaque one-shot render.
When the user is working in an already open Timeline Studio project, prefer its native browser WebMCP tools if the host exposes them. Read references/webmcp-integration.md: inspect the live state, preview the complete supported edit, review its semantic diff, then apply the returned previewId. This route shares the visible editor and differs from the local project-file workflow below. It does not add a remote MCP server or grant permission to upload media.
node scripts/setup-host.mjs --check. Agent-driven Chinese and mixed Chinese/English voiceover uses Timeline Studio's owned browser-local Hojo TTS Light 80M two-voice bundle and does not require a separate Python voiceover capability. If language runtimes or dependencies are missing, show the exact installation plan and obtain explicit user approval before install mode; never treat Skill installation as permission to modify the host or download models..timeline through the command layer or local archive services, and render and verify decoded output when the task changes rendered media. Do not open a browser merely because the editor has a UI.https://video-editor.ai-creator.top/ as the canonical hosted editor only when the user explicitly asks to use the website, provides no local repository or project path, or requires a hosted-only capability.timeline_project_diff before timeline_project_apply with the same revision and operations. The MCP server is a transport over the repository command runner, not a separate editing implementation.package.json for an Agent command script. Do not use npm run ... --if-present as capability detection because it can succeed silently. If the command runner exists, read references/command-contract.md, inspect the project, build a versioned plan, run the structural validator, and use project.diff as the authoritative semantic dry run before project.run.For a marker-only request, read references/timeline-markers.md and follow its inspect → plan → validate → diff → apply → inspect workflow through annotation handoff. Markers, chapter cues, ranges, and notes are project annotations; adding them requires no narration, model download, browser session, or video render. The media-production steps below apply only when the user also requests a media edit.
editing-style replication, AI-generation replication, or a hybrid; reconstruct filters, repetitions, source splits, speed curves, transitions, shots, and timing before building; and explicitly resolve whether the authorized original audio track must be retained. Do not start editing until the replication analysis-completeness gate passes. Use current web search to compare AI video platforms only when generation is required, and use lawful web-sourced footage only when the user has not supplied adequate material.marketing-commerce, website-promotion, or other promotional edits, read references/promotion-narrative-workflow.md. Proactively construct an ambitious evidence-backed umbrella narrative rather than a feature list or kinetic-typography montage. Unless the user explicitly requests a teaser, build a complete problem-to-transformation-to-proof-to-CTA arc and actively find several visually distinct cases—normally at least three—each with its own setup, product action, visible result, and connection to the final payoff. Never invent customers, outcomes, metrics, or product behavior to make the story feel larger.tutorial-demo, vlog-event, marketing-commerce, and narrative-documentary as narrated-by-default categories. Preserve and reuse authorized source speech when it already carries the required story; otherwise author the minimum complete narration needed for context, progression, visible result, consequence, and closure, then synthesize it before timing picture. Do not ask whether narration should exist unless the user explicitly requests a source-only, music-only, or silent treatment; ask only for choices that materially affect language, casting, claims, or delivery.zh_f_qinglan or 若溪 / zh_f_ruoxi—for each narrator or character; never install or use Hojo 40M or MeloTTS for this route. Use a remote service, operating-system voice, or unowned runtime only when the requested language or voice is unavailable locally and the user explicitly approves that fallback. Unless the user requests another delivery, choose the warmest natural storyteller-like match from the eligible local profiles and direct a close, conversational performance with meaningful phrasing; never default to a flat, metallic, or mechanical system-voice effect. Generate narration as separate short breath-group segments, not as one monolithic performance to split afterward. Treat a comma as a sufficient default synthesis boundary, prefer several short phrase clips over one long sentence clip, and use the shared desktop/H5 0.4s gap between adjacent voice clips. Lock the complete segmented audio spine before finalizing scene durations, motion, transitions, captions, or picture cuts; adapt and trim visuals to the measured speech and pauses, never the other way around. Treat runtime only as an outcome measurement and do not target, chase, or align to a preset number of seconds.scripts/validate_edit_plan.mjs <plan.json> for transport-shape errors, then run npm run agent -- project.diff <plan.json> to reject unsupported operations and invalid project-specific edits before applying anything.tutorial-demo, vlog-event, marketing-commerce, and narrative-documentary; it remains optional for other categories unless the brief requires it. If a configured caption has no authorized speech route, omit it or stop with the editable project preserved.audioClipId; do not generate one monolithic narration file and split it after synthesis. Split at sentence-ending punctuation and, by default, at commas, semicolons, colons, em dashes, or another clear spoken pause. Keep a boundary joined only when splitting would create a meaningless fragment or break a proper name, number, URL, or intended bilingual phrase. For free-script generation, keep the first clip at the explicit playhead, append every later clip after the current voiceover-track end, and never reuse an unchanged playhead or 0s; caption-scoped generation stays anchored to its caption. Place adjacent narration clips with the shared desktop/H5 default 0.4s gap, then derive caption timing and picture timing from the accepted audio sequence..timeline archive before a destructive batch.output.render in a command plan or claim that project.run renders video. Use the separate versioned project.render request for its documented portable subset, and use the browser editor for AI generation or unsupported composition features..timeline project and the rendered result video there. Planning, diagnosis, annotation-only work, and an explicit editor-only handoff are exemptions. Annotation-only delivery needs a newly written, inspected .timeline archive; do not render an unchanged video merely to deliver markers.adelay, use adelay=<milliseconds>:all=1 or provide one delay value per channel; a single value with the default all=false delays only the first channel and can pile every later clip into the other channel at time zero. Before delivery, compare left/right activity in the opening window and around every scheduled speech boundary. Reject channel-only early speech, multiple narration clips stacked at the opening, or undocumented interchannel onset skew.0.4s of intentional timeline space. Inspect the isolated speech bus at the opening and reject repeated free-script generations that share 0s, reuse an unchanged playhead, or overlap before their scheduled starts. Also reject overlong multi-clause synthesis, a monolithic narration that was merely cut into ranges, meaningless micro-fragments, accidental overlaps, clipped breath/release tails, or picture timing that forced the accepted speech out of its natural cadence.-18 LUFS integrated and no higher than -2 dBTP, require the loudest-to-quietest segment spread to stay within 1 LU, and keep segment LRA within 5 LU unless an intentional exception is documented. Never accept a narration mix from full-program loudness alone, and do not rely on one-pass normalization of short clips as proof of consistency..timeline and result video under distinct paths, record their modification times and SHA-256 hashes, and compare them with the prior artifacts. Never present, relink, rename, or copy an old render as evidence of the fix..timeline and require exactly one segment at the initial playhead, every later start to equal the preceding end plus the planned gap within timeline precision, and no overlaps. Then inspect or decode the final rendered audio at the same boundaries; project structure alone is insufficient proof that the mix is correct..timeline in Timeline Studio, not only with structural inspection: the main Visuals track must be visible, the first frame must render in Preview, archived media must resolve, and captions/audio/track state must match. Fully decode and verify the rendered video, then return both absolute paths.For editor evaluation, regression work, or any run that exposes friction, read references/e2e-evaluation.md. For automatic-editing evaluation, also read references/auto-edit-scenarios.md and use its fixed category cards, clarification checks, hard gates, and adjacent stress variants. Capture the attempted action, observed result, evidence, fallback, and verification. Classify the finding as product, browser-control, environment, or skill guidance. Update the smallest relevant skill instruction or reference, validate the skill, reinstall the local copy, and rerun the affected scenario plus adjacent smoke tests. Never weaken an assertion merely to make a test pass.
Read references/host-environment.md for host dependency checks and approved installation, and references/voiceover-workflow.md before Agent-driven narration or pre-voiceover generation.
Read references/current-capabilities.md when deciding whether a request can be executed now. Read references/webmcp-integration.md when editing a project already open in a browser that exposes Timeline Studio tools. Read references/mcp-integration.md when connecting or invoking the bundled MCP server, and references/command-contract.md when implementing or invoking the underlying Agent command layer. Read references/local-model-routing.md before model-assisted analysis or enhancement. Read references/remote-video-generation.md before selecting or calling a remote video generator, digital-human service, or programmable composition service. Read references/web-footage-sourcing.md for provider-neutral, current web and short-video footage suggestions. Read references/browser-workflow.md for UI execution, references/auto-edit-workflow.md for category-aware automatic editing, references/promotion-narrative-workflow.md for evidence-backed product and promotional storytelling with closed-loop cases, references/replication-workflow.md for editing-style and AI-generation remakes, references/highlight-tension-workflow.md for peak hierarchy and tension shaping, references/curves-and-subject-effects.md for speed curves, Color Wheels, cutout, outline, and authorized face-swap shot design, references/professional-editing-workflow.md for shared media analysis, generation negotiation, stabilization, enhancement, and artifact delivery, references/auto-edit-scenarios.md for its repeatable category matrix, and references/e2e-evaluation.md for repeated experience-driven testing.
For public explanations, route one question to one page: use docs/agent-video-editing.md for what Timeline Studio is; the platform guide for Codex, Claude Code, GitHub Copilot, or Gemini CLI for discovery and invocation; docs/examples.md for reproducible cases; docs/command-reference.md for exact runner syntax; and docs/comparison.md for FFmpeg, CapCut, and Remotion comparisons. Do not load all public pages unless the user asks for a broad overview.
If a requested operation is unsupported, keep the valid partial timeline unchanged and state the exact missing command or runtime capability.
| Case | Status | Duration (ms) | Turns | Tokens | Tool calls | ||||||||
|---|---|---|---|---|---|---|---|---|---|---|---|---|---|
| Without | With | Δ | Without | With | Δ | Without | With | Δ | Without | With | Δ | ||
case-06 | pass→pass | 9,513 | 6,408 | -33% | 1 | 1 | 0% | 1,517 | 6,352 | +319% | 0 | 0 | — |
case-01 | fail→fail | 18,379 | 6,790 | -63% | 1 | 1 | 0% | 3,555 | 5,589 | +57% | 0 | 0 | — |
case-02 | fail→fail | 27,840 | 8,112 | -71% | 1 | 1 | 0% | 5,470 | 5,813 | +6% | 0 | 0 | — |
case-03 | fail→fail | 6,085 | 7,488 | +23% | 1 | 1 | 0% | 361 | 5,696 | +1478% | 0 | 0 | — |
case-04 | fail→pass | 13,872 | 7,632 | -45% | 1 | 1 | 0% | 2,188 | 6,600 | +202% | 0 | 0 | — |
case-05 | fail→fail | 9,321 | 6,235 | -33% | 1 | 1 | 0% | 1,367 | 6,200 | +354% | 0 | 0 | — |
case-12 | fail→pass | 15,036 | 9,471 | -37% | 1 | 1 | 0% | 2,256 | 6,682 | +196% | 0 | 0 | — |
case-07 | fail→pass | 8,046 | 4,678 | -42% | 1 | 1 | 0% | 1,205 | 5,946 | +393% | 0 | 0 | — |
case-08 | fail→pass | 9,899 | 6,057 | -39% | 1 | 1 | 0% | 1,622 | 6,274 | +287% | 0 | 0 | — |
case-09 | fail→pass | 18,061 | 7,790 | -57% | 1 | 1 | 0% | 1,656 | 6,379 | +285% | 0 | 0 | — |
case-10 | fail→pass | 12,975 | 8,933 | -31% | 1 | 1 | 0% | 1,883 | 6,614 | +251% | 0 | 0 | — |
case-11 | fail→pass | 12,798 | 6,409 | -50% | 1 | 1 | 0% | 1,919 | 6,270 | +227% | 0 | 0 | — |
case-13 | fail→pass | 15,296 | 11,306 | -26% | 1 | 1 | 0% | 2,171 | 6,972 | +221% | 0 | 0 | — |
case-14 | fail→pass | 11,879 | 9,168 | -23% | 1 | 1 | 0% | 1,747 | 6,598 | +278% | 0 | 0 | — |
case-15 | fail→pass | 7,611 | 5,180 | -32% | 1 | 1 | 0% | 1,189 | 5,977 | +403% | 0 | 0 | — |
case-16 | fail→pass | 14,541 | 10,770 | -26% | 1 | 1 | 0% | 2,126 | 6,854 | +222% | 0 | 0 | — |
case-17 | fail→fail | 13,498 | 5,195 | -62% | 1 | 1 | 0% | 1,996 | 6,135 | +207% | 0 | 0 | — |
case-18 | pass→pass | 10,865 | 8,337 | -23% | 1 | 1 | 0% | 1,620 | 6,509 | +302% | 0 | 0 | — |
case-19 | fail→fail | 12,269 | 8,678 | -29% | 1 | 1 | 0% | 1,880 | 6,522 | +247% | 0 | 0 | — |
case-20 | pass→pass | 24,155 | 10,108 | -58% | 1 | 1 | 0% | 2,097 | 6,804 | +224% | 0 | 0 | — |
case-21 | pass→pass | 21,167 | 12,754 | -40% | 1 | 1 | 0% | 4,307 | 7,578 | +76% | 0 | 0 | — |
case-22 | pass→pass | 11,452 | 9,412 | -18% | 1 | 1 | 0% | 2,030 | 6,888 | +239% | 0 | 0 | — |
case-23 | pass→pass | 13,434 | 11,381 | -15% | 1 | 1 | 0% | 2,831 | 7,501 | +165% | 0 | 0 | — |
DecimalAI ran this skill against gemini-3.6-flash twice over the same eval suite — once with the skill loaded and once without — and compared the two runs case by case. 23 cases were attempted, and 20 counted toward the lift figure. The other 3 produced results that are not comparable between the two arms, so they are excluded from the headline rather than averaged into it. The headline lift of +48 percentage points is the difference between those two pass rates over the 20 comparable cases. 1 case got worse with the skill loaded, and it is included in that figure.
The publisher has shipped newer versions since this run, so these numbers describe v4, not the version currently listed.
Without the skill loaded, the model failed this case. With it loaded, the same prompt on the same model passed. This is one improved case from the latest verified run; every case, including any that regressed, is in the table above.
| Model | Method | Date | Lift |
|---|---|---|---|
| gemini-3.6-flash | verified | 8/20/2026 | +44% |
Other measured skills in the registry, with their headline benchmark lift.