PluginBench
Skill
Pass
Audit score 90

tesseract-video

mirage-hq/tesseract

Edit existing footage into finished videos locally with cuts, sound design, and motion graphics.

What is tesseract-video?

Tesseract-video is a local video editing engine for finishing footage with cuts, dialogue preservation, audio mixing, and native motion graphics. Use it to edit and revise video projects when you have source material and need editorial control over pacing, sound, and visual treatment.

  • Make precise cuts and trim footage while preserving dialogue and natural timing
  • Mix and design audio with dialogue, music, sound effects, and foley
  • Add native motion graphics, overlays, and typography as editable layers
  • Review edits frame-by-frame with filmstrip and waveform tools
  • Export finished videos and portable editable project files (.tsrct)
  • Clean up speech audio with EQ, dynamics, and noise reduction

How to install tesseract-video

npx skills add https://github.com/mirage-hq/tesseract --skill tesseract-video
Prerequisites
  • Local execution environment with GPU/filesystem access
  • Tesseract CLI installed (version checked per session)
  • Source footage in supported formats
  • Optional: supplied brand assets, music, or sound libraries
Claude Code
Cursor
Windsurf
Cline

How to use tesseract-video

  1. 1.Resolve the skill root and verify the required Tesseract version
  2. 2.Probe and sample your source material using local media tools to understand its content
  3. 3.Establish the deliverable: audience, aspect ratio, duration, must-keep moments, and brand assets
  4. 4.Plan your edit by deciding for each beat whether to use footage, footage with graphics, or full graphics
  5. 5.Create a new project with `tsrct project create` or check out an existing .tsrct file
  6. 6.Import footage and audio assets using `project import-asset`
  7. 7.Author the edit by placing layers, setting activeRange and sourceRange, and adjusting timing
  8. 8.Read waveforms for music and speech to align cuts to beats and protect dialogue boundaries

Use cases

Good for
  • Editing interview or talking-head footage with eyeline continuity across cuts
  • Timing picture edits to music beats and phrases using waveform analysis
  • Creating ad hooks with visual reveals synced to opening music cues
  • Adding sound design and foley to existing footage for narrative impact
  • Revising existing Tesseract projects with new cuts or audio passes
Who it's for
  • Video editors working with local footage and editorial control
  • Content creators building ads or short-form videos from existing material
  • Producers refining dialogue timing and audio mixing on existing projects
  • Teams needing portable, editable video project files for collaboration

tesseract-video FAQ

Can Tesseract-video generate avatars or create video from scratch?

No. This skill edits existing footage only. It provides no generative-video, avatar, or text-to-video service.

What if I cannot audition the audio in the source material?

Disclose that limit rather than claiming a complete review. Use waveform tools to inspect timing, but acknowledge you cannot confirm audio content by ear.

How do I preserve eyeline continuity when cutting between talking-head shots?

Align the speaker's eye height and screen position across cuts with appropriate framing. Check rendered outgoing/incoming frames and playback while preserving natural head movement.

Can I use Tesseract-video without a local GPU or filesystem?

No. This package requires a local execution environment with GPU and filesystem access. It cannot run on a hosted service or switch rendering engines.

Should I add motion graphics to every scene?

No. Motion graphics are not a quota. For each beat, decide whether footage, footage with graphics, or full graphics best serves the narrative, and only add graphics when they improve the story.

Full instructions (SKILL.md)

Source of truth, from mirage-hq/tesseract.


name: tesseract-video description: Edit existing footage into finished videos locally with Tesseract. Make cuts, preserve dialogue, add purposeful sound design, mix audio, and decide where native motion-graphics scenes or overlays improve the story. Use for video editing and revising Tesseract projects, not avatar or video generation.

Edit video with Tesseract

The CLI may send basic usage telemetry for some commands. You can provide optional attribution via TESSERACT_SKILL=tesseract-video and the other attribution variables. Respect the CLI's telemetry opt-out setting; never enable telemetry on the user's behalf.

Make a deliberate edit from the user's material. Tesseract is the local editing and rendering engine; the agent supplies editorial judgment. The output is both a playable video and its editable project.

Start with the material

  1. Resolve this skill's root from the directory containing this SKILL.md. Read CLI installation and check the required Tesseract version once per session. Use the resolved executable path throughout the task. Scripts and references are relative to the installed skill, not a guessed home directory or development checkout.
  2. Read local operation. Probe the supplied media, sample frames at meaningful points, and listen to its audio using the host's available media tools. Metadata does not tell you what a shot depicts. If audio cannot be auditioned, disclose that limit rather than claiming a complete review.
  3. Establish the deliverable, audience, aspect ratio, duration, must-keep moments, and supplied brand assets from the request. Ask only for missing information that materially changes the edit. For a small correction, preserve the existing edit's style and scope.
  4. Read editorial decisions before planning a new edit. For each beat decide footage, footage with graphics, or full graphics and why. This is not a quota: do not add motion graphics just to demonstrate the engine.

For a new ad or a revision to its opening, read ad hooks. Build a visual hook in the opening seconds from the actual footage, proposition, and evidence. Consider distinct concepts, select the strongest, and map its reveal/motion and purposeful SFX to confirmed opening music cues when music is provided. Render and inspect the hook plus its handoff into the ad before polishing the rest. Preserve the scope of unrelated revisions.

Read waveform editing when music sets the pacing or when trimming talking footage. Generate and open the music waveform before planning picture timing; confirm beats, phrases, and major changes by listening and map visual arrivals to them. For speech, use overview and zoomed source waveforms to choose pause trims, protect quiet word boundaries, and retain natural breathing. Waveform markers are candidates, never automatic cut decisions.

For a substantial new video, share a concise edit plan with the source moments, their narrative purpose, graphic treatment, and audio intent. Routine authorized edits can proceed. Do not require a new approval for every local preview or revision.

Author the edit

  • New edit: tsrct project create creates a portable .tsrct document; import local footage before placing its layers. Existing .tsrct: inspect and check out a copy of its editable JSON. Preserve the original and stable IDs. Save JSON through project commit, or apply supported action batches with project apply, before rendering.
  • Read native authoring before changing project structure. Source time, edit time, and layer time are different clocks. In particular, moving a clip must not silently change which source moment it plays.
  • Keep footage cuts, full scenes, and editable overlays as native layers in the document's single composition. Use groups for scenes. activeRange places a clip in its parent; sourceRange selects the source moment. Do not create base-video, caption, or audio tracks.
  • For talking-head cuts, follow eyeline continuity: align the speaker’s eye height and screen position across cuts with appropriate framing, including punch-ins. Check the rendered outgoing/incoming frames and playback while preserving natural head movement and intentional camera changes.
  • For motion work, read motion design, then only the relevant authoring reference. The examples demonstrate real schema; their visual style is not a required template.
  • For every new edit or sound pass, read sound design. Consider dialogue, music, micro-SFX, structural cues, and foley/ambience; use the roles the piece needs. Add appropriate sounds from supplied/local assets or the procedural helper without per-cue approval. Import the selected files with project import-asset --kind audio, then keep them as editable Audio layers. Use audio and timing for placement, ducking, local preparation, and measurement. Preserve useful source audio and chosen music. Do not produce invented speech or captions.
  • When talking audio needs cleanup, read speech cleanup. Diagnose room noise versus reverb/echo, use restrained EQ/dynamics or local noise reduction as appropriate, and compare at matched loudness. Preserve voice detail and timing; the plugin has no bundled dereverberation model.
  • Keep typography, graphics, footage, and audio native and editable. Import supplied images as assets with project import-asset --kind image, not as a flattened replacement for text, vector graphics, or entire scenes. Use existing footage; this plugin provides no generative-video or avatar service.
  • Respect the user's brand references. Mirage branding identifies this plugin, not every video created with it.

Look up capabilities only as needed

Use the installed CLI's schemas through local operation.

Read capabilities for the feature map and limits, FX authoring for native layers and footage, and motion for supported keyframe and script actions. Search for the specific definition needed rather than dumping the whole schema. The action schema includes supported standalone operations and their field restrictions; use local operation to choose actions versus JSON checkout/commit. The installed schema and actual render are authoritative.

Review and handoff

Use filmstrip review to inspect and correct visual changes. For dialogue cuts and music timing, follow waveform editing. Finish with review and delivery; report actual checks and any limitations. Use absolute local file links and show the video inline when supported. Premiere round-trip, cloud sync, and publishing are not implemented.

If the host cannot execute local commands or access the local GPU/filesystem, explain that this package needs a local execution environment. Do not substitute a hosted service or silently switch rendering engines.