A Weird Machine project in collaboration with Brookline Interactive Group Downloads Tools Who it's for Toolbox Roadmap Source ↗
control-z undo
free media cleaning, prepping, analysis, and finishing tools — for your edit and everything around it

Undo the paywall. Make better media for free.

Journalists, artists, advocates, new filmmakers, and community media stations shouldn't have quality gated behind a paywall. If the content is good, the form should match. Yet the people with the most important stories to tell often have the least access to the tools that make work land — so this suite catches you up, free and open-source, from the technical to the editorial to the creative, built for the best free editor around: DaVinci Resolve.

twelve tools · $0 forever · local only · MIT
get it

Download the Tools.

Plugins that live inside Resolve's node tree, apps that run beside any editor, and the scaffolding to put them to work. All free, all MIT.

1

The app

one window · runs beside any editor
the desktop app · v1.1.0 shipped · v1.5.0 in signing · macOS first, Windows next 100% FREE shipped
Civic Media Studio

Ten tools in one window — scopes on every panel, a batch queue that survives a whole tape shelf, ProRes 4444 out, an OpenFX installer for Hush and Speak built in, and the Meeting Library reading every meeting on your machine as one record. The workbench on either side of your edit.

Rise Scribe Clear Stencil Depth Pivot Highlighter Grabber Index Slate Library
Download v1.1.0
2

OpenFX plugins

they live inside Resolve's node tree · free edition

Two halves of one idea: Hush takes the noise out at the top of the tree, Speak puts the optical character back at the bottom. Clean early, reconstruct late.

night · first node · shipped Hush Quiets the noise — measured, and it shows its work. Hush + Speak, together → or grab the latest Hush release ↗
day · last node · v1.0 Speak Gives the image its voice — H&D curves, halation, grain. Hush + Speak, together → or grab the latest Speak release ↗

One page tells the whole story — the buttons open it; the small links jump straight to each release. macOS and Windows for Hush; macOS for Speak so far. Hush hands its clean-confidence matte to Speak's grain — the handoff, working.

3

Resources

scaffolding, recipes, and the thinking behind it
powergrade · start here100% FREE

The node tree

a blank grading scaffold · .drx · 31 MB

The structure, not the look: a working grade's skeleton — noise reduction, exposure, contrast, balance, sat, a foreground/background split, and a look — wired and empty. Its noise and look stages are compound nodes holding a Studio-or-free choice: a Studio feature if you have it, or the free combo that matches it. Hush and Speak drop onto the free paths.

See the tree + how to use it →
powergrade · one control point100% FREE

Middle Gray Contrast Tool

an 18% gray pivot for your curves · .drx · ~3 KB

A single control point pinned to the exact position of 18% middle gray in DaVinci Wide Gamut / Intermediate (code 344, not the middle of the curve). Sits on the diagonal, so it changes nothing — until you shape contrast around it and your mid-gray stays put instead of drifting darker every time you add punch. Drop it in the node tree's Contrast stage.

Why it matters + how to use it →
fusion pack · 10 setups100% FREE

The template pack

veil blur · cutout · fog · rack focus · depth grade · +5

Ten paste-ready Fusion setups that turn a matte into finished work: privacy-blur inside a tracked Stencil matte, composite on an alpha, mist a shot's far end, rack to a focal plane you didn't shoot, grade near/mid/far separately, 9:16 with a blurred backdrop. Each one paste-tested in a live Fusion comp; built from stock nodes — nothing needs Studio.

See what each does + download →
the paperFREE, OBVIOUSLY

Design study

why transparent tools beat black boxes

The white paper behind the suite's thesis — measurement, texture, and the case for tools that show their work. It ends where Hush and Speak begin.

Read the study
using it with davinci resolve (free) The native home.

1 · Install Hush (and Speak, when it ships) with the one-click installer — they appear in OpenFX on the Color, Edit and Fusion pages.
2 · Run the standalone tools before or after your edit; they hand results back as files Resolve just imports: SRT subtitles, marker & selects EDLs, matte clips, keyframed Fusion .settings.
3 · Every tool section below ends with its exact Resolve recipe.

using it with premiere pro (or anything else) The suite doesn't check your NLE.

The standalone tools are editor-agnostic: Scribe's SRT and EDLs, Clear's cleaned WAVs, Rise/Pivot/Stencil/Depth's ProRes renders and mattes drop into Premiere, Final Cut, or Avid the same way.
The honest exception: Hush and Speak are OpenFX plugins — Premiere doesn't host OpenFX, so run noisy footage through the Suite app instead (or grade in free Resolve; it costs the same as this website).

who is this for?

Built for the people the market skips.

community media centers The home team.

control-z comes out of a working access station — every tool is proven on member footage, tape vaults, and city meetings at Brookline Interactive Group first. Ten donated seats, zero license dollars: captions for compliance (Scribe), rescued rooms (Clear), archives upscaled (Rise), every show vertical (Pivot).

journalists Fast, local, verifiable.

Transcribe the presser on your laptop — not a cloud service (Scribe). Pull the quote as a cut list. Rescue the scanner-room audio (Clear). Matte a face for protection (Stencil). Nothing leaves your machine, which is not a feature so much as a duty of care.

filmmakers Finishing without a finishing budget.

Hush then Speak is a grade sandwich that used to cost $500 in plugins. Depth fakes the rack focus you didn't shoot. Rise saves the archival pickup. Pivot cuts the vertical trailer while you sleep. All of it in the free Resolve you already edit in.

artists & advocates Instruments, not appliances.

Every tool exposes its measurements — which means every tool can be played. Grade through depth mattes, key through noise-confidence, push Rise until reality gets soft. And when the work is advocacy, form is credibility: it deserves both.

the suite

A finishing pipeline, not a plugin pile.

Restore → edit → sound → color → deliver. Lit nodes are shipped; dashed nodes are building in the open. Click a node to read it, then jump to its section — this tree works like the one in Resolve.

mediaINany footage
restore · shipped Riserestores the detail
edit · shipped Scribewrites it all down
edit · shipped Highlighterfinds the moments
edit · shipped Grabberbrings the meeting home
edit · shipped Indexknows where everything is
edit · shipped Meeting LibraryLibrary reads them together
sound · shipped Clearrescues the voice
color · shipped Hushquiets the noise
color · shipped Stenciltraces the subject
color · shipped Depthmeasures the scene
color · beta Speakgives the image its voice
deliver · shipped Pivotfollows the subject
deliver · shipped Slatemakes it official
mediaOUTevery screen
free forever · MIT
works with free resolve
local only
shows its work
honest limitations
restore · shipped

Rise. restores the detail.

SD→HD/4K reconstruction for tape-era archives and punch-in rescue — with a heatmap of what it invented.

try it by eye

Honest resampling vs. reconstruction.

4× upscale · a real 240 px crop · left = lanczos (no synthesis) · right = real-esrgandrag
Real-ESRGAN reconstruction
Lanczos resample
LANCZOS RISE (SYNTHESIZED)
● the right side is labeled synthesis because it IS synthesis eyes and hair sharpen, softness clears — and skin can go plastic; the limitations say so
what it does
  • Real-ESRGAN reconstruction, tiled and seam-free at any resolution
  • Flow-gated temporal stabilization kills per-frame shimmer
  • Interlace guard refuses combed sources instead of sharpening the combs
  • Detail heatmap + A/B wipe: see exactly what was synthesized
  • The same engine powers Pivot's punch-in enhancement
how to use it
  1. rise-cli probe tape.mov — interlace check first (combs + AI = garbage, we refuse politely).
  2. rise-cli up tape.mov --scale 2 --stabilize
  3. In Pivot: tick 'enhance punch-ins' and crops route through the same engine.
technical detail — how Rise actually works

Official Real-ESRGAN x4plus weights converted to a pinned ONNX (the conversion script ships), run through ONNX Runtime (CoreML/DirectML/CUDA) with 512 px tiles and ramped overlap blending — golden-tested seam-free vs whole-frame. Temporal stabilization warps the previous output by Farneback flow and blends where warp error is low. The engine is a frozen API consumed by both the app and Pivot; every result carries which backend actually ran, and synthesis is always labeled synthesis.

modellicensesize
Real-ESRGAN x4 (our ONNX export)BSD-3 (xinntao)64 MB
honest limitations
  • Per-frame model + temporal stabilization ≠ Topaz's temporal models on fast motion.
  • Faces can go uncanny; face restore is opt-in and off by default for journalism.

Replaces: Studio Super Scale ($295) · Topaz Video AI (~$300) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

edit · shipped

Scribe. writes it all down.

Local transcription with speakers, captions, and the paper-edit superpower: select text, get a cut timeline back.

try the paper edit

Click a word. Click another. That's your cut.

real transcript · two synthesized voices, diarized by scribe itselfclick words to mark in/out
pull list: 0
TITLE: Scribe selects FCM: NON-DROP FRAME (mark two words to write your first event)
what it does
  • Whisper-class local transcription with word-accurate timestamps
  • Speaker diarization with editable names, colored across all exports
  • Broadcast-legal captions: SRT/VTT with line-length and CPS presets
  • Marker EDL puts speaker-colored markers on your Resolve timeline
  • Paper edits: select text, export a CMX3600 selects EDL, conform
how to use it
  1. scribe-cli transcribe interview.mov --diarize
  2. Import the .srt onto a subtitle track; import the marker EDL for speaker-colored timeline markers.
  3. Highlight the quotes you want, export a selects EDL, conform in Resolve.
technical detail — how Scribe actually works

faster-whisper (CTranslate2 int8) with Silero VAD and word timestamps; diarization via sherpa-onnx (pyannote segmentation 3.0 + 3D-Speaker embeddings, clustered), majority-overlap assigned per segment. Exports are pure, golden-tested writers: SRT/VTT with word-timed re-blocking to caption presets, speaker-colored marker EDLs (Resolve's Timeline Markers From EDL), and CMX3600 selects EDLs with accumulating record TC and handles. NDF timecode throughout, 23.976 drift documented rather than hidden.

modellicensesize
Whisper large-v3-turbo (faster-whisper)MIT (OpenAI weights)1.6 GB
pyannote segmentation 3.0MIT6 MB
3D-Speaker embeddingsApache-2.040 MB
honest limitations
  • Crosstalk and heavy accents still beat it — a human captioner is the accessibility gold standard.
  • Diarization confuses similar voices in roomy recordings.
  • Post only — live captions are community-captioner's job.

Replaces: Studio Speech-to-Text ($295) · Descript / Rev / Simon Says ($15–30/mo) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

edit · shipped

Highlighter. finds the moments.

Community Highlighter: a public meeting video becomes a highlight reel — read in text, cut by the clock, analyzed like a record.

what it does
  • Reads the meeting from its own captions in seconds; Whisper (with harvested name hotwords) when there are none
  • The Meeting Analyzer: people/places/things with clips and 🔍 Investigate, framing lenses, a cross-reference network, topic heatmap, disagreements, and the town portal's own agendas and minutes
  • ✨ local highlight picks, 🤖 AI picks on your key — every pick clamped to the meeting's real clock
  • Two doors out: a share link the deployed web player opens (nothing uploads), or a staged download-and-stitch MP4 with title cards
  • Full report as markdown + selectable-text PDF; summary and whole-transcript translation in ten languages
how to use it
  1. Paste a meeting URL (or open a local recording) — a 7-hour meeting reads in seconds via its own captions.
  2. Mark moments from the transcript, the search, or the ✨ buttons; they land on a timeline with per-clip trims, speed, and fades.
  3. Export a share link the web player opens, or download-and-cut one MP4 with title cards — only the kept seconds leave YouTube.
technical detail — how Highlighter actually works

Ingest races the watch page against yt-dlp for captions and takes the first winner; transcripts live as sidecars any tool can read. The insight engine is counted, not modeled: extractive briefs, pattern entities with conservative spelling folds, word-boundary framing lenses, sentence-level co-occurrence networks. Exports ride one ffmpeg concat graph — cuts, retimes, fades, and Pillow title cards in a single pass. The town's paper trail comes from its own CivicClerk portal, matched by the meeting's date and name.

honest limitations
  • YouTube gates caption delivery by IP reputation; the suite routes around it honestly (your proxy, or the community caption relay) but a gated bare IP may need one.
  • The AI reel, report narrative, and translations need your own API key — no key ships, nothing requires one.

Replaces: an afternoon of scrubbing (hours per meeting) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

edit · shipped

Grabber. brings the meeting home.

BIG Video Grabber: find a town's recordings where they actually live — CivicClerk portals, zoomgov shares, YouTube — and land them ready to edit.

what it does
  • CivicClerk portal reader — events, recordings, and the published agenda/packet/minutes PDFs
  • A zoomgov share-link resolver of our own (four requests, cookie jar, honest failures)
  • yt-dlp underneath — the other thousand sites, with a nightly self-update on every page open
  • Quality ladder downloads with span-only fetches — only the seconds you keep leave the host
how to use it
  1. Type the town's CivicClerk tenant (brooklinema is the part before .api.civicclerk.com) and a date range.
  2. Every event lists its links — the recording, and the agenda/minutes PDFs the portal publishes.
  3. Fetch at the quality you need; the ladder falls back honestly (best → 4K → … → audio-only).
technical detail — how Grabber actually works

The portal reader speaks CivicClerk's OData dialect defensively: every URL-shaped string in an event is harvested with its field path, video-ish links sorted first, publishedFiles resolved through the portal's own GetMeetingFileStream endpoint. The zoomgov resolver walks share page → meeting id → play info → mp4 with the cookie jar and Referer Zoom expects. Downloads ride the suite's managed yt-dlp nightly, which never routes its own self-update through your proxy.

honest limitations
  • zoomgov.com isn't supported by yt-dlp — the built-in resolver walks Zoom's share pages itself, and when Zoom redecorates, each step fails naming itself.
  • CivicClerk tenants disagree about where recordings hide; the reader harvests every URL-shaped field and shows its work rather than guessing silently.

Replaces: screen recording the player (quality + hours) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

edit · shipped

Index. knows where everything is.

Your footage archive, searchable in plain words — scan folders, search what was said, and jump into the editor with markers.

what it does
  • Folder scans that respect your structure — nothing moves, nothing uploads
  • Plain-word speech search across every transcript sidecar on the shelf
  • FCPXML marker export — search hits become editor markers
  • One-click handoff of any meeting to the Community Highlighter
how to use it
  1. Point Index at a folder; it scans what's there and reads any transcript sidecars beside the files.
  2. Search in plain words — a name, a street, a phrase — and get files with the moments that say it.
  3. Export FCPXML markers and open the hits in your editor, or hand a meeting straight to the Highlighter.
technical detail — how Index actually works

Index is deliberately boring: a filesystem scan plus the .scribe.json sidecars the other tools already write, queried with plain substring and word matches. The interesting part is the contract — because Scribe, the Highlighter, and Index share one sidecar format, any file transcribed anywhere becomes searchable everywhere, and the FCPXML export turns hits into markers Resolve reads natively.

honest limitations
  • Speech search needs words on disk — run Scribe (or the Highlighter's caption read) once per file first.
  • It indexes filenames, metadata, and transcripts — not what's visually IN the frame.

Replaces: MAM/DAM systems ($1k–5k/yr) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

edit · shipped

Meeting Library. Library reads them together.

Every meeting on your machine, read as one record — how the framing moves, who keeps appearing, how any topic travels, and reels cut across meetings.

what it does
  • Framing across meetings — eight civic lenses, each meeting a column, trends called out
  • Topic evolution and entity tracking, with caption misspellings folded into one name that shows its work
  • Meeting comparison side by side — shared topics and names outlined against what each carries alone
  • Discourse analysis — one term traced through every meeting at a per-1k-word rate, receipts included
  • The Montage Maker — a reel cut ACROSS meetings; URL sessions fetch only the picked seconds
how to use it
  1. Open Library in the suite — every meeting the Highlighter has read is already on the shelf.
  2. Read the grids — framing across meetings, topic evolution, entity tracking — and click anything to trace it through the full transcripts.
  3. Pick moments from any trace onto the montage tray and render one reel across meetings, each clip carded with its own meeting's name.
technical detail — how Meeting Library actually works

The Library is an aggregation, not a database: each meeting's insight sidecar (counted framing, topics, entities, pace) is read through the same door the analyzer uses, then folded — entity spellings join under a conservative match (same word count, same initials, high sequence ratio), discourse rates normalize per thousand words so a seven-hour meeting can't out-shout a one-hour one, and the montage stitcher letterboxes mixed resolutions into one frame with a title card naming each clip's meeting.

honest limitations
  • Cross-meeting cards need at least two read meetings; a meeting without words is listed as unread, not invented.
  • The optional AI read sends only the counted digest — dates, tallies, names — never a transcript, and only on your own key.

Replaces: institutional memory (someone's whole job) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

sound · shipped

Clear. rescues the voice.

Voice isolation, de-hum, de-click, room-tone matching, and loudness to spec — the RX-shaped hole in every community edit.

try it by ear

Hear the rescue — and hear what was removed.

synthesized room · a "voice", HVAC-ish noise, and 60 Hz mains hum — all generated in your browserwebaudio · nothing uploaded
hum: present (60/120/180 Hz) room kept: 100% ● the residual button is the whole philosophy: if the voice shows up in there, back off

The real tool runs DeepFilterNet 3 for isolation (this demo just crossfades the synthetic room), auto-detects the hum fundamental, and defaults to keeping 35% of the room — full-wet rarely sounds human.

what it does
  • DeepFilterNet 3 voice isolation with an honest mix-back default (35% room)
  • Auto de-hum: finds 50/60 Hz and notches the whole harmonic series
  • Second-difference de-click that refuses to eat sibilants
  • Room-tone synthesis from a 2-second profile — endless matching tone
  • Loudness to broadcast/podcast/streaming spec, with true-peak honesty
how to use it
  1. clear-cli process interview.wav --isolate 0.65 --loudness broadcast
  2. Auto-detects mains hum (50/60 Hz + harmonics) and notches it.
  3. clear-cli roomtone interview.wav → 30 seconds of matching room tone for your gap fills.
technical detail — how Clear actually works

Isolation runs the official DeepFilterNet 3 binary as a subprocess (48 kHz, real-time class) with our mix-back blend. De-hum detects the mains fundamental by harmonic-sum test against the spectral floor, then zero-phase notch cascades. De-click flags second-difference outliers (no IIR ringing), interpolates runs ≤ 2 ms, and bails honestly if more than 1% of samples flag. Room tone is random-phase resynthesis of a Welch PSD profile — exact spectral envelope, loop-safe tail. Loudness is BS.1770 via pyloudnorm; if the true-peak ceiling fights the target, it says so instead of limiting.

modellicensesize
DeepFilterNet 3 (official binary)MIT/Apache dual60 MB
honest limitations
  • Full-wet isolation sounds processed — the default keeps 35% of the room, on purpose.
  • No spectral repair painting (a cough under a word is still RX territory).
  • Loudness normalize won't limit: if your peaks argue with the target, it says so instead of squashing.

Replaces: Studio Voice Isolation ($295) · iZotope RX ($400–1,200) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

color · shipped

Hush. quiets the noise.

Measured, transparent noise reduction — the tool free-Resolve users are missing most.

try it by eye

Real footage, default settings.

validation scene · 5% gaussian noise · 26.0 → 39.0 dB PSNRdrag
Denoised output
Noisy input
NOISY IN HUSH OUT
● hard edges, fine checker texture and 2-px lines survive
what it does
  • Auto Setup measures the clip and writes every slider — one undo reverts it all
  • Motion-tracked temporal NR (~±8 px) with a hard anti-ghosting knee
  • Noise EQ cuts noise by size: fine grain, clumps, stains, color blotches
  • Live scopes and a Noise Removed view — watch exactly what's being taken
  • Exports a clean-confidence matte Speak (and your grade) can key on
how to use it
  1. Drop Hush Open NR on a node and click Auto Setup — it measures the clip and dials in every slider.
  2. Check the measurement: Step 1 → Scope: Measurements.
  3. Check what's being removed: Step 5 → View → Noise Removed. Featureless static = good.
  4. Set View back to Result. Done.
technical detail — how Hush actually works

Three noise estimators run per frame (temporal difference, fine and coarse Laplacian) into robust histograms, producing separate spatial and temporal sigmas plus a 16-band brightness gain curve. The temporal stage compares 3×3 patches against up to ±3 frames with a hierarchical ±8 px motion search; a hard-knee gate reaches exactly zero past the knee, so mismatches cannot ghost by construction. Spatial cleanup is noise-adaptive NLM with a size-banded Noise EQ, bounded by Detail Rescue. The refine stack (v3.6) rebuilds optical character: coring, acutance, chroma-speckle, and film-matched grain. GPU kernels (Metal/CUDA/OpenCL) are line-by-line ports of one CPU reference, held to ~2×10⁻⁵ agreement by a parity test suite.

honest limitations
  • No motion-compensated temporal NR — Studio keeps an edge on fast pans.
  • Classical (NLM), not AI: very competitive at normal noise, less magical on starlight footage.

Replaces: Resolve Studio NR palette ($295) · Neat Video (~$160) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

color · shipped

Stencil. traces the subject.

Click an object once, get a matte for the whole shot — delivered as matte clips any page of free Resolve can use.

try the qc loop

The matte tells you where it's guessing.

simulated shot · one click on frame 1 → propagated matte · amber = model unsurescrub · click an amber frame to re-prompt
frame 1/24 confidence 0.97 ● scrub across the strip — the subject motion-blurs mid-shot and confidence drops
what it does
  • One click on any frame propagates a matte across the whole shot (SAM 2.1)
  • Confidence timeline glows amber exactly where the model is unsure
  • Corrective clicks re-propagate — QC the minutes that need it
  • Grow / feather / de-flicker post chain, tuned for grading mattes
  • Exports luma mattes and ProRes 4444 alpha that free Resolve just uses
how to use it
  1. Point at the thing: one include click (add excludes if needed).
  2. Propagate — SAM 2.1 tracks it across the shot on your GPU.
  3. Export a luma matte (Color page → Add Matte) or ProRes 4444 with alpha.
technical detail — how Stencil actually works

SAM 2.1 hiera-small propagates prompted masklets bidirectionally through a shot at 720p analysis resolution (PyTorch on Metal/CUDA; ONNX port planned). Confidence per frame is the model's mean in-mask probability — the amber QC strip. Post chain: 3-frame morphological majority (de-flicker), despeckle, grow/shrink, gaussian feather; mattes upsample to source res on export. Shot-bounded by the shared cut detector so propagation never crosses an edit.

modellicensesize
SAM 2.1 hiera-smallApache-2.0 (Meta)176 MB
honest limitations
  • Soft-matte quality on hair and motion blur — it's a roto assistant, not a keyer.
  • Heavyweight install (PyTorch) until the ONNX port lands.
  • Prompt granularity is real: click a face, get a face; click a jacket, get a jacket.

Replaces: Studio Magic Mask ($295) · Mocha Pro (~$700) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

color · shipped

Depth. measures the scene.

A depth matte for every clip — fog, rack focus, and depth-graded atmosphere through a paste-in Fusion template pack.

try the probe

Move your cursor. Read the scene.

real depth · a street canyon, measured by the shipping enginehover to probe · slider adds fog
depth under cursor:
● fog is depth-shaped: the far end of the street mists first, the near towers stay clear — this is the Fog template from the pack, in your browser
what it does
  • A relative depth matte for every clip, temporally smoothed per shot
  • Edge-guided upsampling keeps depth edges locked to image edges
  • 10-bit gray ProRes mattes the Color page keys through directly
  • Five paste-in Fusion templates: fog, rack focus, depth grade, parallax, haze
  • False-color probe view with per-shot range readouts
how to use it
  1. depth-cli run scene.mov → a 10-bit gray depth matte clip beside your footage.
  2. depth-cli templates → five paste-ready Fusion setups: Fog, Rack Focus, Depth Grade, Parallax 2.5D, Haze Light.
  3. Color page: Add Matte, key through the depth bands.
technical detail — how Depth actually works

MiDaS-small (256 px, MIT) behind a model-agnostic engine — Video-Depth-Anything-Small is the planned v0.2 backend. Temporal EMA resets at every cut; per-shot normalization uses 2/98 percentiles so hot pixels can't crush the range. Upsampling is a He-style guided filter against the source luma, keeping depth edges on image edges. Export packs 0–1 depth into 10-bit gray ProRes 4444 (near = white by default, invertible, gamma-mappable).

modellicensesize
MiDaS-smallMIT (Intel ISL)64 MB
honest limitations
  • Relative depth, not metric — mattes and atmosphere, not measurement.
  • Mirrors, glass, and scale tricks fool it.
  • v0.1 backend is per-frame with temporal smoothing; the video-native model is planned.

Replaces: Studio Depth Map + Relight ($295) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

color · beta

Speak. gives the image its voice.

Film character for the last node: Hurter–Driffield tone curves, subtractive color, halation and grain — with every curve on screen.

the pair, in the tree

Denoise early. Reconstruct late.

node 1MediaInyour clip
node 2HushNR · first
node 3Gradeyours
node 4Speakfilm · last
deliverOutRec.709 / P3 / HDR
davinci wide gamut · one look, every deliverable
hush's matte rides downstream → speak lays grain where the image was cleaned most
what it does
  • Real Hurter–Driffield curves: negative stage, printer-light timing in genuine printer points, print stage
  • Subtractive color in the density domain — highlight chroma self-compresses toward paper white
  • Halation injected as exposure BEFORE the curve, so blown highlights go white-hot in the core and red at the edges
  • Grain as density noise per dye layer: loud in shadows, grainless at paper white, never negative
  • The Hush handoff — grain lands exactly where the denoiser cleaned deepest
how to use it
  1. Install Speak, then put it on the LAST node of your grade — Hush opens the tree, Speak closes it.
  2. Pick a tone scale and printer lights; the H&D scope plots the exact curve your pixels are using.
  3. For the handoff: enable "Export Clean Matte to Alpha" in Hush ≥ 3.7 and "Use Incoming Matte" in Speak's grain.
technical detail — how Speak actually works

Speak models the photochemical chain rather than filtering the picture. A negative stage applies closed-form Hurter–Driffield characteristic curves, printer lights time the color in genuine printer points, and a print stage lands the result — all plotted live from the same math the pixels take, so the scope cannot disagree with the render. Saturation happens in the density domain (dye density adds where light multiplies), with an inter-image coupler for dye cross-talk. Halation is injected as exposure before the curve on a multi-scale scatter field with a power-law skirt, making it physically self-limiting. Grain is per-RGB-layer density noise, bandpassed at a physical pitch and boiling every frame, optionally keyed to Hush's clean-confidence matte. Metal on macOS, OpenCL on Windows/Linux, all agreeing with a CPU reference to ~2e-5; identity at default and alpha always passes through untouched.

honest limitations
  • Early beta (v0.2): machine-verified module by module, but not yet field-tested through a full Resolve grade — don't cut a feature on it this week.
  • macOS binaries only so far; the Windows/OpenCL build follows.
  • No film-stock preset library yet — you build the look from the curves themselves.

Replaces: Dehancer (~$380) · FilmConvert Nitrate (~$140) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

deliver · shipped

Pivot. follows the subject.

Auto-reframe your 16:9 masters to 9:16, 1:1, 4:5 — with an editor's eye, not a cropping algorithm's.

try the actual solver

Drag the subject. The camera is Pivot.

live playground · the same math that ships: deadzone · jerk-limited moves · no overshootdrag the dot
camera: holding moves: 0 ● amber dot = subject · slate box = the 9:16 crop · shaded band = deadzone

The offline app also looks 12 frames ahead (a luxury a live demo can't have), so real renders start their moves slightly early — like an operator who read the script.

what it does
  • Shot-aware subject tracking: faces first, whole people as fallback
  • An editor's camera: deadzone hold, eased jerk-limited moves, no drift
  • Path-trace view shows every detection and every camera decision
  • Per-shot overrides: punch, follow, or center — re-solved instantly
  • Renders natively or exports keyframed Fusion .setting for free Resolve
how to use it
  1. Open a clip, pick your aspects, hit Analyze — shots, faces, and a solved camera path.
  2. Scrub the path trace; override any shot to punch / follow / center.
  3. Render natively (no silent upscale) or export a keyframed .setting for the Fusion page.
technical detail — how Pivot actually works

Analysis decodes once at 360p: shot detection (adaptive luma-diff), YuNet face detection every 2nd frame, YOLOX-s persons only where faces are absent, greedy IoU/distance tracking. Per shot, a subject is selected (size × centrality × persistence) and solved: shots under 2 s or under the motion floor become static punches at a trimmed median; the rest get the follow controller — deadzone with hysteresis, 12-frame lookahead anticipation, braking-parabola velocity setpoint clamped by v-max/a-max. Overshoot is pinned ≤ 4 acceleration quanta by golden tests. Renders are frame-accurate (EOF flush handled) with audio stream-copied, never re-encoded.

modellicensesize
YuNet face detectionMIT (OpenCV Zoo)0.2 MB
YOLOX-s person detectionApache-2.0 (Megvii)34 MB
honest limitations
  • Faces and people only so far — sports and wildlife will miss (saliency model on the roadmap).
  • Punch-ins past native resolution are honest resampling until you enable the Rise engine.

Replaces: Studio Smart Reframe ($295) · auto-clip SaaS tools ($15+/mo) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

deliver · shipped

Slate. makes it official.

Station graphics kit: lower thirds, slates, bars & tone, countdowns — typeset once, rendered broadcast-clean with real alpha.

what it does
  • Lower thirds, full slates, countdown clocks, and bars & tone from one panel
  • ProRes 4444 out with straight alpha (yuva444p12le) — keys clean over anything
  • PNG stills and looping GIFs for web and social from the same template
  • The suite's font discovery — your station typeface, not a bundled lookalike
how to use it
  1. Open Slate in the suite and pick a template — lower third, slate, countdown, or bars.
  2. Type the names; the preview is the render.
  3. Render ProRes 4444 (alpha), PNG, or GIF and drop it in the editor's bin.
technical detail — how Slate actually works

Typography renders through Pillow at the output's own pixel size (no scaling pass), composited to premultiplied-safe straight alpha, then encoded by the suite's LGPL ffmpeg: ProRes 4444 for the editor, PNG for stills, palette-optimized GIF for the web. Countdowns render per-frame from the same layout engine, so the type never shifts between seconds.

honest limitations
  • Static graphics and countdowns — kinetic type still belongs to Resolve's Text+.
  • Font discovery uses the fonts your system already has; it won't fetch webfonts.

Replaces: hand-rolled every time (hours) — honest ballparks; the paid tools are good tools, you just shouldn't need them.

↑ back to the pipeline

design values

What we refuse to compromise.

the premiseFree is the feature.

MIT-licensed, commercial use included. No tiers, trials, seats, accounts, or "free for now." A price of zero is a design constraint we build around, not a promotion.

the boundaryLocal is a right.

Footage and audio never leave your machine. Models run on your hardware, downloaded once with their license shown and their hash verified. Nothing phones home — there is no home to phone.

the methodShow the work.

Every tool renders its own measurements: noise scopes, camera paths, confidence strips, residual monitors, depth probes. Glass boxes build skill; black boxes build dependence.

the voiceHonest about limits.

Each tool names what the paid equivalent still does better — in writing, on this page. Synthesis is labeled synthesis. A tool that can't admit its losses can't be trusted with your wins.

the doorEasy first, deep later.

Three clicks to a great default; every parameter still reachable two clicks deeper. Auto-setup writes into visible controls you can override — automation as a teacher, never a ceiling.

the rootBuilt with community media.

Proven on member footage, tape vaults, and city meetings at a working access station before release. If it doesn't survive a Tuesday-night edit bay, it doesn't ship.

under the hood

Architecture, strategy, and proof.

Two engine families, one covenant: plugins that live inside Resolve's node tree, and standalone tools that hand it files it already understands.

control-z/ ├─ OpenFX · C++ & GPU Hush · Speak — one CPU reference implementation, │ three GPU ports (Metal / CUDA / OpenCL) held to │ ~2×10⁻⁵ agreement by a parity test suite ├─ standalone · Python Pivot · Stencil · Scribe · Clear · Rise · Depth │ └─ czcore shot detection · media IO · hash-pinned model │ store · exports · app shell shared by every tool ├─ interchange SRT · marker & selects EDL · keyframed Fusion │ .setting · ProRes matte clips · sidecar JSON — │ formats the free version of Resolve just reads └─ suite app · next one window, scopes everywhere, batch queue, OpenFX installer, ProRes up to 4444 out
modelsPermissive or nothing.

Every AI model is MIT/Apache/BSD-licensed, mirrored with a pinned SHA-256, and stored once for the whole suite. Non-commercial and login-gated checkpoints are rejected on principle — the model card on each tool says exactly what runs.

proofTested like broadcast gear.

Golden tests pin the math (solver physics, notch depths, timecode); GPU ports are parity-tested against a single reference; every tool is verified on real 4K footage before it's called working. The CHANGELOG says what's honest and what's deferred.

strategyEngines compound.

Rise's upscaler already powers Pivot's punch-ins; Scribe's speech stack will power Minutes and Babel; Hush's texture math becomes Speak. Each wave of tools makes the next one cheaper — that's the whole roadmap.

the rest of the free stack

Other paywall-undoers we recommend.

control-z only builds what doesn't already exist for free. These do other paywalled jobs, excellently — link them, fund them, love them.

From our own shop, outside the suite: community-captioner — free browser-based live captions for OBS (Scribe's live sibling). Ask at Community AI.

Start with Hush. Stay for the pipeline.