Journalists, artists, advocates, new filmmakers, and community media stations shouldn't have quality gated behind a paywall. If the content is good, the form should match. Yet the people with the most important stories to tell often have the least access to the tools that make work land — so this suite catches you up, free and open-source, from the technical to the editorial to the creative, built for the best free editor around: DaVinci Resolve.
Plugins that live inside Resolve's node tree, apps that run beside any editor, and the scaffolding to put them to work. All free, all MIT.
Ten tools in one window — scopes on every panel, a batch queue that survives a whole tape shelf, ProRes 4444 out, an OpenFX installer for Hush and Speak built in, and the Meeting Library reading every meeting on your machine as one record. The workbench on either side of your edit.
Two halves of one idea: Hush takes the noise out at the top of the tree, Speak puts the optical character back at the bottom. Clean early, reconstruct late.
One page tells the whole story — the buttons open it; the small links jump straight to each release. macOS and Windows for Hush; macOS for Speak so far. Hush hands its clean-confidence matte to Speak's grain — the handoff, working.
The structure, not the look: a working grade's skeleton — noise reduction, exposure, contrast, balance, sat, a foreground/background split, and a look — wired and empty. Its noise and look stages are compound nodes holding a Studio-or-free choice: a Studio feature if you have it, or the free combo that matches it. Hush and Speak drop onto the free paths.
See the tree + how to use it →A single control point pinned to the exact position of 18% middle gray in DaVinci Wide Gamut / Intermediate (code 344, not the middle of the curve). Sits on the diagonal, so it changes nothing — until you shape contrast around it and your mid-gray stays put instead of drifting darker every time you add punch. Drop it in the node tree's Contrast stage.
Why it matters + how to use it →Ten paste-ready Fusion setups that turn a matte into finished work: privacy-blur inside a tracked Stencil matte, composite on an alpha, mist a shot's far end, rack to a focal plane you didn't shoot, grade near/mid/far separately, 9:16 with a blurred backdrop. Each one paste-tested in a live Fusion comp; built from stock nodes — nothing needs Studio.
See what each does + download →The white paper behind the suite's thesis — measurement, texture, and the case for tools that show their work. It ends where Hush and Speak begin.
Read the study↗1 · Install Hush (and Speak, when it ships) with the one-click installer — they
appear in OpenFX on the Color, Edit and Fusion pages.
2 · Run the standalone tools before or after your edit; they hand results back as
files Resolve just imports: SRT subtitles, marker & selects EDLs, matte clips,
keyframed Fusion .settings.
3 · Every tool section below ends with its exact Resolve recipe.
The standalone tools are editor-agnostic: Scribe's SRT and EDLs, Clear's cleaned
WAVs, Rise/Pivot/Stencil/Depth's ProRes renders and mattes drop into Premiere,
Final Cut, or Avid the same way.
The honest exception: Hush and Speak are OpenFX plugins — Premiere doesn't host
OpenFX, so run noisy footage through the Suite app instead (or grade in free
Resolve; it costs the same as this website).
control-z comes out of a working access station — every tool is proven on member footage, tape vaults, and city meetings at Brookline Interactive Group first. Ten donated seats, zero license dollars: captions for compliance (Scribe), rescued rooms (Clear), archives upscaled (Rise), every show vertical (Pivot).
Transcribe the presser on your laptop — not a cloud service (Scribe). Pull the quote as a cut list. Rescue the scanner-room audio (Clear). Matte a face for protection (Stencil). Nothing leaves your machine, which is not a feature so much as a duty of care.
Hush then Speak is a grade sandwich that used to cost $500 in plugins. Depth fakes the rack focus you didn't shoot. Rise saves the archival pickup. Pivot cuts the vertical trailer while you sleep. All of it in the free Resolve you already edit in.
Every tool exposes its measurements — which means every tool can be played. Grade through depth mattes, key through noise-confidence, push Rise until reality gets soft. And when the work is advocacy, form is credibility: it deserves both.
Restore → edit → sound → color → deliver. Lit nodes are shipped; dashed nodes are building in the open. Click a node to read it, then jump to its section — this tree works like the one in Resolve.
SD→HD/4K reconstruction for tape-era archives and punch-in rescue — with a heatmap of what it invented.
Official Real-ESRGAN x4plus weights converted to a pinned ONNX (the conversion script ships), run through ONNX Runtime (CoreML/DirectML/CUDA) with 512 px tiles and ramped overlap blending — golden-tested seam-free vs whole-frame. Temporal stabilization warps the previous output by Farneback flow and blends where warp error is low. The engine is a frozen API consumed by both the app and Pivot; every result carries which backend actually ran, and synthesis is always labeled synthesis.
| model | license | size |
|---|---|---|
| Real-ESRGAN x4 (our ONNX export) | BSD-3 (xinntao) | 64 MB |
Replaces: Studio Super Scale ($295) · Topaz Video AI (~$300) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Local transcription with speakers, captions, and the paper-edit superpower: select text, get a cut timeline back.
faster-whisper (CTranslate2 int8) with Silero VAD and word timestamps; diarization via sherpa-onnx (pyannote segmentation 3.0 + 3D-Speaker embeddings, clustered), majority-overlap assigned per segment. Exports are pure, golden-tested writers: SRT/VTT with word-timed re-blocking to caption presets, speaker-colored marker EDLs (Resolve's Timeline Markers From EDL), and CMX3600 selects EDLs with accumulating record TC and handles. NDF timecode throughout, 23.976 drift documented rather than hidden.
| model | license | size |
|---|---|---|
| Whisper large-v3-turbo (faster-whisper) | MIT (OpenAI weights) | 1.6 GB |
| pyannote segmentation 3.0 | MIT | 6 MB |
| 3D-Speaker embeddings | Apache-2.0 | 40 MB |
Replaces: Studio Speech-to-Text ($295) · Descript / Rev / Simon Says ($15–30/mo) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Community Highlighter: a public meeting video becomes a highlight reel — read in text, cut by the clock, analyzed like a record.
Ingest races the watch page against yt-dlp for captions and takes the first winner; transcripts live as sidecars any tool can read. The insight engine is counted, not modeled: extractive briefs, pattern entities with conservative spelling folds, word-boundary framing lenses, sentence-level co-occurrence networks. Exports ride one ffmpeg concat graph — cuts, retimes, fades, and Pillow title cards in a single pass. The town's paper trail comes from its own CivicClerk portal, matched by the meeting's date and name.
Replaces: an afternoon of scrubbing (hours per meeting) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
BIG Video Grabber: find a town's recordings where they actually live — CivicClerk portals, zoomgov shares, YouTube — and land them ready to edit.
The portal reader speaks CivicClerk's OData dialect defensively: every URL-shaped string in an event is harvested with its field path, video-ish links sorted first, publishedFiles resolved through the portal's own GetMeetingFileStream endpoint. The zoomgov resolver walks share page → meeting id → play info → mp4 with the cookie jar and Referer Zoom expects. Downloads ride the suite's managed yt-dlp nightly, which never routes its own self-update through your proxy.
Replaces: screen recording the player (quality + hours) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Your footage archive, searchable in plain words — scan folders, search what was said, and jump into the editor with markers.
Index is deliberately boring: a filesystem scan plus the .scribe.json sidecars the other tools already write, queried with plain substring and word matches. The interesting part is the contract — because Scribe, the Highlighter, and Index share one sidecar format, any file transcribed anywhere becomes searchable everywhere, and the FCPXML export turns hits into markers Resolve reads natively.
Replaces: MAM/DAM systems ($1k–5k/yr) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Every meeting on your machine, read as one record — how the framing moves, who keeps appearing, how any topic travels, and reels cut across meetings.
The Library is an aggregation, not a database: each meeting's insight sidecar (counted framing, topics, entities, pace) is read through the same door the analyzer uses, then folded — entity spellings join under a conservative match (same word count, same initials, high sequence ratio), discourse rates normalize per thousand words so a seven-hour meeting can't out-shout a one-hour one, and the montage stitcher letterboxes mixed resolutions into one frame with a title card naming each clip's meeting.
Replaces: institutional memory (someone's whole job) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Voice isolation, de-hum, de-click, room-tone matching, and loudness to spec — the RX-shaped hole in every community edit.
The real tool runs DeepFilterNet 3 for isolation (this demo just crossfades the synthetic room), auto-detects the hum fundamental, and defaults to keeping 35% of the room — full-wet rarely sounds human.
Isolation runs the official DeepFilterNet 3 binary as a subprocess (48 kHz, real-time class) with our mix-back blend. De-hum detects the mains fundamental by harmonic-sum test against the spectral floor, then zero-phase notch cascades. De-click flags second-difference outliers (no IIR ringing), interpolates runs ≤ 2 ms, and bails honestly if more than 1% of samples flag. Room tone is random-phase resynthesis of a Welch PSD profile — exact spectral envelope, loop-safe tail. Loudness is BS.1770 via pyloudnorm; if the true-peak ceiling fights the target, it says so instead of limiting.
| model | license | size |
|---|---|---|
| DeepFilterNet 3 (official binary) | MIT/Apache dual | 60 MB |
Replaces: Studio Voice Isolation ($295) · iZotope RX ($400–1,200) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Measured, transparent noise reduction — the tool free-Resolve users are missing most.
Three noise estimators run per frame (temporal difference, fine and coarse Laplacian) into robust histograms, producing separate spatial and temporal sigmas plus a 16-band brightness gain curve. The temporal stage compares 3×3 patches against up to ±3 frames with a hierarchical ±8 px motion search; a hard-knee gate reaches exactly zero past the knee, so mismatches cannot ghost by construction. Spatial cleanup is noise-adaptive NLM with a size-banded Noise EQ, bounded by Detail Rescue. The refine stack (v3.6) rebuilds optical character: coring, acutance, chroma-speckle, and film-matched grain. GPU kernels (Metal/CUDA/OpenCL) are line-by-line ports of one CPU reference, held to ~2×10⁻⁵ agreement by a parity test suite.
Replaces: Resolve Studio NR palette ($295) · Neat Video (~$160) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Click an object once, get a matte for the whole shot — delivered as matte clips any page of free Resolve can use.
SAM 2.1 hiera-small propagates prompted masklets bidirectionally through a shot at 720p analysis resolution (PyTorch on Metal/CUDA; ONNX port planned). Confidence per frame is the model's mean in-mask probability — the amber QC strip. Post chain: 3-frame morphological majority (de-flicker), despeckle, grow/shrink, gaussian feather; mattes upsample to source res on export. Shot-bounded by the shared cut detector so propagation never crosses an edit.
| model | license | size |
|---|---|---|
| SAM 2.1 hiera-small | Apache-2.0 (Meta) | 176 MB |
Replaces: Studio Magic Mask ($295) · Mocha Pro (~$700) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
A depth matte for every clip — fog, rack focus, and depth-graded atmosphere through a paste-in Fusion template pack.
MiDaS-small (256 px, MIT) behind a model-agnostic engine — Video-Depth-Anything-Small is the planned v0.2 backend. Temporal EMA resets at every cut; per-shot normalization uses 2/98 percentiles so hot pixels can't crush the range. Upsampling is a He-style guided filter against the source luma, keeping depth edges on image edges. Export packs 0–1 depth into 10-bit gray ProRes 4444 (near = white by default, invertible, gamma-mappable).
| model | license | size |
|---|---|---|
| MiDaS-small | MIT (Intel ISL) | 64 MB |
Replaces: Studio Depth Map + Relight ($295) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Film character for the last node: Hurter–Driffield tone curves, subtractive color, halation and grain — with every curve on screen.
Speak models the photochemical chain rather than filtering the picture. A negative stage applies closed-form Hurter–Driffield characteristic curves, printer lights time the color in genuine printer points, and a print stage lands the result — all plotted live from the same math the pixels take, so the scope cannot disagree with the render. Saturation happens in the density domain (dye density adds where light multiplies), with an inter-image coupler for dye cross-talk. Halation is injected as exposure before the curve on a multi-scale scatter field with a power-law skirt, making it physically self-limiting. Grain is per-RGB-layer density noise, bandpassed at a physical pitch and boiling every frame, optionally keyed to Hush's clean-confidence matte. Metal on macOS, OpenCL on Windows/Linux, all agreeing with a CPU reference to ~2e-5; identity at default and alpha always passes through untouched.
Replaces: Dehancer (~$380) · FilmConvert Nitrate (~$140) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Auto-reframe your 16:9 masters to 9:16, 1:1, 4:5 — with an editor's eye, not a cropping algorithm's.
The offline app also looks 12 frames ahead (a luxury a live demo can't have), so real renders start their moves slightly early — like an operator who read the script.
Analysis decodes once at 360p: shot detection (adaptive luma-diff), YuNet face detection every 2nd frame, YOLOX-s persons only where faces are absent, greedy IoU/distance tracking. Per shot, a subject is selected (size × centrality × persistence) and solved: shots under 2 s or under the motion floor become static punches at a trimmed median; the rest get the follow controller — deadzone with hysteresis, 12-frame lookahead anticipation, braking-parabola velocity setpoint clamped by v-max/a-max. Overshoot is pinned ≤ 4 acceleration quanta by golden tests. Renders are frame-accurate (EOF flush handled) with audio stream-copied, never re-encoded.
| model | license | size |
|---|---|---|
| YuNet face detection | MIT (OpenCV Zoo) | 0.2 MB |
| YOLOX-s person detection | Apache-2.0 (Megvii) | 34 MB |
Replaces: Studio Smart Reframe ($295) · auto-clip SaaS tools ($15+/mo) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
Station graphics kit: lower thirds, slates, bars & tone, countdowns — typeset once, rendered broadcast-clean with real alpha.
Typography renders through Pillow at the output's own pixel size (no scaling pass), composited to premultiplied-safe straight alpha, then encoded by the suite's LGPL ffmpeg: ProRes 4444 for the editor, PNG for stills, palette-optimized GIF for the web. Countdowns render per-frame from the same layout engine, so the type never shifts between seconds.
Replaces: hand-rolled every time (hours) — honest ballparks; the paid tools are good tools, you just shouldn't need them.
MIT-licensed, commercial use included. No tiers, trials, seats, accounts, or "free for now." A price of zero is a design constraint we build around, not a promotion.
Footage and audio never leave your machine. Models run on your hardware, downloaded once with their license shown and their hash verified. Nothing phones home — there is no home to phone.
Every tool renders its own measurements: noise scopes, camera paths, confidence strips, residual monitors, depth probes. Glass boxes build skill; black boxes build dependence.
Each tool names what the paid equivalent still does better — in writing, on this page. Synthesis is labeled synthesis. A tool that can't admit its losses can't be trusted with your wins.
Three clicks to a great default; every parameter still reachable two clicks deeper. Auto-setup writes into visible controls you can override — automation as a teacher, never a ceiling.
Proven on member footage, tape vaults, and city meetings at a working access station before release. If it doesn't survive a Tuesday-night edit bay, it doesn't ship.
Two engine families, one covenant: plugins that live inside Resolve's node tree, and standalone tools that hand it files it already understands.
Every AI model is MIT/Apache/BSD-licensed, mirrored with a pinned SHA-256, and stored once for the whole suite. Non-commercial and login-gated checkpoints are rejected on principle — the model card on each tool says exactly what runs.
Golden tests pin the math (solver physics, notch depths, timecode); GPU ports are parity-tested against a single reference; every tool is verified on real 4K footage before it's called working. The CHANGELOG says what's honest and what's deferred.
Rise's upscaler already powers Pivot's punch-ins; Scribe's speech stack will power Minutes and Babel; Hush's texture math becomes Speak. Each wave of tools makes the next one cheaper — that's the whole roadmap.
control-z only builds what doesn't already exist for free. These do other paywalled jobs, excellently — link them, fund them, love them.
The transcoder. Batch encodes that used to be a Compressor license.
Shutter EncoderThe swiss-army converter: ProRes proxies, rewraps, burn-ins, conforms.
LosslessCutInstant trims with zero re-encode — the tape log bench, reborn.
OBS StudioBroadcast-grade capture, switching and streaming. The Tricaster tax, undone.
AudacityQuick audio surgery on anything, since forever.
NatronNode compositing in the Nuke idiom — and it hosts OpenFX, so Hush runs there too.
Blender3D, tracking, and a real compositor. The whole VFX department, donated.
darktableRaw photo developing that undoes the Lightroom subscription.
KdenliveA capable, honest NLE for machines Resolve won't run on.
From our own shop, outside the suite: community-captioner — free browser-based live captions for OBS (Scribe's live sibling). Ask at Community AI.