The wedge: one recording produces a shareable video and an accurate step-by-step SOP at the same time, cross-linked, so nobody has to turn one into the other afterwards.
Loom gives you a video you still have to turn into a doc. Scribe gives you a doc but no video. Knovvy gives you both from a single recording, cross-linked so viewers can jump between watching and reading.
Everything Scribe does, plus a real video from the same capture, and a 2-seat minimum instead of 5.
A shareable video and an accurate SOP, not a recording you have to manually document later.
One-click browser capture with auto-editing, no heavyweight pipeline to learn.
Worth understanding, because the guide is built differently depending on how you recorded.
Start a guide capture with "Record this tabβs video too" enabled and you get both at once: the steps come from real DOM events, and the video comes from the tab itself.
This is the strongest version. The SOP is click-accurate because it was never inferred, and the video is the same session rather than a separate take.
Steps keep recording across other tabs, but the video stays on the tab you started in. The toggle says so in the extension rather than leaving you to discover it.
Finish a screen recording that had narration and Knovvy builds a companion SOP from the transcript, then cross-links the two so each carries a banner pointing at the other.
Be clear-eyed about the difference: a guide derived from spoken words is inference, and it will never be as exact as one built from click events. It is the right tool for a walkthrough you narrated, not a replacement for capture.
Guide capture runs in your browser. Add the free Chrome extension, click through your workflow, and the steps, screenshots and highlights are written for you.
No. Dual-capture works on the free tier with no card and no AI credits, because it is the product rather than an upsell. Free-tier video carries a watermark and is capped at 20 minutes.
The browser capture. Its steps are built from real DOM events, so click targets and field names are exact. A companion guide built from a video transcript is inferred from spoken words and will be less precise by nature.