Comparison
Cutroom vs Riverside: clean capture, and still no finished ad
Riverside solves capture. Remote guests recorded locally, a separate track per speaker, broadcast audio off bad wifi. Cutroom starts at the file and turns one filmed read into an ad. This page covers where capture stops, what one batch does to a take, and what a finished minute costs here. By the end you will know which of the two your bottleneck actually is.
- speaker, one mixed track
- 1
- upload ceiling, enforced
- 3 min
- credits per output minute
- 20
A guest on bad wifi arriving as a clean track is worth the subscription on its own
Getting broadcast-quality audio and video out of people who are not in the same room. It records each participant locally at full quality and uploads in the background.
So a guest on poor wifi arrives as a clean track rather than a pixellated mess. Anyone who has salvaged an interview ruined by a dropped connection knows exactly what that is worth.
Separate tracks per speaker is the detail that matters most in post. You can fix one person's levels without touching the other. You can cut around a cough that lands under somebody else's sentence.
That cannot be added later if the recording was mixed down at source. It is the one decision in the whole pipeline you genuinely cannot undo.
What a clean track is not is an ad. Everything after the recording is the half Cutroom does.
Count hours captured against ads published. That ratio is the diagnosis
A clean recording is a floor, not a finish. Once the file exists the questions change. Which forty seconds of this earns attention. What gets thrown away.
Where does the viewer need to see something instead of a face. How fast should it move. None of that is answered by a better microphone or a second track.
That is why the two categories coexist happily. Recording infrastructure has no opinion about where a cut lands inside a spoken sentence. It was never meant to.
The mistake here is a sequencing mistake. People buy excellent capture. They accumulate excellent recordings. Then they stall at the edit. The bottleneck moved and the budget did not follow it.
Count it on your own drive tonight. Hours captured against ads published. If the first number is large and the second is small, you have diagnosed yourself without reading further.
This is the whole editor
Highlight a phrase and a clip lands on those exact words. No timeline, no keyframes, no layers.
Cutroom takes one speaker, three minutes, vertical, and returns a finished MP4
It takes one talking-head take, one speaker, up to three minutes. Back comes a finished 9:16 MP4 with captions burned in and b-roll placed on the phrases that need showing.
You direct on the transcript. Delete a line and the cut rebuilds. Highlight a phrase and footage covers exactly those words. Change the pace and the whole piece re-cuts. Trim the opening.
Eleven actions, no timeline, no layers, no frame-level nudging of anything. Coverage is capped at 40 percent on chill, 45 on normal and 52 on fast, so your face stays on screen for most of the ad.
State the limits plainly, because they decide whether the pairing works. Three minutes maximum. One speaker. Vertical only. One mixed audio track. No recording of any kind here.
A two-person interview is not a shape this product understands. Anything from a conversation has to be reduced to one person talking before it arrives.
- Single speaker, vertical, three minutes or less
- Six caption packs and five clip transitions
- 100 credits a batch, 20 credits per output minute
- 7 day trial, 300 credits, no card, tested on a real take
Where the capture tool stops and this one starts
Nothing in the first two boxes happens here. Nothing in the last three happens in a recording tool.
Record the guest
Not us
Edit the episode
Not us
Film a separate read
Under 3 min, alone
Direct on the transcript
100 CR batch
Export
9:16 MP4, 20 CR a minute
Record the show, then film four separate reads. A conversation and a pitch are different shapes
Record the show properly, with separate tracks per speaker, and edit it as an episode wherever you do that now. Nothing here changes any of it.
Separately, once a month, film four short reads to camera about what your guests made you realise. One take each, under a minute, no guest and no back and forth.
Those are ad material. The episodes are not, even when the words are nearly identical. The shape is what differs, and the shape is what a cold feed judges in the first second.
Bring each read here, correct the transcript, cut it and export it. Four ads from one sitting, at roughly 120 credits each. Each one can be re-cut later for the cost of the export alone.
One more reason to keep them apart. A guest owns half of an interview, so clearing a clip for paid use is a permissions question with an email attached. A read you filmed alone is not.
Four ads from one sitting, against a month of credits
Basic is 2,500 credits a month for $39.99. One sitting of four reads uses under a fifth of it.
- One finished minute120 CR
- Four reads in one sitting480 CR
- Basic, every month2,500 CR
Capture is one purchase. The finished ad is a different one.
Pick Riverside if the problem is capture. Remote guests, multiple speakers, audio that has to survive editing, a show shipping on a schedule.
Pick Cutroom if the problem is the ad. A filmed read exists and it has to become something a stranger will watch in a paid feed without being asked twice.
The facts a buyer needs. Cutroom records nothing. There is no guest workflow, no scheduling, no local capture, no separate audio stems, no audio-only path.
It expects one person on video, talking to a camera, for three minutes or less. Output is a 9:16 MP4. There is no timeline underneath and no frame-level nudging.
If your recording quality is already fine on a phone, you may not need a capture tool at all. Running both makes sense when you record interviews and also run ads with your own face in them.
Separate pipelines, separate budgets. Ours: 100 credits a batch, 20 per output minute, $39.99 a month for 2,500 credits. Read theirs at the source.
Riverside and Cutroom, row by row
Two rows go to Riverside, and they own the recording. The other six are what turns that file into an ad.
| Riverside | Cutroom | |
|---|---|---|
| Records remote guests | Local capture, clean tracks | No recording of any kind |
| More than one speaker | A separate track per person | Single speaker only |
| Cuts a read into an ad | Clipping, not directing | Drafted cut, captions burned in |
| Captions burned into the export | Add them somewhere else | Six packs, restyleable |
| Speaker held on screen by a cap | Nothing to cap | 40, 45 or 52 percent b-roll |
| Delete a line, cut rebuilds | Trim the clip by hand | The transcript is the surface |
| B-roll on phrases you name | Capture, not placement | Highlight words, clip lands |
| Cheap fifth edit | Re-clip the same episode | 20 credits per output minute |
Questions people ask
- Can I record inside Cutroom?
- No. There is no recording feature. You film elsewhere and upload the take.
- Can I bring a podcast episode into Cutroom?
- Only if you reduce it first to one speaker and three minutes or less. Full episodes and two-person conversations sit outside what the product handles.
- Does Cutroom need studio-quality footage?
- No. A phone in decent light is the assumption. Clear audio helps the transcription, which is what everything downstream is built on.
- Who should buy Riverside instead of us?
- Anyone whose bottleneck is capture: remote guests, more than one speaker, or audio that arrives too compressed to edit. Fix the recording first, then bring one clean read here to be cut.
Great capture and a finished ad are two different purchases. Cutroom is the second one, and it costs about 120 credits.