Skip to content

Feature

No camera and no face, and the script is still your job

The numbers first, because faceless is sold as effortless and it is not. A presenter renders from one photo and one audio track at 350 credits a minute, on Basic at $39.99 a month. A generated read is 70 credits a minute. What no module replaces is the script, and this page is honest about that part.

Per rendered presenter minute
350 CR
Per generated voice-over minute
70 CR
Basic, the plan avatars live on
$39.99
The most speech one video takes
3 min

The prices before the pitch, because faceless is sold as effortless

A rendered presenter costs 350 credits per minute of output, and it needs Basic at $39.99 a month for 2,500 credits. That is the biggest number in the product, so it goes first.

A generated read is 70 credits per minute, for when nobody wants to record the audio. The batch that cuts any video is 100 credits, and export is 20 credits per finished minute.

So a sixty second faceless video runs from about 120 credits filmed behind the camera, to about 540 with a rendered presenter and a generated read.

Read those numbers before the pitch, because the pitch is real but not magic. What the render buys is a face without a booking. What it does not buy is a video without work.

A sixty second faceless video, itemised by build

Every build includes the 100 credit batch and the export at 20 credits a minute. The render is the expensive part.

  • Filmed behind the camera120 CR
  • Presenter, your own recording470 CR
  • Presenter, generated read540 CR

Build one: a photo and an audio track become the presenter

You hand the module one front-facing photo and one audio track of the script. It returns a lip-synced take of that person delivering your words.

That take enters the editor like filmed footage. It gets transcribed and cut, footage lands on the words that need showing, captions come from one of six packs, and the export is a 9:16 MP4.

The face has to be one you may use. Yours, or somebody who agreed in writing, and never a public figure. A stock portrait only works when its licence explicitly covers synthetic video.

Write the read short and cut away to footage often. A rendered face is most convincing in bursts, and a long unbroken hold is where viewers start to study it.

This is the whole editor

Highlight a phrase and a clip lands on those exact words. No timeline, no keyframes, no layers.

CLIPS · 5I have thisexact conversationeverysingle week. Somebody sits down and says,oh yeah, I takecinnamonevery day.And honestly, doc, I have no idea if it works.So let me tell you what isin that capsule.a clip lands on these wordscut from the editTAKING CINNAMONEVERY DAY?is it doing anythingHeadlineMusicCaptionsTHIS VIDEOLength25.0sClips5Words removed18Export video

Build two: keep every face out of frame, and keep the voice

The uploader takes any video with speech in it, and nothing requires a face in the frame. Prop the phone over the desk or the product and talk while you show it.

Up to three minutes of that batches for 100 credits, and it comes back cut, captioned and covered like any other take. Your voice carries it. Your face is nowhere.

Footage coverage runs to 40, 45 or 52 percent of the video depending on pace, and the captions do the visual work a face would have done. On this build the caption pack matters more than on any other kind of video.

This build runs on Lite at $19.99 a month, because it uses no render. It is the cheapest honest answer to the faceless search, which is why it is on this page at all.

The camera was never the work. The script was, and it still is

Faceless gets sold as effortless, and removing the camera does remove something real: the styling, the nerves, the reshoots. It removes none of the actual work.

Somebody still decides what the video argues. Somebody still writes it in sentences a voice can carry. Somebody still reads it, or pays 70 credits a minute for a voice with no opinion of the words.

The machine cuts what was said. The render moves a mouth in time with the audio it was given. Neither one improves a script, and a weak script comes back as a well-produced weak video.

Write with concrete nouns, because the footage search reads the words. Name the product, the number, the before and the after. Abstract lines give the search nothing to place.

The faceless build, in the order it actually happens

The first step has no price on it, and it decides more of the result than the four that do.

  1. Write the script

    The part no module does

  2. Get the read

    Record it, or 70 CR a minute

  3. Add one photo

    A face you may use

  4. Render

    350 CR a minute, on Basic

  5. Cut, cover, caption

    Inside the 100 CR batch

  6. Export

    9:16 MP4, 20 CR a minute

The long uploads of a faceless channel are the wrong shape here

Faceless channel usually means long uploads. Ten minute compilations, narrated stories, stock montage with a synthetic voice front to back. Say it plainly: that is not what this builds.

The source ceiling is three minutes of speech, the export is 9:16 vertical, and footage never covers more than 52 percent, because a person, a hand or a product stays the spine of the video.

For the long uploads you want a timeline editor and a music library subscription, and no purchase here changes that.

What fits is the channel's vertical end. The Shorts, the promos, the ads that point at the channel. One three minute session at the desk becomes those, on the numbers at the top of this page.

Questions people ask

Do I ever have to show my own face?
No. One build renders a presenter from a photo of somebody who consented, at 350 credits a minute on Basic. The other films past you: hands, product, screen, with your voice carrying it, through the normal 100 credit batch.
Whose photo can the presenter use?
Yours, a person who agreed in writing, or a licensed likeness whose licence explicitly covers synthetic video. Never a public figure. The consent question does not disappear because the video is generated.
Can it write the script for me?
No, and be suspicious of tools that say yes cheaply. The machine cuts what you said and renders what you wrote. The argument itself has to come from somebody who knows the product and the customer.
Can it read the script out loud for me?
Yes. A generated read is 70 credits per minute, and it goes straight into the presenter render. Recording it yourself costs nothing extra and usually sounds more like a person, which is the entire point of a presenter.
Will it make my ten minute faceless YouTube videos?
No. Three minutes of speech is the ceiling and 9:16 is the only export. Long faceless uploads belong in a timeline editor with a music subscription. Bring the vertical work here: the Shorts, the trailer, the ad.
Which plan do I need?
The presenter build needs Basic at $39.99 a month for 2,500 credits. The behind-the-camera build runs on Lite at $19.99 for 1,250. Both export with nothing on the picture, and cancelling is one click.

The render replaces the camera for 350 credits a minute. Nothing on this page replaces the person with something to say.

3 videos free, no card3 finished videos free in your first 7 days, no card. They carry a Cutroom mark; Lite at $19.99/month removes it