Feature
No camera and no face, and the script is still your job
The numbers first, because faceless is sold as effortless and it is not. A presenter renders from one photo and one audio track at 350 credits a minute, on Basic at $39.99 a month. A generated read is 70 credits a minute. What no module replaces is the script, and this page is honest about that part.
- Per rendered presenter minute
- 350 CR
- Per generated voice-over minute
- 70 CR
- Basic, the plan avatars live on
- $39.99
- The most speech one video takes
- 3 min
The prices before the pitch, because faceless is sold as effortless
A rendered presenter costs 350 credits per minute of output, and it needs Basic at $39.99 a month for 2,500 credits. That is the biggest number in the product, so it goes first.
A generated read is 70 credits per minute, for when nobody wants to record the audio. The batch that cuts any video is 100 credits, and export is 20 credits per finished minute.
So a sixty second faceless video runs from about 120 credits filmed behind the camera, to about 540 with a rendered presenter and a generated read.
Read those numbers before the pitch, because the pitch is real but not magic. What the render buys is a face without a booking. What it does not buy is a video without work.
A sixty second faceless video, itemised by build
Every build includes the 100 credit batch and the export at 20 credits a minute. The render is the expensive part.
- Filmed behind the camera120 CR
- Presenter, your own recording470 CR
- Presenter, generated read540 CR
Build one: a photo and an audio track become the presenter
You hand the module one front-facing photo and one audio track of the script. It returns a lip-synced take of that person delivering your words.
That take enters the editor like filmed footage. It gets transcribed and cut, footage lands on the words that need showing, captions come from one of six packs, and the export is a 9:16 MP4.
The face has to be one you may use. Yours, or somebody who agreed in writing, and never a public figure. A stock portrait only works when its licence explicitly covers synthetic video.
Write the read short and cut away to footage often. A rendered face is most convincing in bursts, and a long unbroken hold is where viewers start to study it.
This is the whole editor
Highlight a phrase and a clip lands on those exact words. No timeline, no keyframes, no layers.
Build two: keep every face out of frame, and keep the voice
The uploader takes any video with speech in it, and nothing requires a face in the frame. Prop the phone over the desk or the product and talk while you show it.
Up to three minutes of that batches for 100 credits, and it comes back cut, captioned and covered like any other take. Your voice carries it. Your face is nowhere.
Footage coverage runs to 40, 45 or 52 percent of the video depending on pace, and the captions do the visual work a face would have done. On this build the caption pack matters more than on any other kind of video.
This build runs on Lite at $19.99 a month, because it uses no render. It is the cheapest honest answer to the faceless search, which is why it is on this page at all.
The camera was never the work. The script was, and it still is
Faceless gets sold as effortless, and removing the camera does remove something real: the styling, the nerves, the reshoots. It removes none of the actual work.
Somebody still decides what the video argues. Somebody still writes it in sentences a voice can carry. Somebody still reads it, or pays 70 credits a minute for a voice with no opinion of the words.
The machine cuts what was said. The render moves a mouth in time with the audio it was given. Neither one improves a script, and a weak script comes back as a well-produced weak video.
Write with concrete nouns, because the footage search reads the words. Name the product, the number, the before and the after. Abstract lines give the search nothing to place.
The faceless build, in the order it actually happens
The first step has no price on it, and it decides more of the result than the four that do.
Write the script
The part no module does
Get the read
Record it, or 70 CR a minute
Add one photo
A face you may use
Render
350 CR a minute, on Basic
Cut, cover, caption
Inside the 100 CR batch
Export
9:16 MP4, 20 CR a minute
The long uploads of a faceless channel are the wrong shape here
Faceless channel usually means long uploads. Ten minute compilations, narrated stories, stock montage with a synthetic voice front to back. Say it plainly: that is not what this builds.
The source ceiling is three minutes of speech, the export is 9:16 vertical, and footage never covers more than 52 percent, because a person, a hand or a product stays the spine of the video.
For the long uploads you want a timeline editor and a music library subscription, and no purchase here changes that.
What fits is the channel's vertical end. The Shorts, the promos, the ads that point at the channel. One three minute session at the desk becomes those, on the numbers at the top of this page.
Questions people ask
- Do I ever have to show my own face?
- No. One build renders a presenter from a photo of somebody who consented, at 350 credits a minute on Basic. The other films past you: hands, product, screen, with your voice carrying it, through the normal 100 credit batch.
- Whose photo can the presenter use?
- Yours, a person who agreed in writing, or a licensed likeness whose licence explicitly covers synthetic video. Never a public figure. The consent question does not disappear because the video is generated.
- Can it write the script for me?
- No, and be suspicious of tools that say yes cheaply. The machine cuts what you said and renders what you wrote. The argument itself has to come from somebody who knows the product and the customer.
- Can it read the script out loud for me?
- Yes. A generated read is 70 credits per minute, and it goes straight into the presenter render. Recording it yourself costs nothing extra and usually sounds more like a person, which is the entire point of a presenter.
- Will it make my ten minute faceless YouTube videos?
- No. Three minutes of speech is the ceiling and 9:16 is the only export. Long faceless uploads belong in a timeline editor with a music subscription. Bring the vertical work here: the Shorts, the trailer, the ad.
- Which plan do I need?
- The presenter build needs Basic at $39.99 a month for 2,500 credits. The behind-the-camera build runs on Lite at $19.99 for 1,250. Both export with nothing on the picture, and cancelling is one click.
The render replaces the camera for 350 credits a minute. Nothing on this page replaces the person with something to say.