Videos
Faceless video maker
Faceless is not a style, it is a constraint you accept for a reason. Nobody has to be available, nothing has to be filmed twice when a price changes, and the same video can be reissued in another language without anybody appearing to speak it. The trade is that every second has to be carried by type, motion and voice.

What holds the screen instead of a person
Mostly type, moved deliberately. The statement role alone covers a single shouted line, three stacked words, a two line setup and payoff, a phrase cycling through several, an echoed ghost word, a broadcast lower third and a dictionary style definition. Around it sit counting numbers, building charts, ticking checklists, drawn funnels and cycles, a typing prompt bar and a conversation in message bubbles. None of it needs a camera and none of it needs stock footage.
The voice is doing more work than usual
With no face, the narration is the presence. Voices come from Google's Chirp3-HD set and are chosen by style and gender, in four styles: energetic, instructional, warm and neutral. The director is told to pick from the mood of your brief rather than the video type, so a sensitive subject gets warm or neutral even in a promo. In the builder you can go further and search the whole live voice catalogue per scene, plus set the delay, the playback speed and whether that scene captions itself.
Captions are not optional here
A faceless video is usually watched muted first. Captions are on by default and generated live from the narration, letters typing on slightly ahead of the voice in a fixed dark pill so they stay legible over any scene. They sit 44 pixels from the bottom in the wide cut and 120 in the vertical one, and cap at 72 percent of the frame width so a long line wraps rather than running edge to edge.
Where faceless is the wrong call
Two cases. Anything where trust comes from a specific person, a founder's apology, a surgeon explaining a procedure, a testimonial that needs a name and a face, is weaker without them. And anything where the product is a physical experience, food arriving at a table, a class in progress, reads flat as type and drawn graphics. Those want real footage or, in the second case, real photographs carrying the scenes.
How it works, in three steps
Step 1
Decide the through line
With no presenter, the script is the spine. Write the one sentence the viewer should repeat afterwards, and put it in the brief as the closing line.
Step 2
Let it choose the voice from the mood
Leave the voice unspecified in the brief and describe the tone instead. The style and gender are picked from that rather than defaulted by video type.
Step 3
Check the muted pass
Play it without sound in the preview. If a scene stops making sense, its on screen copy is leaning on the narration and needs rewriting.
The full walkthrough with screenshots is in the guide Add voiceover, captions and music to a video.
Limits worth knowing
- No footage is licensed in. Everything on screen is drawn by the engine, uploaded by you, or generated to a prompt.
- The synthesised voice is expressive but not directable line by line. There is no emphasis markup or pause control beyond delay and speed.
- Two languages do not have the premium voice tier at all and fall back to an older one: Malay and Filipino.
- Captions render as one centred pill. There is no per word karaoke styling, alternative placement or custom caption font.
Questions people ask
Does faceless mean no images at all?
No. It means no presenter. Screenshots, product photos and generated illustrations all still work, and a video built from your screenshots with a voice over it is faceless by this definition.
Can I record my own voice instead?
The scene narration is synthesised from text in the voice dialog. If you want your own voice, you would need to bring it as an uploaded audio track, and the automatic captions are derived from the narration text rather than from the audio.
How is this different from the creator format?
The creator format generates a person and animates them speaking, which needs the Pro plan and is capped under 30 seconds. Faceless has no such gate or ceiling and can run to twelve scenes.
Will it look the same as everybody else's?
That is the real risk with faceless video. The counters here are the 21 looks, 73 palettes, 27 fonts and 39 animated backgrounds, plus the fact that scene layouts are picked per role from 425 rather than from a fixed template.
Make your own video
The button opens the generator with this use case already described. Change the wording to match yours, generate, then edit anything you like.
Create a video with OneCraftRelated pages
AI video generator
You write a paragraph about what you are promoting. What comes back is not a storyboard or a set of stock clips to assemble, but a video that already plays: scenes picked and ordered, copy written into every box on screen, a narration track spoken aloud and captions running under it.
Text to video with voiceover
The interesting problem in turning text into a video is not writing the scenes. It is timing: a spoken line and a visual cut are two different clocks, and every text to video tool has to decide which one wins. Here the voice is treated as one continuous track and the picture is fitted around it.
AI creator video or faceless video
These are the two ways to make a video without filming anything, and they are not interchangeable. One casts a person and animates them talking. The other never shows anybody. They have different plan gates, different length ceilings, different scene libraries and different failure modes.
More finished work of this kind is on the video examples hub.