Skip to content
Blog

How-to

How to Write a Facecam Script: Structure, Template, and Tips

Learn how to write a facecam script that sounds natural: structure, hook, transitions into your screen recording, and tips for recording day.

Macreto-Team 6 min read

The best time to write a facecam script is after you’ve recorded your screen: watch what happens in the footage, then write your hook, transitions, and wrap-up in your own words. A good facecam script is short, sounds spoken rather than written, and tells viewers before each step what’s coming and why it matters.

In this guide, you’ll get the structure, a fill-in-the-blanks template, and the techniques that make a script sound natural on camera.

What is a facecam script?

A facecam script is the text you say while your face is visible in the video. In tutorials, AI explainers, and course lessons, it complements your screen recording: the recording shows the how, and the facecam delivers the why, adds context, and guides viewers through the video.

That makes it different from two related formats:

Format It’s… Its job
Facecam script seen and heard Guide, add context, build trust
Voiceover script only heard Narrate what’s happening on screen in real time
Blog post read Cover everything, work as a reference

The most common mistake is writing a facecam script like a blog post. What looks polished on paper often sounds stiff on camera.

Why you should write the script after recording your screen

A lot of creators write first and record second. For tutorials, the reverse usually works better:

  • Your script matches the footage. You only describe steps that actually appear in the recording.
  • The surprises have already happened. If a menu looks different than expected or a step takes longer, you already know when you sit down to write.
  • Recording stays relaxed. While you click, you focus on clicking instead of trying to phrase things at the same time.

For how this fits into the full workflow, see how to make a YouTube tutorial.

The structure: five parts of a facecam script

1. Hook: the problem and the result

Your first few sentences decide whether people stay. Name the problem the way your viewer experiences it, and show what the result looks like. No introductions, no “before we get started.”

2. Promise: what will viewers learn?

One sentence that tells people where they’ll end up: “By the end of this video, you’ll have…” Optionally, add a quick overview of the steps. It helps viewers orient themselves and lines up nicely with chapters.

3. Transitions: before every step

Each section of your screen recording gets a short transition. It answers two questions: What’s about to happen? Why does it matter? One to three sentences is enough.

4. Context: tips and pitfalls

This is where your experience comes in. What mistake do people make all the time? Which setting do most people miss? It’s the part a plain walkthrough can’t give you, and often the reason viewers recommend your video over another one.

5. Wrap-up: the result and the next step

Sum up the result in one sentence and point to a logical next step, like a related video. One call to action is good. Ten is too many.

Template: a facecam script for a tutorial

Copy this template and fill it in. Replace everything in square brackets.

HOOK
Ever run into this: [the problem in one sentence, from the viewer's side]?
In this video, I'll show you how to [result] – here's what it looks like at the end.
PROMISE
We'll do this in [number] steps: [step 1], [step 2], [step 3].
TRANSITION TO STEP 1
First, [what happens]. This matters because [reason].
TRANSITION TO STEP 2
Now that [result of step 1], let's tackle [step 2].
Heads up: [common pitfall].
TRANSITION TO STEP 3
Almost there. Last up, [step 3].
WRAP-UP
And that's it – you now have [result]. If you want to [related topic] next,
check out [video/chapter].

Writing for the ear: the rules that matter most

Spoken language plays by different rules than written language. These make the biggest difference:

  1. Short sentences. One idea per sentence. Long, nested sentences are hard to follow by ear and hard to deliver.
  2. Your own voice. Use the phrases you’d use on your channel anyway. If you’re casual on camera, be casual in the script.
  3. Active verbs. “Click Save,” not “The Save button should be clicked.”
  4. Make numbers and jargon speakable. Spell out acronyms the first time and round long numbers when the content allows it.
  5. Repetition is fine. In print, repeating yourself looks redundant. On camera, it helps people keep the thread.
  6. Read it out loud. Rewrite every spot where you stumble.

Phrases worth replacing

Instead of… Try…
“In today’s video, we’re going to be taking a look at how to…” “I’ll show you how to…”
“It’s important to note that…” “Heads up:…”
“As previously mentioned…” Just say the point again, briefly
“In the following section…” “Next up…”

Revising your script: two passes before recording

The first draft is rarely the best one. Two quick passes turn a decent script into a good one.

Pass 1: Content

Put the script next to your screen recording and go section by section:

  • Does every transition match what happens on screen right after it?
  • Are there lines that just describe what viewers can already see? Cut them.
  • Is the “why” missing anywhere? Add it.
  • Do menu paths, labels, and version numbers match the footage?

Pass 2: Sound

Now read it out loud, ideally at the pace you plan to record:

  • Where do you run out of breath mid-sentence? That’s where a period belongs.
  • Which words feel awkward to say? Swap them for your own.
  • Does the hook sound like the way you’d open a real conversation?

If someone can listen in for a few minutes, even better. A person who doesn’t know the topic will notice right away where an explanation is missing.

Recording with a teleprompter

A teleprompter displays your script right next to the camera, so you can read while still looking into the lens. A few tips for recording day:

  • Keep the text close to the lens. The closer it is to the camera, the less your eyes give away that you’re reading.
  • Slow down a little. Slightly slower than usual is better. You can trim pauses in the edit, but you can’t fix rushed delivery.
  • Record in sections. Hook, each transition, and the wrap-up as separate takes. A flub then only costs you one short section.
  • Pause after a flub. Then restart the sentence. That makes the right take easy to find in the edit.
  • Bring a bit more energy. Normal energy often reads as flat on camera. A little more emphasis and expression than in everyday conversation is usually right.

Common facecam script mistakes

  • Too long: When the facecam explains what the screen recording already shows, the video drags.
  • Too generic: Lines like “This is a super exciting topic” say nothing. Concrete pitfalls and results do.
  • Out of sync with the footage: The script announces a step that looks different in the recording. That’s why you record first and write second.
  • Not your voice: A script that doesn’t sound like you is obvious the moment you start reading it.

How Macreto helps you write facecam scripts

Macreto writes your facecam script from your screen recordings and a research step. First, Macreto researches your topic on the web and summarizes it with sources. You review the research and approve it. Then Macreto goes through your tutorial recordings and writes a facecam script with a hook, transitions, and a wrap-up, in your own voice and tone.

To make the script sound like you, Macreto can learn from your previous videos, via file upload or YouTube link: how you address viewers, your go-to phrases, your openings, and your pace. More on that on the learn your voice page. On recording day, you open the teleprompter in your browser and control it from your phone. You can also export the script as a TXT or DOCX file.

Frequently asked questions

How long should a facecam script be?

+

For tutorials, the facecam script is usually much shorter than the video, because the screen recording carries a lot of the weight. The facecam handles the hook, transitions, and wrap-up. The simplest check is to time yourself reading it out loud.

Should I read my script word for word or speak freely?

+

Both work. Reading word for word from a teleprompter saves takes and keeps you on topic. Talking from bullet points feels looser but tends to add repetition and filler words. Many creators read the hook verbatim and speak more freely after that.

How do I look natural reading from a teleprompter?

+

Write the way you talk and place the teleprompter as close to the lens as possible. Read the script out loud several times beforehand so you know the ideas, not just the words. Short pauses at the end of sentences sound more natural than a constant pace.

What's the difference between a facecam script and a voiceover script?

+

A voiceover script is only heard, while a facecam script is seen and heard. On camera, eye contact, facial expression, and direct address matter more, while a voiceover narrates closely along with what's happening on screen.

#script#facecam#teleprompter#tutorial

Your next video almost makes itself

We're letting creators in in small groups. Beta creators get one month of the Creator plan for free.

Get a beta spot