How to Make a Silent Product Demo With On-Screen Steps
A 10-minute video guide for a silent product demo with five on-screen step captions. Ban covered buttons. Mute-on phone check. ChatGPT, CapCut, Runway.
Anyone shipping a product clip can make a silent product demo with on-screen steps in about 10 minutes: lock five step captions, record or generate a mute-friendly walkthrough, and place each caption where hands and UI stay readable. Run the caption plan in ChatGPT or Claude, then cut in CapCut, Descript, or Runway before the model writes a voiceover you do not need.
A silent product demo with on-screen steps is a mute-safe walkthrough viewers can follow without sound, not a cinematic trailer and not a talking-head pitch. If the captions cover the button being clicked, you never did the job. Five clear step lines beat a script nobody hears on the subway.
This is not cut a demo video so the cursor never jumps. Cursor continuity is about edit seams. This guide is about mute-safe step captions. It is not burned-in captions that do not cover hands. That guide places captions on any hand video. Here the captions are numbered product steps. It is not a five-second product spin video. A spin shows the object. This demo shows the workflow.
What you need
Your product UI or physical product, a phone or screen recorder, five step outcomes in plain verbs, ChatGPT or Claude for caption copy, and CapCut, Descript, or Runway for timeline placement. Ten minutes for the caption pass plus your normal cut time. Headphones optional—sound stays off for the viewer check.
Banned before you start: relying on voiceover, tiny captions under 5% height, captions over primary buttons, inventing steps the UI does not support, and trailer music that implies sound is required.
Build a mute-safe stepped demo
1. Lock five step captions before you record
List the five outcomes as short imperative lines. Paste into ChatGPT or Claude: “Rewrite these five product steps as on-screen captions. Max 6 words each. Number them 1–5. Ban marketing adjectives. Ban inventing clicks not in my list. Output caption text only. Then stop.”
Good: “1. Open Projects. 2. Tap New brief. 3. Paste goal. 4. Choose template. 5. Export PDF.” Bad: “1. Unlock magical productivity in seconds.” Magical is not a click.
2. Record or generate the walkthrough mute-first
Screen-record the five clicks slowly, or generate a clean UI walkthrough in Runway only if you already have approved product frames. Keep the cursor path calm. Do not talk while recording if you want a true silent master.
Instruct yourself: one action per beat, pause one second on the result, no zoom spam. If you use Runway for B-roll inserts, keep them under one second and never replace a real click with a fantasy UI.
Keep captions off the action zone
Reserve the lower-third or a safe top bar that does not cover primary buttons, form fields, or hands. If your UI is bottom-heavy, put captions top-center. Test on a phone screen, not only a desktop monitor.
3. Place numbered captions on the timeline
In CapCut, Descript, or your editor, drop captions 1–5 so each appears before the click and holds through the result. Font large enough to read muted on a phone. High contrast. No karaoke animation that distracts from the cursor.
Prompt for polish if needed: “Suggest safe caption positions for a mobile-first silent demo. Ban covering primary CTAs. Keep numbers visible.” Then apply positions yourself—do not trust auto-layout blindly.
4. Mute-on cold check on a phone
Watch the full cut with sound off. Can a stranger complete the five steps from captions alone? If step 3 is unclear, rewrite only that caption: “Replace caption 3 only. Max 6 words. Ban fluff.”
Hard fail: any step that needs audio; any caption over the button; any invented sixth step; any music cue that implies narration.
5. One human pass, then export
Export a silent master (or a version with optional soft bed that is not required). Do not let the model “add a VO script” after the fact if your distribution is mute-heavy feeds.
Check traps: long captions, covered UI, missing numbers, trailer energy, steps the product cannot do. Fix by regenerating caption copy only, then re-time. If viewers still need sound to understand step 4, you never made a silent product demo with on-screen steps. Stop.
What good looks like
Trigger: You are about to ship a demo that fails on mute in the feed.
Input: Five step verbs, ChatGPT/Claude captions, screen record or Runway inserts, CapCut/Descript timeline, phone mute check.
Output: Numbered on-screen steps, readable on phone, no covered buttons, zero required voiceover.
Stop: After one mute-on phone check. No VO rescue. No invented clicks.
Treat the demo like a mute instruction card, not a brand film. Models love cinematic intros. Delete intros that delay step 1. Keep captions short so thumbs can follow on a noisy commute. When editors “helpfully” auto-caption a discarded VO, strip it if your master is silent. The mute check exists to catch covered UI and unclear steps. Skip that check and you will ship a demo nobody can follow without headphones. That wastes views.
Distribute the silent master to mute-heavy feeds first, then optionally add a soft bed that is not required for comprehension. Keep a caption style guide: one font, one size, one number format across all five steps. If you localize later, regenerate captions with the same six-word cap rather than stretching English lines. When stakeholders ask for a VO cut, keep the silent master as the source of truth so captions never drift from the clicks. Models love rewriting steps into slogans. Restore verbs that match the UI labels exactly.
Takeaways
- Lock five step captions at ≤6 words before you record.
- Place numbers where buttons and hands stay clear.
- Mute-on phone check. Ban inventing clicks.
- Export a silent master. Do not rescue with VO after the fact.
Frequently Asked Questions
Do I need voiceover at all?
No. The point is mute-safe steps. Optional soft music is fine only if captions alone still teach the flow.
How many steps should I use?
Five is the default for feeds. If your flow needs more, split into two demos instead of tiny unread captions.
Can Runway invent UI screens?
No. Use real product frames. Invented UI teaches clicks that do not exist.
Where should captions sit on mobile?
Away from primary buttons—often top-center if the UI is bottom-heavy. Verify on a phone, not only desktop.
What if a step needs more than six words?
Split the action or simplify the UI beat. Long captions fail on mute in the feed.