How to Add Subtitles to Vertical Video Without Covering Faces or Key Visuals

Learn how to place subtitles in vertical videos for readability and accessibility without covering faces, products, or key visuals.

*No credit card required
Tablet on a desk shows a video portrait with a subtitle bar covering the person's eyes
CapCut
CapCut
Aug 11, 2026

Place subtitles in the frame's safe area, usually the lower third or top two-thirds depending on the shot, and always move them if they block a face, product, or action. In vertical video, readable captions work best when they are timed to the audio, kept to 1-2 lines, and checked on a phone-sized screen before export.

Vertical video leaves less usable space than widescreen, so caption placement is not just a styling choice; it is part of accessibility and viewer comprehension. If you are editing social clips, product demos, or short lessons in CapCut, the practical workflow is: generate captions, review the transcript, reposition the text, and confirm that the final layout protects the main subject instead of competing with it. CapCut's online subtitle tool follows a similar workflow: generate captions, review the transcript, and adjust the text placement manually.

Why Subtitle Placement Matters in Vertical Video

Two tablets show a man on video, with subtitles placed at the top and bottom on the right and covering him on the left.

Vertical frames are narrow, and that makes captions more likely to cover faces, hands, text overlays, or product shots. Mobile screens can also crop the edges of the frame, so text that seems safe on a desktop preview may still feel crowded on a phone. Keeping captions in a deliberate safe zone helps preserve both readability and visual focus.

The accessibility case is straightforward: captions are a synchronized text display of spoken dialogue and other audible information, and they are intended to help viewers who are deaf or hard of hearing follow the audio. For prerecorded synchronized media, captions are required, and open captions and closed captions serve different workflow needs.

What Goes Wrong When Captions Sit Too Low

The most common mistake is defaulting to bottom-center placement without checking what is already happening in that part of the frame. If a speaker's face, a hand gesture, a logo, or a product sits low in the shot, bottom captions can make the edit feel cluttered. Some caption tools also have placement limits, so you may need to move the text earlier in the edit rather than trying to fix it at export.

A good rule is to treat captions as part of the composition, not as an afterthought. That means framing the shot first, then placing the subtitle layer where it preserves the subject and the action.

Start With Auto Captions, Then Reposition Manually

AI captioning can save time, but it should be treated as a first draft. Auto-generated captions often need review for timing, punctuation, spelling, speaker changes, and missing sound cues. CapCut's auto caption workflow can help generate subtitles from the footage, and then you can customize their placement and styling in the editor.

For a basic workflow, import or upload the clip, place it on the timeline, generate subtitles from speech, and then review the transcript against playback. CapCut's online editor supports upload, timeline editing, and export in the same workflow, which makes it easier to adjust captions before rendering the final vertical clip.

A Practical Editing Sequence

    1
  1. Upload the vertical clip and place it on the timeline.
  2. 2
  3. Generate subtitles from the audio or transcript.
  4. 3
  5. Check timing, spelling, and line breaks.
  6. 4
  7. Reposition captions away from faces, hands, products, and on-screen text.
  8. 5
  9. Play the full edit on a phone-sized screen.
  10. 6
  11. Export only after the captions stay readable without blocking key visuals.

That sequence matters because the best subtitle layout depends on where the motion, faces, and overlays sit in each scene, not just on the caption text itself.

Where Captions Should Go in a Vertical Frame

Three vertical frame layouts show subtitles placed at the bottom, top, and center around a silhouette placeholder.

A common default for vertical video is bottom-center placement, but that is only a starting point. If the lower frame contains a face, product, or other critical visual, move the subtitles to a cleaner area of the screen. When the bottom is crowded with embedded text or controls, top-center placement can work better.

For short-form video, caption placement should prioritize the least crowded area of the frame. Section 508 guidance also notes that on-screen text should often live in the top two-thirds because captions usually occupy the lower third, which reduces overlap with the dialogue area.

Safe Placement Rules That Hold Up Well

    1
  1. Keep subtitles away from faces and gestures.
  2. 2
  3. Avoid covering product shots, charts, or demo steps.
  4. 3
  5. Use the empty space in the frame, not the most active part of the scene.
  6. 4
  7. Shift captions between scenes if the subject changes location.
  8. 5
  9. Reframe or resize the clip if the shot is too crowded for text.

These rules are especially useful for creators who post tutorials, product explainers, talking-head clips, or educational reels where the subject may move through the frame.

Subtitle Styling That Stays Readable on Mobile

Tablet on a desk with a tall vertical rectangle on screen, beside a pause symbol card, color swatches, and a ruler

Placement solves only half the problem. Captions also need enough contrast, size, and timing to stay legible on a phone. Section 508 guidance recommends readable body text styling, typically a sans serif font such as Helvetica or Arial, with about 18-point text, white text on a black translucent background, and no more than two lines when possible.

Timing matters too. Captions should stay on screen long enough to read, and speech that runs much faster than normal can become hard to follow. The same guidance notes that speech over 180 words per minute may be too fast for captions, so fast dialogue may need trimming or tighter subtitle segmentation.

Styling Choices That Help, Not Hurt

    1
  1. Use high contrast between text and background.
  2. 2
  3. Keep line length short and breaks natural.
  4. 3
  5. Avoid oversized text that covers the subject.
  6. 4
  7. Use one consistent caption style across the edit.
  8. 5
  9. Check that resizing or font changes do not move text onto important visuals.

If your workflow uses subtitles for translation rather than accessibility, keep in mind that subtitles and captions are not identical. Captions include spoken dialogue and relevant sounds for accessibility, while subtitles are often used for translation and may omit those extra cues.

How CapCut Fits Into the Workflow

CapCut is useful when you want one editor that can generate subtitles, let you adjust the caption track, and then export the finished vertical video. The practical value is not that captions appear automatically, but that the text can be generated and then customized so it fits the frame, the pacing, and the subject placement. The tool page also describes upload, timeline editing, and export steps that fit a simple short-form workflow.

That said, AI-assisted captions still need human review. Auto-captioning can miss words, speaker changes, punctuation, or non-speech audio, and caption position still needs to be checked against faces and other key visuals. In other words, CapCut can reduce manual work, but it does not remove the need to inspect the final composition.

When AI Help Is Enough, and When It Is Not

AI captioning is most useful when:

    1
  1. the speech is clear,
  2. 2
  3. the scene changes are simple,
  4. 3
  5. the frame has obvious empty space,
  6. 4
  7. and the final export is reviewed on a mobile screen.

Manual adjustment becomes more important when:

    1
  1. the speaker moves through the frame,
  2. 2
  3. product demos use the lower third,
  4. 3
  5. multiple people appear on screen,
  6. 4
  7. or the video includes charts, overlays, or other embedded text.

A Simple Checklist Before Export

Use this checklist before you publish the video:

    1
  1. Generate captions from the audio.
  2. 2
  3. Review the transcript for accuracy.
  4. 3
  5. Put captions in the safest empty area of the frame.
  6. 4
  7. Move captions if they cover a face, hand, product, or text overlay.
  8. 5
  9. Test the video on a phone-sized screen.
  10. 6
  11. Export only when captions stay readable across the whole edit.

That sequence is especially useful for short-form social, marketing, and education videos, where one bad caption placement can hide the main point of the clip.

Comparison Table: Caption Options and Placement Choices

Table comparing caption types, best use, strengths, limitations, and placement notes.

The table is a planning tool, not a fixed rulebook. Your shot composition should decide the final caption position, especially on vertical videos where the lower frame often carries the most visual weight.

Key Takeaways

The best subtitle workflow for vertical video is simple: generate captions, verify the transcript, and then place the text where it does not cover faces or key visuals. In practice, that often means using the lower third only when it is actually empty, and moving captions higher or to another safe area when the frame is crowded.

CapCut can fit this workflow well because it supports subtitle generation, timeline editing, custom placement, and export in one editor. The key is to treat AI captions as a starting point, then check readability, sync, and composition before publishing.

Q: Where Should Subtitles Go In Vertical Video If A Face Is In Frame?

A: Put them in the clearest open space, not over the face. If the lower frame is crowded, move the captions higher or to another low-clutter area so the subject stays visible.

Q: Can Auto Captions Be Used Without Manual Review?

A: They can speed up the first draft, but they should still be checked for spelling, timing, missing sound cues, and placement. Auto captions are a starting point, not the final quality check.

Q: What Subtitle Style Works Best For Short-Form Vertical Clips?

A: A readable, high-contrast style with short lines, synced timing, and a safe position in the frame works best. Closed captions are usually the most flexible choice because viewers can turn them on or off.

Practical Next Steps

Before your next export, frame the shot with captions in mind: leave space where the subtitle will sit, import or generate the transcript, and move the text if it blocks faces, products, or other important details. If you are using CapCut, use its caption generation and editing tools to place subtitles after the visual composition is already set, not before it.

Hot and trending