アプリに戻る

How to Export One Talking-Head Video as 9:16 and 16:9

October 9, 2026 · FilmeeAi Blog

Why one recording needs two different frames

If you make course lessons, product explainers, training videos, or talking-head YouTube content, you already know the problem: YouTube, LinkedIn, and most LMS players expect 16:9 horizontal video, while TikTok, Instagram Reels, YouTube Shorts, and most mobile viewing expect 9:16 vertical video. The talk is identical. The frame is not. Cutting both versions by hand usually means re-editing the clip twice: reframing, re-exporting subtitles, re-adding music, re-attaching the outro, and re-checking that nothing got cropped out of frame.

For a single five-minute lesson, that duplication can easily cost another 30-45 minutes of manual work per format, on top of the first edit. Multiply that by a course with twenty lessons, or a training team publishing weekly, and the vertical version quietly stops getting made at all — which is exactly why so much long-form talking-head content never shows up as Shorts or Reels.

What actually changes between the two versions

The talk, the music, and the subtitles don't change. What changes is:

  • Framing: a horizontal frame shows your shoulders and some background; a vertical crop usually tightens on your face and upper torso, so anything positioned near the left or right edge in the original recording can fall outside the vertical crop.
  • Caption placement: captions that sit safely at the bottom of a 16:9 frame can land directly under a platform's comment bar, like button, or caption overlay in 9:16, becoming unreadable.
  • Outro and overlay elements: a subscribe-button overlay or an outro clip designed for a wide frame needs to be re-checked in a tall one, or it ends up partially cropped.

Everything else about the edit — the words being said, the music bed, the trimmed silence — should stay identical between the two exports. That consistency is the whole point of doing this from one source edit rather than two separate projects.

Edit once, export twice: the core idea

The most time-efficient approach is to run the raw recording through one automatic edit, then generate both aspect ratios from that same edit rather than starting over. This is the core value of a tool like FilmeeAi: you upload a talking-head recording once, and it automatically handles background music (up to three tracks with crossfades), burned-in subtitles, your own outro video appended with volume matched to the rest of the clip, silence trimmed at the very start and end with a fade in/out, skin smoothing, brightening and face-contour slimming, and a subscribe-button overlay — then lets you export the finished result as vertical (9:16) or horizontal (16:9).

It also transcribes the talk, splits it into scenes, and inserts AI explainer animation that matches what's being said, which is worth checking in both frame shapes since animated inserts can sit differently depending on orientation. There's a separate, lighter-weight feature for turning a single line of text into a narrated storybook-style video, but that's a different workflow from exporting your own recorded talking-head footage and isn't the focus here.

Step-by-step: exporting the same talking-head video in both formats

  1. Record your talking-head video once. Keep yourself centered in the frame with headroom above your head and some margin on both sides — this margin is what a vertical crop will use later, so don't frame tight to the edges.
  2. Upload the raw recording to your editing tool. Do this once; you don't need a second recording for the second aspect ratio.
  3. Let the automatic transcription finish, then read through the generated subtitles and fix any names, numbers, or technical terms it misheard.
  4. Set your edit options once: pick up to three background music tracks, decide whether to trim the silence at the very start and end, choose skin smoothing/brightening/slimming levels, and toggle the subscribe-button overlay.
  5. Attach your outro video. Its volume will be automatically matched to the rest of the clip so it doesn't suddenly jump louder or quieter.
  6. Run the first export at 16:9 (horizontal). Watch the preview scene by scene, paying attention to where the AI explainer animation appears and whether any on-screen text sits close to the frame edge.
  7. Without rebuilding the edit, switch the export setting to 9:16 (vertical) and generate the second render from the same project. Review it specifically for cropped elements and caption placement against the bottom of the frame.
  8. Download both finished files, then do a final side-by-side check: play each one at actual size on a phone screen (for the vertical) and a desktop or TV-sized window (for the horizontal) before publishing.

Note what this workflow does not do: it won't remove filler words or cut out pauses in the middle of your recording. If your talk has long "um" stretches or dead air mid-sentence, that has to be re-recorded or trimmed before upload — the automatic trimming only applies to silence at the very beginning and end of the clip.

Real cost and time math before you export

Compositing AI explainer animation onto a talking-head video costs about 15 credits per minute of finished video. That cost applies per export, so producing both a horizontal and a vertical version of the same talk means paying for two downloads, not one — the editing work itself isn't what's billed, the finished download is.

For a 5-minute lesson with explainer animation: 15 × 5 = 75 credits for the horizontal export, and another 75 credits for the vertical export, for a total of 150 credits for both formats. New accounts get 200 free credits on sign-up with no credit card required, which covers both exports of that same 5-minute video with 50 credits left over. For a 10-minute talk, each export runs about 150 credits, so both formats together cost roughly 300 credits — more than the free allowance, which is where a paid plan (starting at $19/month) comes in. Credits roll over month to month and are only spent when you actually download a finished video, so a render you decide not to keep costs nothing, and a failed render is free as well — you can re-try a problematic export without burning credits.

Budgeting rule of thumb: minutes of finished video × 15 × number of aspect ratios you plan to export = credits needed.

Common mistakes that ruin a dual-format export

  • Framing too tight in the original recording. If you fill the horizontal frame edge-to-edge with your shoulders, the vertical crop has nowhere to pull from and will cut off part of your face or body. Fix: record with visible margin on both sides and above your head, even if the horizontal version looks slightly "loose" — it won't once it's published.
  • Assuming captions that work horizontally will work vertically. Subtitles positioned safely above the bottom edge in 16:9 often land inside the comment/caption zone that TikTok, Reels, and Shorts overlay on top of vertical video. Fix: preview the vertical export specifically, at phone size, before publishing — not just on a desktop monitor where everything looks fine.
  • Re-recording a second take "for the vertical version." This doubles your filming time and guarantees small inconsistencies between the two versions — different pacing, slightly different wording, a different background. Fix: record once, run one edit, and export both aspect ratios from that single project.
  • Forgetting that explainer animation cost is per export. Teams sometimes budget credits for one format and are surprised the second export costs the same again. Fix: use the credits-per-minute math above before you start a batch, not after.

Pre-flight checklist

  • Subject centered in the frame with visible margin on both sides and above the head
  • Audio clean enough for an accurate automatic transcript (background noise degrades subtitle accuracy, not just sound quality)
  • Outro video ready and uploaded, volume already reasonable so auto-matching has something sensible to match to
  • Music tracks chosen (up to three, with crossfades) and roughly leveled against your voice
  • Decided which platforms need which format: YouTube main uploads, LinkedIn, and most LMS players want 16:9; TikTok, Reels, Shorts, and Stories want 9:16
  • Credits budgeted using minutes × 15 × number of formats, with your free 200-credit balance or paid plan checked first
  • Time set aside to preview the vertical export on an actual phone screen, not just a desktop window

Scaling this across many lessons or training videos

Course creators and internal-comms teams rarely need just one dual-format video — they need this done consistently across a whole library, often on a recurring schedule. The same edit-once, export-twice logic applies at scale: the main cost driver is the explainer-animation compositing fee per minute per export, so a 10-lesson course at 5 minutes each, exported in both formats, runs roughly 1,500 credits total (75 credits × 2 formats × 10 lessons), which is a useful number to have before you commit a training budget. For teams that already script or trigger video production from other tools, this kind of pipeline can also be run straight from an AI assistant or developer workflow: FilmeeAi can be connected through an MCP-compatible assistant at filmee.app/mcp, which can take a direct link to an already-recorded talking-head file (.mp4, .mov, or .webm, up to 15 minutes — a YouTube page link won't work), add subtitles, and insert matching explainer animation, then hand back a link to the finished video; setup details are at filmee.app/developers.

Frequently asked questions

Do I need to record my talking-head video twice, once for each format?

No. Record once, with enough margin around you in the frame, and run that single recording through one automatic edit. The vertical and horizontal versions are two separate exports of that same edited project, not two separate recordings or two separate edits.

Will exporting in both formats remove filler words or awkward pauses in the middle of my talk?

No. Automatic editing of this kind trims silence only at the very start and end of the recording, with a fade in/out, and does not cut filler words or dead air in the middle of the talk. If mid-recording pauses are a problem, address them before uploading, either by re-recording that section or trimming it yourself first.

How much does it actually cost to get both a vertical and horizontal version of a 5-minute video?

Using the rate of about 15 credits per minute of finished video for explainer-animation compositing, a 5-minute talking-head video costs roughly 75 credits per export. Getting both a 16:9 and a 9:16 version means two exports, so about 150 credits total. The 200 free credits given on sign-up cover that comfortably, with 50 credits left over, since credits are only spent when you download a finished video and failed renders cost nothing.

FilmeeAi turns a single line of text into a finished anime video with narration and BGM — and can drop AI explainer animation straight into your own talking-head footage. Sign up and you get free credits, no card required.

Make a video for free →

See what the AI actually produces in the gallery.

← Back to all articles