How to film a Q&A video with both iPhone cameras
A reaction video works in picture-in-picture because two things are genuinely happening at once: the thing in front of you, and your face responding to it. A Q&A does not have that built in. It is one person, in one place, answering one question after another — nothing on the other side of the frame is doing anything by itself, which means the second camera in a dual recording does not automatically earn its space the way it does in a reaction or a vlog. Left unplanned, the inset ends up pointed at a wall or an empty half of the room for the entire video.
The fix is not more cameras, it is giving the second camera something specific to hold: the questions themselves, as physical objects, so a viewer has an actual reason to glance at that corner between answers. Setting the phone up to record alone at all — where to put it, what to check before you start, how long to set the countdown — is covered in the guide on recording a dual-camera video by yourself; this page is about what to do with the second frame once the phone is in position.
The second camera needs its own job
Picture-in-picture, split screen, either layout in DualCam assumes there are two things worth looking at. In a two-person interview that is obviously true — there are two faces. In a reaction it is true because the object or the screen you are reacting to keeps changing on its own. A Q&A is neither: it is your face, in the same spot, for the length of the video. If the back camera is just pointed at whatever happens to be behind you, the small window becomes decoration instead of content, and split screen becomes an empty half of the frame competing with your face for exactly no reason.
The straightforward fix is to give the back camera something that actually has content in it: a stack of question cards, sticky notes on a board, a printed list, or a notebook you are working through one line at a time. None of that requires special equipment — a handful of index cards on the desk in front of you is enough — but it turns the second feed into something a viewer is meant to look at, rather than a frame that is filled because the layout requires one.
Picture-in-picture, with the questions in the small window
Picture-in-picture is the better default here, for the same reason it works for most single-speaker formats: your face fills the frame as the main shot, and the small window carries the one other thing worth seeing — the card you are about to read from. Pick a corner that is not where your hand naturally rests when you pick up a new card, so your own arm does not block the one piece of the frame you added on purpose.
DualCam lets you swap which feed is the main view while recording, the same as moving the inset or resizing it, so a natural rhythm is to bring the cards to the main view for the second or two it takes to read a question clearly, then swap back to your face as the main shot for the answer. That is a deliberate choice available mid-take, not a fixed layout you commit to for the whole recording.
Split screen only when the notes deserve equal billing
A top-and-bottom split gives the cards genuinely equal weight with your face, which is worth it if the physical act of reading a question is part of the performance — a game-show pace where pulling a card and reading it cold is the point, not just a means of prompting yourself. For a calmer format where the cards are a reference rather than a reveal, equal billing mostly just wastes half the frame on a stack of paper sitting still.
Left-right split is the one to be careful with here specifically. It crops roughly a quarter off each side of a vertical canvas to fit both feeds side by side, and a card held anywhere near the edge of frame is exactly what that crop removes first — text near the border of the card can disappear from the recording even though it was visible in your hand a moment before. Picture-in-picture, or a top-and-bottom split, does not have that particular failure mode.
Say the question out loud — the track only remembers what you said
Dual camera recording writes one audio track, not one per camera, and it captures whatever the room actually sounds like — there is no separate channel carrying “the text on the card” the way there might be a caption file in an edited video. If you read a question silently and just start answering, the only record of what was actually asked is whatever a viewer can make out from a small, often slightly soft inset window, not the words themselves.
Reading the question aloud before you answer it is therefore not just a courtesy for anyone listening without watching — a commute, a podcast feed, an accessibility read-out — it is the only way the file itself actually contains the question. It also gives you a clean, repeatable cue point: state the question, pause half a beat, answer. That half-beat is useful later if you ever do want to trim or reorder answers, because it marks exactly where one question ends and the next begins.
One sitting, cut at every new question
A dual-camera recording runs straight through from record to stop with no pause in between, so fifteen questions answered back to back are either one long file or fifteen short ones — there is no in-between where you pause through the boring parts. Treating each new question as the start of a new take, the same way an interview treats each new topic, keeps the built-in recording library full of short clips you can pick from afterward instead of one long recording you have to scrub through to find where question eleven starts.
It also keeps the resolution choice honest. DualCam encodes at roughly 75 MB per minute regardless of whether you shoot at 720p or 1080p, and a Q&A sitting that runs twenty or thirty minutes across many short takes adds up in storage the same way any long shoot does. Because a Q&A rarely has fast motion for the extra resolution to preserve — it is mostly a still face and a still card — 720p is a reasonable default here, with 1080p worth the cost only if you plan to crop tightly into individual answers afterward.
Common follow-up questions
Do I need someone else there to ask the questions?
No. Most Q&A videos are one person working through cards, sticky notes or submitted comments alone, which is different from a two-person interview where both people are actually on camera. If someone else is physically asking the questions on camera with you, that is closer to an interview setup, and a split screen or the interview-specific framing advice applies instead.
Should I use picture-in-picture or split screen for a Q&A video?
Picture-in-picture, by default, with your face as the main shot and the question cards in the small window. Split screen is worth it only when reading the card is part of the performance and deserves equal weight with your face — otherwise it spends half the frame on a stack of paper that is not doing anything.
Should I read the question out loud if it is already written on the card?
Yes. A dual-camera recording has one audio track, and there is no separate record of the text on a card the way a caption file would carry it — if you do not say the question, the file does not contain it. Reading it aloud also gives you a clean point to cut between one question and the next.
Want to just do this?
DualCam records the iPhone front and back cameras at the same time and writes one finished MP4 while you shoot. Free, no account, no ads, nothing leaves the phone.