Three modes, three different skills
Most people assume a mock interview is a mock interview. It is not. Typing an answer, speaking it aloud, and speaking it on camera stress-test completely different muscles. Text checks whether your reasoning is sound. Voice checks whether you can deliver that reasoning in real time without rambling. Video adds body language, framing and the pressure of being watched. Practising only one leaves a blind spot the real interview will find.
Text mode: structure and reasoning
Text is the lowest-pressure entry point. You can think before you commit a word, so it is the best place to fix the substance of an answer — the logic of a system design, the structure of a behavioural story, the correctness of your code. On Wrexa Edge the executable coding round pairs naturally with text: you write, run against test cases, and iterate without a clock in your head.
- Pros: no audio nerves, easy to review the transcript, great for content and structure.
- Cons: hides pace and filler-word problems, does not train live recall, feels nothing like a phone screen.
- Maps to: asynchronous screens, take-home questions and online assessments (OAs).
Voice mode: pace and clarity
Voice is where most candidates discover their real weakness. Answers that read cleanly fall apart when spoken — you trail off, you fill silence with "um" and "like", you answer a question you were not asked. Voice trains you to think aloud, to signpost ("there are three trade-offs here"), and to land a conclusion before you run out of breath.
- Pros: builds live recall, exposes filler words and pacing, closest thing to a recruiter call.
- Cons: no visual feedback, easy to forget the structure you nailed in text.
- Maps to: recruiter and hiring-manager phone screens.
Video mode: presence and framing
Video is the full simulation. Now you are managing eye contact, framing, lighting and the odd sensation of talking to a camera while still keeping your answer tight. It is the most uncomfortable mode, which is exactly why it is the most valuable one before an on-site or a virtual final round.
- Pros: trains eye contact, posture and framing, highest transfer to virtual on-sites.
- Cons: highest setup cost, hardest to do casually, most tiring.
- Maps to: virtual on-sites, panel rounds and final-round loops.
Which mirrors your real interview?
Match the mode to the round you are actually facing. A first-stage phone screen is voice, so rehearse in voice. A remote on-site loop is video, so at least one of your practice runs should be on camera. An OA or take-home is text. If you do not yet know the format, ask the recruiter — and until you find out, default to voice, because it is the most common early filter.
A progression that builds confidence
Do not jump straight to video and demoralise yourself. Ladder up:
- Run the interview in text first to get the content right.
- Repeat the same interview in voice to fix pace and filler words.
- Finish in video once the words come easily, so you can spend attention on presence.
Because Wrexa Edge scores every run on the same rubric, you can watch the same answer improve as you climb the ladder, and your practice dashboard tracks the trend across all three modes.
Quick tips for each mode
- Text: write in structured chunks, not a wall — use a clear opening line, then the reasoning, then a conclusion.
- Voice: slow down, pause instead of saying "um", and always finish with a one-sentence summary.
- Video: put the camera at eye level, look at the lens not the screen, frame yourself from the chest up in even light.
Pick your mode and start a free run on the AI mock interview. If you want the full pre-interview routine, read how to prepare for an AI mock interview, and to sidestep the usual traps see 10 common AI mock interview mistakes. Company-bound? Try the Google SDE board or the general SDE role loop.
Frequently asked questions
Do I need a webcam and mic to practise?
No. You can do the entire interview in text with just a keyboard. Voice needs a microphone and video needs a camera, but you can start free with text and add the harder modes as you build confidence.
Is video always the best mode to practise in?
Only if your real round is on camera. Video trains presence, but if you are facing a phone screen or an OA you will get more transfer from voice or text. Match the mode to the format you will actually sit.
Does the mode change how I am scored?
The rubric evaluates the substance and delivery of your answers, not the medium. The same evidence-backed feedback applies across text, voice and video, so you can compare your progress fairly across modes. Practice does not predict hiring outcomes.