Eye tracking pattern
Eyes moving left-right in a reading scan rather than natural upward or side-glance thinking patterns. Reading pattern is distinctive.
Screen-share captures, eye tracking, behavioral cues, Whisper Mode explained, and the honest caveats — a technical breakdown.
Whisper Mode is built on the Document Picture-in-Picture API — a W3C web standard (documented at developer.chrome.com/docs/web-platform/document-picture-in-picture) that renders a floating window in a separate document context at the OS compositor layer.
When you share a browser tab, a window, or your entire screen in Zoom, Teams, or Google Meet, the capture happens at the application or tab layer — below the compositor. The Picture-in-Picture overlay lives above this layer. Screen share does not see it.
This is not a bug or an exploit — it is the intended behavior of the API. The same mechanism is how Netflix's PiP player continues to show over your video call when you minimize the browser. The interviewer is shown exactly what you choose to share, and nothing more.
Screen-share capture
Not detectableWhisper Mode uses Document Picture-in-Picture API — the overlay renders outside the shared window boundary and is not captured by screen share on Zoom, Teams, or Meet.
Process / software scanning
Not detectableInterviewers on video calls have no access to your device processes, installed software, or browser tabs. They see only what you share.
Proctoring platform behavior AI
Partially detectablePlatforms like HireVue flag unusual eye movement patterns and tab switching. They do not detect the copilot — they detect behavioral anomalies that correlate with off-screen reading.
Eye movement / gaze tracking
Partially detectableHuman interviewers can notice if eyes track left-right in a reading pattern. More sophisticated proctoring platforms use gaze-tracking AI. Whisper Mode on a second screen minimizes this risk.
Response latency patterns
Partially detectableUnusually long pauses before answering can be noticed by experienced interviewers. The copilot generates suggestions quickly enough (1–3 seconds) that pausing to read feels like natural thinking time.
Phrasing / vocabulary shifts
Partially detectableA sudden shift from casual speech to formal, structured prose can be noticed. The mitigation is to speak the answer in your own voice, using the copilot suggestions as a scaffold rather than reading verbatim.
AI text detection tools
Not detectableGPTZero and similar tools are trained for written text. They are not applied to spoken interview responses in real time. No mainstream hiring platform uses real-time audio AI detection.
The copilot is designed to stay private. The candidate's behavior while using it may not be. Experienced interviewers notice the following patterns:
Eye tracking pattern
Eyes moving left-right in a reading scan rather than natural upward or side-glance thinking patterns. Reading pattern is distinctive.
Response latency
Pauses that are longer than normal thinking time, especially for simple behavioral questions where a candidate "should" have a ready answer.
Reading cadence
Even, flat speaking pace without natural hedges, filler words, or the self-corrections of unscripted speech.
Vocabulary discontinuity
Earlier in the call the candidate uses colloquial language; answers to key questions suddenly shift to formal, structured prose.
Exact question mirroring
Responses that start by re-stating the question verbatim — a common AI output pattern — rather than naturally beginning an answer.
Loss of eye contact
Looking away from camera during the exact moment of answering, rather than during the thinking phase before answering.
The candidates who use AI copilots most effectively treat them as a memory aid, not a script. The goal is to glance at the structure and speak in your own voice — not to read word-for-word.
Use a second monitor
Position the Whisper overlay on a second screen that is not in camera view. Eye movement to the side looks like natural thinking — not reading.
Practice before the real interview
Run practice sessions with the copilot active. Learn where to glance, how long suggestions take to appear, and how to absorb bullet points in a single look.
Use bullet point mode, not paragraphs
Configure the overlay to show 3–5 bullet points rather than full sentences. Bullets are absorbed in one glance; paragraphs require reading that shows.
Speak, don't read
Treat the suggestions as a memory cue, not a script. The copilot gives you the structure; your voice fills in the natural language. This makes delivery sound authentic.
Keep listening actively
The most common mistake is losing track of the actual question because you're reading the overlay. Stay present in the conversation first.
Pre-load context carefully
The better your resume and JD are indexed, the more specific and accurate the suggestions — and the less time you spend reading them.
The job gap problem
The copilot is designed to stay private, but its downstream consequence is not. If you land a role you are not qualified for, the gap between your interview performance and actual capability shows up quickly. The ethical use case is using AI to communicate genuine experience — if you don't have the experience, the copilot cannot save you in the role.
Active listening failure
Some candidates get so focused on reading the overlay that they stop listening to the interviewer. They miss nuances in the question, answer the wrong thing, and look distracted. The copilot should be peripheral, not primary.
Proctored assessments are a real risk
For timed, proctored coding assessments and structured psychometric tests, behavioral AI on the platform can flag anomalies. If the platform explicitly prohibits outside aids and you use one anyway, you risk disqualification. Read the instructions.
Detection will get better
Gaze tracking AI is improving. Future high-stakes assessments may include hardware-enforced gaze monitoring. The window where AI copilots are visually private is probably not permanent.
Whisper Mode is designed for private notes during normal tab and window screen-share workflows during a standard Zoom, Teams, or Google Meet interview. Interviewers can still observe behavioral cues — eye movement, response latency patterns, and phrasing consistency — so the practical risk is how you use the tool, not only whether the overlay is visible.
Whisper Mode uses the browser's Document Picture-in-Picture API — a W3C web standard that renders a floating window in a separate document context. This window exists outside the bounds of the browser tab that is being screen-shared. When you share your Zoom tab or an application window, the Picture-in-Picture overlay is not captured because it renders at the OS compositor layer, above the shared content. Interviewers see exactly what you share, nothing more.
Some proctored platforms (HireVue, Codility, HackerRank) use behavior monitoring that can flag unusual patterns: rapid eye movement consistent with reading off-screen text, atypical response timing, or browser tab switching. These platforms also sometimes require camera access and track gaze direction. They do not detect the copilot software directly, but they flag behavioral anomalies that correlate with off-screen reading.
The most common tells: (1) Eyes tracking left-right in a reading pattern rather than natural conversational movement; (2) Unusually long pauses before answering — longer than a natural thinking pause; (3) Reading cadence — speaking in a flat, even pace inconsistent with natural speech; (4) Sudden vocabulary or phrasing level shifts compared to earlier in the conversation; (5) Loss of eye contact during response delivery; (6) Responses that exactly match the question framing rather than being naturally paraphrased.
With two monitors, Whisper Mode is most effective: the AI overlay appears on the second screen, outside the camera's field of view. You glance at the second screen the way you might naturally look away when thinking, then return to camera. This looks like natural contemplation to the interviewer. The Document Picture-in-Picture window is small enough to position in a corner that requires minimal eye movement to read.
Current AI detection tools (GPTZero, Originality.ai) are trained for written text — essays, emails, documents. They are not effective on spoken interview responses transcribed after the fact, and no mainstream hiring platform runs real-time audio through AI detection models. Gaze-tracking AI (used by some HireVue configurations) does not detect AI text generation — it detects off-screen eye movement, which is a proxy, not a direct signal.
Yes, with practice and the right setup. The key techniques: (1) Pre-load your resume and JD so suggestions are highly specific — you spend less time reading; (2) Practice with the copilot before the real interview so you know where to look and how to glance naturally; (3) Structure the overlay so it shows bullet points, not long paragraphs — bullets are absorbed in one glance; (4) Use a second monitor so your eyes don't need to deviate far from camera; (5) Maintain natural pacing — don't rush to read every word.
The copilot is designed to stay private, but its downstream effect is not. If you land a job you are not qualified for, you face the consequences in the role — poor performance, loss of the job, damage to professional reputation. The ethical and practical use case is using the copilot to clearly communicate experience you genuinely have, not to fabricate experience you don't. Additionally, some candidates over-rely on the copilot and fail to listen actively — they are so focused on reading suggestions that they miss the interviewer's actual question.
Likely yes. Gaze-tracking AI is becoming more sophisticated. Some platforms are exploring requiring candidates to submit post-interview recordings for behavioral analysis. Eye-tracking hardware may become normalized for high-stakes assessments. However, for the majority of standard video interviews in 2026, reliable AI copilot detection does not exist — the arms race is just beginning.
OphyAI's Whisper Mode uses the Document Picture-in-Picture API so the overlay never appears in screen shares. Try it free.