You hardened the software. The overlay is invisible, the dock icon is gone, the audio routes through a virtual cable. Then you blow the interview in the first ninety seconds because your eyes drift to the same spot above the webcam before every answer, and the hiring manager has seen it twice this week.
The 2026 consensus across r/cscareerquestions, r/recruitinghell, and Teamblind is blunt: Cluely, Final Round AI, Interview Coder, and OphyAI are genuinely hard to detect at the binary level. The candidates using them are easy to detect at the behavioral level. This article is about the four tells that do the outing, why each one happens mechanically, and what to actually rehearse before your next Zoom round.
The detection economy moved from the overlay to the webcam
Nobody is fingerprinting your Cluely install anymore, because they do not need to. Behavioral scoring is now a product category, and it works.
Fabric, one of the enterprise vendors in this space, claims roughly 85% detection across a study of 19,368 interviews using gaze tracking, response-timing consistency, keystroke dynamics, and language patterns. None of those signals require seeing the overlay. They require seeing you. A Blind commenter put it in one sentence: "even if some of these tools are undetectable from a technical perspective, it's obvious when someone is saying words they don't actually understand and those candidates get reported."
The economy around this has gotten specific. Cluely was founded April 20, 2025 in New York by Roy Lee, Neel Shanmugam, and Alex Chen with a "cheat on everything" launch campaign before pivoting to the softer "AI meeting assistant" framing. Patrick Shen at Columbia then shipped Truely, explicitly marketed as the anti-Cluely. The vendors are fighting over the binary. The interviewers stopped caring about the binary.
Across 19,368 interviews, using gaze, timing, keystrokes, and language patterns. No overlay fingerprinting required.
The pool using these tools is the pool they expose fastest
The entry-level cohort is about 31x the size of the senior one, and it is the exact cohort buying Cluely and Final Round AI subscriptions. In Refolk's index of professional profiles, there are 363,733 US software, backend, and full-stack engineers at the Entry Level and Senior tiers combined. 352,528 of them, roughly 97%, are entry level. The senior slice is about 11,205 people.
| Segment | Count | Source |
|---|---|---|
| US SWE / Backend / Full-Stack (Entry + Senior) | 363,733 | Refolk index |
| US SWE (Entry Level only) | 352,528 | Refolk index |
| US Senior SWE (derived) | ~11,205 | Refolk index |
| US Technical Recruiters | 115,532 | Refolk index |
| SWE candidates per recruiter | 3.15 : 1 | Derived |
| Fabric behavioral detection rate | ~85% / 19,368 interviews | interviewcoder.co |
| Final Round AI Trustpilot | 3.5 / 5 (249 reviews) | ie.trustpilot.com |
Entry-level engineers also have the shortest casual-conversation baseline before a polished answer lands, which is why polish discontinuity (tell #3 below) is loudest in the junior pool. The tool sells hardest to the people it exposes fastest. That is not an accident of marketing. It is the mechanism.
Tell #1: Reading verbatim
Reading verbatim is the easiest tell to catch and the hardest to coach out, because it is a cadence problem, not a content problem. Humans speaking spontaneously stress unexpected words, revise mid-sentence, and drop filler. Humans reading a feed stress nothing and revise nothing.
A hiring manager on Blind described a recent round like this: "a candidate was so blatantly cheating when I interviewed her. She was literally reading and couldn't answer any follow up questions." The follow-up collapse is the detector. The reading cadence is just what tipped them off to probe.
Why it happens
Cluely-class tools surface a complete answer in a side panel. The temptation to parrot it is enormous, especially under pressure, and the fastest way to hit the latency floor (see tell #4) is to just read. The tool rewards speed; the interview penalizes it.
How to fix it
- Never speak the first sentence the overlay produces. Paraphrase every answer out loud before delivering it.
- Rehearse your top ten behavioral stories cold, without any assistant running, and open every answer with your own first sentence before glancing at anything.
- If you forget what you just said, you were reading. Record a mock and listen.
Tell #2: Off-camera eye drift
Eye drift means your gaze lands in the same off-camera spot before every answer, and the interviewer notices by the third question. This is the tell behavioral-detection vendors score most aggressively because it is the most visually obvious and the easiest to automate against.
The mechanics are simple. Your overlay lives in a fixed region of your screen. Your eyes go there. You are not making awkward eye contact because you are shy; you are making awkward eye contact because the answer is in the upper-right quadrant of your left monitor, and your pupils are a laser pointer at it.
How to fix it
- Pin the overlay directly beneath your webcam. If the tool allows a transparent overlay on top of the Zoom tile, use that. Any horizontal offset from the camera lens is visible.
- Rehearse scanning the overlay in your peripheral vision instead of fixating on it. The brain can parse 20 to 30 words in a glance if you have practiced; it cannot if you have not.
- Blink and look away deliberately between sentences. Natural speakers do not stare at one spot for 40 seconds.
The deeper problem is that eye drift compounds with reading cadence. Together they are unmistakable. Separately a good interviewer still catches both.
Tell #3: Polish discontinuity from your casual baseline
Polish discontinuity is the gap between how you talk during the "tell me about yourself" warmup and how you talk during the first real question. If the warmup sounded like a human and the system-design answer sounds like a McKinsey slide, the delta is the tell.
This is the one that hurts juniors the most. If you have been in industry three years, your "casual" register already includes architecture vocabulary, so the delta is small. If you are a new grad whose casual register is "yeah so like, I kind of just used Firebase," and the next sentence is "I'd horizontally scale the write path through a sharded consumer group on Kafka with idempotent producers," your interviewer heard a different person start talking. The 352,528 entry-level engineers in the pool are the ones for whom this gap is widest.
How to fix it
- Build a glossary of the five to ten technical terms you actually use in your own voice. Deliver answers in that vocabulary, not the overlay's.
- Keep your casual warmup casual. Do not try to sound impressive in the opener; you are setting a baseline you will have to match.
- Read your own writing out loud: PRs, Slack messages, commit descriptions. That is your true voice. The answers have to sound like it.
This is also where resume prep bleeds into interview prep. If your resume is a GPT-flavored word salad and your spoken answers suddenly sound exactly like it, you have fused two tells into one. The resume should sound like you first. Refolk writes a resume from your own history and the specific posting you are applying to, so the vocabulary matches the way you actually talk about your work. If the paper and the person match, the interviewer has nothing to flag.
The tool is the easy part now. Your webcam is the hard part, and the webcam has been winning.
Tell #4: Pause-then-perfect delivery
Pause-then-perfect is the pattern where every single answer begins with 5 to 15 seconds of silence followed by a fluent, structured response. Humans don't do that. Humans ramble, correct, hedge, and arrive at structure. Cluely-class tools pause, then deliver.
Why it happens mechanically
Independent testing has clocked response delays of 5 to 90 seconds from the Cluely overlay itself. That means the "unnaturally consistent pause" is partly a product artifact, not a user hesitation. No amount of personal rehearsal eliminates a latency floor baked into the tool. Final Round AI holds a 3.5 out of 5 Trustpilot across 249 reviews, which is roughly what you'd expect from a product whose UX is one slow network call away from unusable.
How to fix it (within the limits)
- Think out loud through the pause. "Okay, so the first thing I want to figure out is the write pattern," spoken while the overlay catches up, kills the dead-air signature.
- Vary your pause length deliberately. Fast answer on question two, longer on question three, back to fast on question four. Consistency is the giveaway.
- Prepare the first 20 seconds of every common question without any assistant. Buy yourself a warmup runway so the overlay catches up invisibly.
The enforcement answer is physical, not algorithmic
Even perfect behavioral stealth has a shrinking ceiling, because the enforcement response is moving off the laptop entirely. A Blind post reports Apple will require in-person final rounds for all SWEs starting July, specifically to shut down overlay tools. Treat it as rumor, but a directionally consistent one. A separate Blind thread lists the current state cleanly: "interview coder, cluely and leetcode wizard are all detectable on hackerrank and coderpad; ultracode and final round still work for the time being but a lot of faang companies including Google are bringing in person on-sites back."
The implication for your prep calendar is specific. If the tool only gets you to an onsite where you cannot use it, you are rehearsing for a round you will never sit. The 115,532 US technical recruiters in Refolk's index (one for every 3.15 SWE candidates in the active pool) are the gatekeepers to those onsites, and they talk. A reported interview is a scorched reference.
Spend your rehearsal hours on the problems you will actually face in person: a whiteboard with a marker, a real engineer asking real follow-ups, no overlay to catch you.
The three-round rehearsal that actually fixes all four tells
A rehearsal plan that works runs three rounds, cold, before you turn any assistant on.
- Round one, baseline. Record yourself answering ten behavioral and five technical questions with no assistant. Watch it back. Note your pause length, your eye position, your natural vocabulary. That is your floor.
- Round two, assistant on, overlay only. Run the same fifteen questions with Cluely or Final Round visible but muted. Deliver answers in your own words, using the overlay as a safety net you glance at twice per answer, not a teleprompter.
- Round three, mock interviewer. Have a friend run the same questions and interrupt with follow-ups. Follow-ups are the universal detector; if you cannot defend an answer two levels deep, cut it from your rehearsal set.
If the overlay stays on in round three and you still handle follow-ups cleanly, you have earned the right to run it live. If you cannot, no amount of binary-level stealth saves you. The 85% detection figure Fabric reports is not a software problem. It is a performance problem. The volume pressure that pushes candidates to Cluely in the first place is also fixable upstream: Refolk tailors every application to the posting and scores how well you actually fit, so fewer applications earn more at-bats and the overlay stops looking necessary.
FAQ
Are Cluely and Final Round AI actually detectable in 2026?
The tools themselves are hard to detect at the binary level on most platforms, which is the consensus across r/cscareerquestions and Teamblind threads. HackerRank and CoderPad flag some older tools like Interview Coder and LeetCode Wizard, but Cluely-class overlays mostly evade binary detection. The detection that catches candidates is behavioral: gaze, cadence, polish gap, and pause pattern. Fabric alone claims 85% behavioral detection across 19,368 interviews without touching the binary.
Why do interviewers catch me even when my overlay is invisible?
Because they are not looking for your overlay. They are watching your eyes drift to the same off-camera spot, your pause-then-perfect-answer cadence, and the gap between your casual warmup voice and your suddenly-polished technical answers. Follow-up questions then confirm it: if you cannot defend the answer two levels deep, the interviewer writes it up.
Does the "pause-then-perfect" tell have a fix?
Partially. The pause is baked into tool latency (5 to 90 seconds measured on Cluely), so you cannot fully eliminate it. You can mask it by thinking out loud during the pause, varying answer speed deliberately across questions, and prewriting the first 20 seconds of every common question so the overlay catches up invisibly.
Will in-person interviews make these tools useless?
At the final-round stage, yes, and that direction is accelerating. Blind posts flag Apple reinstating in-person SWE finals starting July, and Google is reported to be following. The strategic implication is that overlay craft gets you to an onsite you cannot overlay through, so the better investment is rehearsing cold and getting more targeted applications out so you earn more real at-bats.