Why we don’t count your filler words
Most speaking apps count every “um”. The interview evidence points somewhere else: what costs you is the long silence before you start — not the fillers along the way.
Everyone says “um”. The commonly cited base rate is around six disfluencies per hundred words in ordinary speech — filler words are the normal texture of unscripted talk, not a defect. Experimental interview evidence isolating filler frequency as a cost is thin. We audited it looking for a reason to coach it, and did not find one.
What the evidence actually penalises
The well-supported delivery variable is delay. In recruiter-rated real interviews, pauses before answering averaged under two seconds for the first question and grew with each one — and longer pre-answer delays both degraded answer quality and independently reduced hireability. The harm lives in the multi-second silence before you begin — not in mid-answer pausing, which interview research has barely studied, and not in filler count.
Vocal delivery as a whole does carry modest, real signal: a composite of pitch variability, rate, pauses and amplitude predicted supervisor-rated job performance at r = .18–.20 — small effects, and notably predictive of the job itself, not just the interview rating.
The advice we refuse to give
“Speak faster to sound competent” is extrapolated from persuasion research — a 1975 study of computer-resynthesised voices, which measured competence and benevolence, not interviews and not credibility. “Mirror the interviewer” fares no better: the payoff to ingratiation depends on who is rating you — self-presentation styles that pay off in the Western interview context are, in the researchers’ own words, highly selective, and impression management correlates far more strongly with interview ratings (r = .47) than with actual job performance (r = .15).
And as scoring policy, never as a measured claim: we do not let accent, fluency or hesitation count feed a competence score. Second-language speakers hesitate more because word retrieval costs more, not because they know less — coaching them to “cut every um” trains anxiety, not articulation. That policy is written into the same rubric this product scores with.