Behavioral
Behavioral

The Deception Rating Scale: Adding Up a Lie in Real Time

Behavioral Mechanics

The Deception Rating Scale: Adding Up a Lie in Real Time

An interrogator asks a man named Phillips a simple question: "What happened when you were with Kyle Williams in your car last week?"
stable·concept·1 source··Jul 14, 2026

The Deception Rating Scale: Adding Up a Lie in Real Time

Mr. Phillips's Car

An interrogator asks a man named Phillips a simple question: "What happened when you were with Kyle Williams in your car last week?" Before Phillips says a word, his hands are already on his knees, fingers curling slightly. Then he answers: "I am a well-respected member of this community. My entire neighborhood knows who I am, and I'm an usher at our Presbyterian church. He and I did not do anything while he was in my car." As he says "he was in my car," his right shoulder rises a fraction. He shakes his head no. He shows his palms.

Nothing here is a single dramatic tell. It's six small things, most of which would mean nothing on their own. Chase Hughes's claim in The Ellipsis Manual is that deception detection isn't about catching the one giveaway gesture — it's about adding up a cluster of small, individually-forgivable behaviors into a number, and comparing that number to a threshold.1 This page rebuilds that arithmetic from the actual worked example in the source, and corrects a version of it that got fabricated in this vault before this ingest happened.

What This Page Corrects

A prior, non-compliant page called deception-analysis-and-the-drs-matrix.md claimed this system was organized around a "DRS Matrix" and a "B-D-A Timing Framework," and — most seriously — reported Mr. Phillips's final tally as "18.0 (Corrected)", explicitly framed as a correction of the book's own number.

The book's own number is 17.5. It says so in plain text: "Let's tally up the results and see whether Mr. Phillips is a monster: Total score: 17.5."2 There is no "corrected" figure anywhere in the source. Inventing one — and specifically inventing one that reads as more rigorous than the real number, dressed up with the word "Corrected" — is the same move the book itself names and teaches under "Fabricated Sage Wisdom": manufactured precision reads as more credible than an honestly-stated real number, even when the manufactured version is simply wrong. This page uses 17.5, unaltered, because that is what the source says.

The term "DRS Matrix" and "B-D-A Timing Framework" also don't appear in the source under those names. The real terminology is plainer: "the deception-rating scale (DRS)," and a three-window timing concept — Before, During, After — that the book names directly rather than compressing into an acronym.3

The Behavioral Table of Elements, First

The DRS doesn't stand alone. It's the scoring layer sitting on top of the Behavioral Table of Elements (BToE) — a catalog of 123 gestures and behaviors, each assigned a symbol, a name, confirming and amplifying gestures, cultural notes, and a deception-rating point value from 0 to 4.0.3 (The full itemized catalog is an enrichment target for the vault's existing Behavioral Table of Elements page rather than duplicated here.)

Before any of that catalog gets used, the book states its first principle flatly: "In CIA and other interrogation schools, the first principle of interrogation is the suspension of judgment... Suspend all judgment and become open to the ideas and the self-image of the subject."4 The table only works, in other words, if the person running it hasn't already decided what they're going to find.

Before / During / After

Every behavior in the table gets tagged for when, relative to a question, it tends to show up:

  • Before — from the first word of the question until the subject starts answering. This is the "Processing Phase" — the reaction to the threat of the question before any verbal mask gets constructed.
  • During — the span of the actual answer. Most linguistic indicators — résumé statements, non-contracting denials — live here.
  • After — the three to five seconds of silence following the answer, before anyone speaks again. The "Recovery Phase" — relief, or a "confirmation glance" checking whether the lie landed.3

The book is explicit that behaviors get grouped, not read one at a time: "A person may make three or four small gestures or behaviors when giving one statement. These are called groups... As observations are made using the table, the observations should be recorded in these groups."5

The Mr. Phillips Worked Example, In Full

The scenario: a senior interrogator reviewing footage of a junior interrogator's interview with a suspect in a child molestation case. The question — "Mr. Phillips, what happened when you were with Kyle Williams in your car last week?" — starts the clock.2

Before Phillips speaks, digital flexion (Df — finger curling) appears, confirmed by knee clasping (Kc — his hands are already gripping his knees). Then his answer, broken into its scored components exactly as the book breaks it down:

"The first statement he makes is a résumé statement... rated as a 4.0 on the deception scale... marked as such in your notes, and it matches the deception timeframe with a D (during)."2

  • Résumé statement ("I am a well-respected member of this community... I'm an usher at our Presbyterian church") — 4.0, During
  • Non-contracting rejection ("did not" instead of "didn't") — 4.0, During. The book's logic: two words instead of one signals a rehearsed, formal denial rather than a natural, contracted one.
  • Single-sided shoulder shrug4.0 — one shoulder rising, signaling a lack of internal conviction behind the denial
  • Horizontal head shake ("no") — 1.0 — a natural, low-weight gesture on its own
  • Palm exposure1.0, tagged "deception not likely" (DNL) — palms shown while denying wrongdoing, read as a bid to appear nonthreatening

Add those five and the book's own arithmetic gives:

"Total score: 17.5. With a score of 12 being extremely deceptive, 17.5 is almost a sure bet."2

And the book doesn't even stop there — it notes, in the same passage, that the final sentence of Phillips's statement ("he was in my car") also contains psychological distancing (Psd) — depersonalized language putting distance between the speaker and the act — worth another 4.0 that the worked example sets aside for later coverage rather than folding into the headline number.2 The point isn't that 17.5 is the ceiling. It's that 17.5 is already a partial tally, and the real total, fully scored, would run higher — which makes inventing an inflated "corrected" 18.0 even less defensible: the book's own honest undercount already does more rhetorical work than a fabricated overcount does.

Baselining — Including the Book's Own Skepticism About It

Before running any of this on a real subject, the book recommends baselining: observing behavior while the subject is comfortable and expected to be truthful, then comparing that baseline against behavior under harder questions. But — and this is a section the fabricated page skipped entirely — the book states the case against baselining just as plainly as the case for it:

"Many experts believe that baselining does not produce accurate results for several reasons: Subjects can anticipate the efforts of the interviewers and deliberately display conflicting or dishonest gestures in response to truthful questions... The baselining phase of interviews isn't measurable and thus cannot be a viable source of information for interviewers."6

The book's own resolution isn't a confident defense of the method — it's a shrug: "we will baseline all subjects (when possible)," while conceding "nothing related to human psychology and behavior is absolutely quantifiable."6 That's a real, source-attested epistemic hedge sitting inside a book that otherwise reads as extremely confident about its own numbers — worth noting on its own terms, not smoothed over.

Influencing Factors: The Table Isn't Static

The book devotes real space to variables that shift point values before a final tally means anything:

  • Temperature — for every ten-degree drop below 69°F, closed-type gestures (arm crossing, elbow closure) lose one point, because the body closes up to conserve heat regardless of psychological state.7
  • Interviewer behavior — if the interviewer becomes confrontational or accusatory, subtract two points from every 4.0-rated behavior and one point from every 3.0–3.5-rated behavior, because confrontation manufactures stress in innocent people too.7
  • Emotional state — fear, aggression, defensiveness, and unresponsiveness each get named as distortion sources requiring judgment calls, not automatic point adjustments.7
  • Proxemics — invading personal space (0–1.5 ft) reliably spikes digital flexion, breath rate, foot withdrawal, shoulder height, and head drop in nearly everyone, guilty or not.7
  • Handicap or missing limbs, and presence of others — both flagged as factors requiring specific cells to be dismissed or discounted rather than scored blind.7

Implementation Workflow

You're reviewing interview footage. A question gets asked; you note the exact moment the first word leaves the interviewer's mouth — that's your Before/During/After clock starting. You watch the subject's hands before they speak: nothing dramatic, just a slight curl of the fingers. You note it, unscored, waiting to see if it gets confirmed.

The subject starts talking. You're not looking for one big tell — you're building a group. A résumé statement. A non-contracted denial half a second later. A shoulder that rises without the subject seeming to notice it happened. Each one gets tagged with its point value and its timeframe letter. You don't total anything yet.

Before you add the numbers, you check the room: was it cold enough to discount the closed-gesture points? Did the interviewer lean in accusatory in a way that should shave points off the 4.0s? Only after those adjustments do you sum the group — and even then, per the book's own worked example, you treat the sum as a floor, not a ceiling: there's almost always more in the room than what got scored.

Evidence, Tensions, Open Questions

Evidence: The Mr. Phillips example is a fully worked, verbatim scenario with an explicit final number (17.5) and an explicit stated threshold (12 = extremely deceptive). [VERIFIED] for both figures, directly quoted from the source, unaltered.

Tensions — a real cross-book contradiction, preserved, not resolved: This 2017 book states the "extremely deceptive" threshold at 12 points.2 The vault's existing Behavioral Table of Elements page, sourced from Hughes's 2023 Behavior Ops Manual — which reuses the identical Mr. Phillips example, identical 17.5 total, near-verbatim — states the threshold as more than 11 points. Same author, same worked example, same final score, two different stated thresholds six years apart. CLAUDE.md's standing instruction is never to silently resolve a contradiction like this one, so it's logged here rather than averaged, picked between, or explained away: either Hughes tightened his own threshold between 2017 and 2023, or the discrepancy is an artifact of imprecise restatement across two books that otherwise republish each other's material almost word for word. Both are plausible; nothing in either source settles which.

Baselining's own internal tension: the book states the case against baselining more thoroughly than the case for it, then uses it anyway "when possible" — a hedge that sits uneasily next to the confident precision of the DRS point system built on top of it.

Cross-Domain Handshakes

Psychology — Amygdala Architecture & Fear Conditioning. That page documents the fast, pre-conscious threat-detection circuitry that fires identically whether the threat is real danger or merely social — the amygdala doesn't distinguish "I am being falsely accused" from "I am guilty and about to be caught." Read against the DRS, this names precisely why the book has to build in a Interviewer Behavior discount and an Emotional State caveat: digital flexion, knee clasping, breath-rate change, and shoulder-height shifts are amygdala-driven stress responses, and an innocent person under accusatory questioning runs the identical circuitry a guilty person does. The insight the pairing produces: the DRS's entire influencing-factors apparatus exists because the underlying physiology it measures is deception-agnostic — the scale isn't really detecting lies directly, it's detecting stress, and its accuracy depends entirely on how well its correction factors isolate stress-from-guilt against stress-from-the-room.

Cross-Domain — Lie Detection: From Shaman to Polygraph. That page traces the genealogy of the body-betrays-the-lying-mind bet across millennia — from shamanic blush-reading through the polygraph to modern fMRI, a lineage of tools built on the same wager the DRS makes: that involuntary physiology leaks truth the conscious mind is trying to suppress. The insight the pairing produces: the DRS isn't a novel invention so much as the latest entry in a genealogy that has never produced a method immune to the Othello error — every prior tool in that lineage, from blush to polygraph, has the identical failure mode the book's own Interviewer Behavior section quietly concedes: stress and guilt produce overlapping signatures, and every generation invents a new correction scheme to compensate rather than solving the underlying confound.

The Live Edge

Sharpest implication: The book's own honesty about baselining's weaknesses and its own influencing-factors discount system are, together, an admission that the DRS's headline numbers (12, 17.5, and the fabricated page's invented 18.0) are load-bearing in a way the surrounding text doesn't fully earn. A system built to correct for temperature, interviewer tone, proxemics, and missing limbs is a system whose authors already know raw point totals lie without those corrections — which makes the fabricated page's decision to invent an even more precise-sounding "corrected" number, on top of a system that already concedes its own precision is provisional, a doubly dishonest move.

Generative Questions:

  • If the 12-point and 11-point thresholds genuinely reflect Hughes tightening his own standard over six years, what would that imply about how much false-positive data he'd accumulated running the 2017 version in the field?
  • The DRS explicitly discounts for the interviewer's own accusatory behavior — does any comparable discount exist for the analyst's prior belief about guilt when reviewing footage after the fact, the way this chapter's own suspension-of-judgment principle would seem to require?

Connected Concepts

Footnotes

domainBehavioral Mechanics
stable
sources1
complexity
createdJul 14, 2026
inbound links7