Professional Examiners Trained to APA Standards
140+ Professional Testing Locations Across the U.S. & Canada
Trusted by 10,000+ Clients, Attorneys & Organizations
LieDetectorTest.com Private & Confidential Polygraph Provider
Research ledger
The Comparison Question Test: Does It Work and If So How?

HomePolygraph Research › The Comparison Question Test: Does It Work and…

Catalogue entry · Laboratory & Analog Studies

The Comparison Question Test: Does It Work and If So How?

Heinz Offe; S. Offe — Law and Human Behavior,

2007Published
Experimental participants in CQT paradigmSample size
4Cited by
Key findings

Investigated the theoretical mechanisms underlying CQT effectiveness. Examined the role of comparison questions in calibrating physiological responses. Provided experimental evidence for the psychological processes driving CQT accuracy.

Abstract

Investigation of the theoretical mechanisms underlying CQT effectiveness, examining the role of comparison questions in calibrating physiological responses.

Methodology

Experimental study examining physiological response patterns to relevant and comparison questions. Analyzed whether comparison questions create the differential response patterns predicted by CQT theory.

Detailed summary

Offe and Offe investigated the fundamental question of how and why the CQT works at a mechanistic level. By examining the role of comparison questions in creating differential physiological response patterns, they provided experimental evidence for the psychological processes that drive CQT accuracy. Their work contributed to the theoretical understanding of why guilty individuals show larger responses to relevant questions while innocent individuals show larger responses to comparison questions.

Implications for polygraph practice

Understanding the mechanism behind CQT effectiveness strengthens its scientific credibility and helps optimize test design. The finding that comparison questions create predictable differential patterns validates the fundamental CQT architecture.

Comprehensive study analysis

An in-depth, original analysis of this research study's methodology, findings, and significance for the polygraph profession.

Background & Context

The Comparison Question Test has been the dominant polygraph technique in criminal investigations since the 1940s, yet a fundamental question had long persisted: Does the CQT work based on the theoretical mechanisms described by its proponents, and if so, how? This 2007 study by Heinz and Susanne Offe represents a critical experimental investigation into the foundational assumptions underlying CQT effectiveness.

Traditional CQT theory holds that comparison questions serve a crucial calibration function—creating differential response patterns where guilty individuals show larger physiological reactions to relevant questions, while innocent individuals show larger responses to comparison questions. However, critics had questioned whether these theoretical mechanisms actually drive CQT accuracy or whether successful outcomes resulted from other factors entirely. This study directly tested whether manipulating comparison questions through explanation and discussion would produce the effects predicted by theory.

Research Design & Methodology

The researchers conducted a mock crime study involving 65 participants—35 who chose to participate as guilty subjects and 30 as innocent subjects. The mock crime scenario involved the theft of a monetary voucher (€50), providing a realistic context for examining physiological responses during polygraph testing.

The experimental design manipulated two key testing conditions to examine their impact on CQT effectiveness:

  • Explanation condition: Whether comparison questions were explained during the pretest interview or not
  • Re-discussion condition: Whether comparison questions were re-discussed between charts, which had no effect on identification rates
  • Physiological measures: Standard CQT physiological recording including electrodermal, cardiovascular, and respiration responses
  • Subjective ratings: Ratings of subjective stress due to relevant and comparison questions as indicators of question significance

Statistical analysis used a three-factor ANOVA (guilt×explanation×re-discussion) with total score of physiological measures as the dependent variable. The comparison questions used probable-lie format, asking about past behaviors "prior to 1999" to create temporal separation from the relevant questions about the mock crime.

Results & Key Findings

Identification rates of approximately 90% for both guilty and innocent participants were achieved in groups with explanation of comparison questions compared to groups without explanation. This represents the study's most striking positive finding regarding CQT accuracy when proper procedures are followed.

The physiological data analysis revealed important patterns:

  • Main effect of guilt: F(57,1)=44.89, p=.000, partial η²=.440, demonstrating strong discrimination between guilty and innocent participants
  • Interaction effect: The guilt×explanation interaction indicating higher negative scores for guilty participants with explanation missed significance (F(57,1)=3.425, p=.069, partial η²=.057)
  • Re-discussion manipulation: Re-discussing comparison questions between charts had no effect on identification rates
  • Question significance: The significance of comparison questions was hardly affected by the different testing conditions, and when effects were detectable, they contradicted theoretical expectations in their direction

These findings challenged conventional wisdom about how the CQT operates. While accuracy was high, the significance of comparison questions was hardly affected by testing conditions, and when effects were detectable, they contradicted theoretical expectations in their direction.

Discussion & Significance

The Offe and Offe findings present a paradox that has significant theoretical implications: the CQT achieved excellent accuracy (~90%) despite the fact that manipulating comparison questions did not produce the theoretically predicted effects on their subjective significance. The study shows that calibration of comparison questions, which depends on examiner skills, may not be as relevant to examination accuracy as previously believed—the significance of different question types results almost exclusively from the fundamental difference in significance of relevant questions for guilty versus innocent participants, not from manipulation of comparison question significance by the examiner.

This suggests that CQT effectiveness may stem primarily from the inherent salience of relevant questions to guilty individuals rather than from the examiner's ability to artificially elevate the importance of comparison questions for innocent examinees. The research provides experimental evidence that standardized procedures can be highly effective, potentially reducing reliance on individual examiner skill in manipulating psychological set.

The high accuracy rates achieved with proper explanation procedures validate the fundamental CQT architecture while simultaneously questioning traditional explanations of its mechanism. This has spurred ongoing theoretical development, including Ginton's later work on "relevant issue gravity" as an alternative explanatory framework that better accounts for these empirical findings.

Limitations & Considerations

Several methodological limitations affect the generalizability of these findings. The Offe and Offe study had relatively few subjects and thus had relatively low statistical power to find small effects, which may explain why some theoretically predicted interactions failed to reach statistical significance.

  • Self-selection bias: Volunteers were permitted to decide whether they wanted to participate as guilty or innocent subjects, rather than random assignment, potentially introducing personality differences between groups
  • Laboratory setting: Mock crime scenarios lack the genuine stakes and emotional intensity of real criminal investigations
  • Participant pool: Volunteers differ from actual criminal suspects in motivation, stress levels, and potentially in deceptiveness capability
  • Single crime scenario: The voucher theft represents only one type of offense; results may vary with different crime types

Practical Applications

The practical takeaway for polygraph examiners is clear and reassuring: proper explanation of comparison questions during the pretest interview is essential and produces excellent accuracy. The 90% accuracy rate achieved with adequate explanation validates current best practices in CQT administration and suggests that standardized protocols can be highly effective without requiring exceptional examiner artistry in psychological manipulation.

For consumers and legal professionals, this research provides qualified support for CQT validity when administered according to established protocols. However, the finding that comparison question manipulation doesn't work as theoretically described raises questions about traditional examiner training emphasis. The results suggest that transparency and clear explanation—rather than subtle psychological manipulation—may be the key to effective CQT administration, potentially supporting more standardized, less examiner-dependent procedures in operational settings.

Read the original study

The analysis above is original editorial content based on our review of this research. For the complete study including full data, methodology details, and author discussion, access the original publication below.

Cited by [4]

Other studies in our database that reference this paper.

Related research

Other studies in this category that may be of interest.

Join Our Examiner Network

APA-trained examiners using validated techniques can apply to join the LieDetectorTest.com network.

Apply now →

Keep reading the ledger.

Every peer-reviewed study on polygraph and deception detection we track — catalogued, searchable and citable.

Need to book now? Our online booking system is open 24/7. Speak directly with our team about your test or booking.