For more than three decades, the standard way to make a person measurably stressed in a laboratory has been to sit them in front of an unresponsive panel and ask them to sell themselves for a job. An article collecting accounts from 46 researchers who have run that experiment reports something the protocol never accounted for: the people playing the panel are carrying a burden of their own.
The article appears in Comprehensive Psychoneuroendocrinology and was written in honor of Clemens Kirschbaum, the German psychologist whose team introduced the Trier Social Stress Test in 1993. Rather than presenting new data, the authors compiled anecdotal reports from researchers about participant and panel member reactions, problems in running the protocol, and recommendations for designing future stress studies.
Forty-Six Researchers Wrote Down What Happens Behind the Panel Table
The central observation is straightforward and, in the authors' framing, overdue. The test places a substantial burden not only on participants but also, in a different way, on the researchers serving as panel members.
That is not a small population. One methodological review counted more than 4,000 studies using the test between 1993 and 2007 alone, and every session requires trained staff to sit in judgment and withhold every ordinary social cue: no nodding, no smiling, no encouragement.
The paper raises a question the field has not answered. If the panel role is itself demanding, does imposing stress on another person produce a physiological stress response in the person doing the imposing? Nobody has measured it systematically.
The accounts collected go beyond that single question, covering participant reactions, practical difficulties in running the protocol as written, and suggestions for how future stress induction studies should be designed. The contributions came in response to an open request, which means the sample reflects who chose to reply rather than a representative cross-section of everyone who has ever run the test.
A Job Interview Designed to Go Badly on Purpose
The protocol is deliberately simple. In the original published description, first reported in 1993, a participant gets an anticipation period of about ten minutes, then a test period of about ten minutes in which they deliver a free speech and perform mental arithmetic in front of an audience. The speech is usually framed as an argument for why they should be hired. Notes are taken away. When the speech ends early, the committee sits in silence until the time is up. The arithmetic, typically serial subtraction, restarts from the beginning on every error. A camera and microphone are part of the setup, and saliva is collected before, during, and after to track the rise and fall of cortisol.
The committee is usually two adults, and a modified version exists for children from about the age of seven. Two ingredients do the work: social evaluative threat and uncontrollability. A review of the protocol's principles and practice notes that other laboratory stressors, including public speaking on its own, tend to produce variable or absent responses in the hypothalamic-pituitary-adrenal axis. Qualitative work published in PLOS One confirmed that participants experience both components as stress-inducing.
Control conditions have been engineered to match the physical and cognitive load without the threat. A placebo version uses free speech and simple arithmetic while removing the evaluation and the unpredictability.
The Open Question Nobody Has Measured
There is a plausible biological basis for the question the paper raises. Empathic and vicarious stress responses are documented, and a person maintaining deliberate coldness toward someone visibly struggling is doing effortful emotional labor. Whether that translates into measurable cortisol release in panel members is unknown.
The practical implication is not that the test should be abandoned. It is that the panel is a variable rather than a fixture. If panel members habituate, burn out, or drift toward warmth over hundreds of sessions, that could quietly affect how strongly the test works across a long study, without anyone recording the drift in the methods section.
What This Means for the Thousands of Studies Built on the Test
Methodological scrutiny of this protocol is not new. A systematic review of the methodology found no consensus on when saliva samples should be taken, how many, or how far apart. Across the 35 studies it examined closely, researchers reported 38 different collection times, with the first sample drawn anywhere from immediately before the test to 280 minutes ahead of it.
Adding panel effects to that list matters because the test underpins a large share of what is known about how acute stress affects memory, immune function, decision-making, and mental health risk. If the intensity of the stressor varies with who is sitting behind the table, effect sizes across studies may reflect that as much as the biology under investigation.
The paper's content is anecdotal by design. These are collected recollections from experienced researchers, not measured outcomes, and the authors present them as a starting point for questions worth testing rather than as findings. The article itself is worth reading for anyone who has ever volunteered for a psychology study and wondered what the people behind the desk were thinking.
Key Questions Answered
What is the Trier Social Stress Test?
A standardized laboratory procedure introduced in 1993, in which a participant delivers a mock job interview speech and performs mental arithmetic in front of an unresponsive committee.
What did the new article report?
Drawing on accounts from 46 researchers, it reports that the procedure burdens not only participants but also the researchers who serve as panel members.
Is this based on measurements?
No. The article compiles anecdotal reports and recommendations rather than new experimental data, and the authors present it as such.
Why does the panel role matter scientifically?
If panel members change how they behave over time, the strength of the stress induction could vary between sessions and studies, affecting how well results compare.
What remains unknown?
Whether imposing stress on another person produces a measurable physiological stress response in the person doing it. The authors raise this as an open question.
Why does this test get used so heavily?
It reliably activates the hypothalamic-pituitary-adrenal axis by combining social evaluative threat with uncontrollability, which few laboratory procedures achieve consistently.