Research › Hypotheses Tested
Artul.ai Research LibraryStudy No. 34Hypotheses TestedUpdated 2026-08-28

Hypothesis Tested: Answers go deeper than the script

By Artul.ai Research Group · n = 200 earnings calls · First published 2026-08-28

When an executive stops reading the script and actually answers the question, does the call get better? To find out, we had an AI read 144,500 earnings-call transcripts from the 165,000-call corpus and score each one on candor, specificity, and related measures. The 200 calls where answers went noticeably deeper than the prepared remarks form the group studied here.

These calls scored higher on specificity (7.76 vs 7.65) and confidence (7.57 vs 7.35), with slightly lower evasion (2.42 vs 2.66). They co-occur most with "Volume About to Step Up" at 1.55x the base rate (39.5% vs 25.5%), and less with rehearsed-sounding calls (0.72x). Guidance was raised in 33.5% of cases versus 24.4% in the base. Post-call returns were less negative: median -5.0% versus -6.9% overall.

These AI-read scores are noisy and the pattern is rare, so treat the profile differences cautiously. Return figures are descriptive history from a 22,449-call sample, not a trading signal; our own forward tests falsified directional prediction. Language models also partially remember famous stocks' histories, which contaminates any backtest of this kind.

Matching Calls
200
13.8% of corpus
12m Excess vs SPY (median)
-5.0%
sample baseline -6.9%
Beat SPY
43%
baseline 41% · n=68

Behavior profile

MeterThis groupAll calls
Candor7.00/96.92/9
Evasion2.42/92.66/9
Specificity7.76/97.65/9
Stress2.09/92.33/9
Promotion5.08/95.09/9
Confidence7.57/97.35/9

Guidance actions in this group

ActionThis groupAll calls
raised33.5%24.4%
maintained46.5%51.4%
lowered7.5%12.0%
withdrawn0.5%1.0%

What travels with it

SignalLiftIn groupOverall
Volume About to Step Up1.55×39.5%25.5%
The Question Left Hanging0.66×27.5%41.7%
Calls That Read Rehearsed0.72×27.1%37.8%

Share of calls by year

YearShare
20150.09%
20160.18%
20170.23%
20180.11%
20190.01%
20200.00%
20210.13%
20220.15%
20230.24%
20240.10%
20250.00%

Recent examples

TickerCallDateGrade
TXRHQ2 20242024-07-25A
FTAIQ2 20242024-07-24A
DAOQ1 20242024-05-23B
VITLQ1 20242024-05-09A
TALOQ1 20242024-05-07B+
FSPQ1 20242024-05-01D
WECQ1 20242024-05-01A
ROCKQ1 20242024-05-01B+

This hypothesis has its own page with every matching company: see the full question.

Cite this study Artul.ai Research Group (2026). “Hypothesis Tested: Answers go deeper than the script.” Artul.ai Earnings-Call Research Library, Study No. 34. https://artul.ai/research/hypothesis-answers-go-deeper-than-the-script

Related studies

Hypothesis Tested: Identified near-term contributorsHypothesis Tested: Uncontested runwayHypothesis Tested: Coming out of the tunnelHypothesis Tested: Repeat expansion inside the base Hypothesis Tested: Earned edge, pressed harderHypothesis Tested: Second demand front open and fund
Not investment advice. Artul.ai publishes AI-generated earnings-call quality grades and expected-volatility estimates — never buy or sell recommendations. We tested over 1,600 predictive hypotheses against 165,000 transcripts; the honest result, including what failed, is documented in our methodology.