Research › Hypotheses Tested
Artul.ai Research LibraryStudy No. 41Hypotheses TestedUpdated 2026-08-28

Hypothesis Tested: Engine explained, runway named

By Artul.ai Research Group · n = 103 earnings calls · First published 2026-08-28

When a company explains why a metric moved and names the runway ahead, calls read differently: more candor, more specificity, less stress. This study tests that pattern across 165,000 earnings-call transcripts read by an AI, isolating 103 calls where both moves appear together — about 26.3% of the 391-call base group. For anyone who parses management language for signal, the question is whether this combination marks genuinely better communication or just better-rehearsed stories.

The profile shifts are modest but consistent: specificity runs 7.85 vs 7.63 and confidence 7.76 vs 7.27, while stress drops to 1.96 from 2.42. Guidance skews favorable — 32.0% raised vs 21.0% in the base, and only 7.8% lowered vs 15.1%. Co-occurrence lifts cluster around reassurance and scale: "Skeptic Reassured" appears at 91.3% vs 68.3%, and "A Tiny Fraction of the Market" at 40.8% vs 30.4%. But the returns comparison undercuts the story: 41 calls with forward returns show a median of −6.8%, actually better than the −11.0% base median yet hardly a triumph, with 36.6% beating versus 36.2%.

These fields come from AI reads of transcripts and are noisy measures of tone, not ground truth. The return figures are descriptive history from a 22,449-call sample, not a trading signal — our own forward tests falsified directional prediction. Language models also partially remember famous stocks' histories, which contaminates any backtest built on their outputs. Treat the pattern as a description of how these calls read, not a forecast.

Matching Calls
103
26.3% of corpus
12m Excess vs SPY (median)
-6.8%
sample baseline -11.0%
Beat SPY
37%
baseline 36% · n=41

Behavior profile

MeterThis groupAll calls
Candor6.92/96.91/9
Evasion2.52/92.71/9
Specificity7.85/97.63/9
Stress1.96/92.42/9
Promotion5.30/95.07/9
Confidence7.76/97.27/9

Guidance actions in this group

ActionThis groupAll calls
raised32.0%21.0%
maintained55.3%51.2%
lowered7.8%15.1%
withdrawn0.0%1.3%

What travels with it

SignalLiftIn groupOverall
Volume About to Step Up1.47×37.9%25.8%
Early Products Growing Fast1.38×49.5%35.8%
A Tiny Fraction of the Market1.34×40.8%30.4%
Skeptic Reassured1.34×91.3%68.3%
Results Worse Than Direction0.45×21.4%47.8%
The Question Left Hanging0.59×26.2%44.2%
When the CFO Dominates0.66×9.7%14.6%

Share of calls by year

YearShare
20150.11%
20160.06%
20170.05%
20180.09%
20190.01%
20200.00%
20210.05%
20220.11%
20230.12%
20240.07%
20250.00%

Recent examples

TickerCallDateGrade
CRGOQ1 20242024-05-20C+
WRBYQ1 20242024-05-09A
ECPGQ1 20242024-05-08B
AZEKQ2 20242024-05-08B+
LINCQ1 20242024-05-06B+
AESQ1 20242024-05-03C+
TTIQ1 20242024-05-01A
LTRXQ3 20242024-04-29C

This hypothesis has its own page with every matching company: see the full question.

Cite this study Artul.ai Research Group (2026). “Hypothesis Tested: Engine explained, runway named.” Artul.ai Earnings-Call Research Library, Study No. 41. https://artul.ai/research/hypothesis-engine-explained-runway-named

Related studies

Hypothesis Tested: Identified near-term contributorsHypothesis Tested: Uncontested runwayHypothesis Tested: Coming out of the tunnelHypothesis Tested: Answers go deeper than the scriptHypothesis Tested: Repeat expansion inside the base Hypothesis Tested: Earned edge, pressed harder
Not investment advice. Artul.ai publishes AI-generated earnings-call quality grades and expected-volatility estimates — never buy or sell recommendations. We tested over 1,600 predictive hypotheses against 165,000 transcripts; the honest result, including what failed, is documented in our methodology.