Research › Hypotheses Tested
Artul.ai Research LibraryStudy No. 46Hypotheses TestedUpdated 2026-08-28

The Numbers Are Fine; Everything Else Is Pending: Flagging Unbooked Growth on Calls

By Artul.ai Research Group · n = 426 earnings calls · First published 2026-08-28
Abstract

We studied 992 earnings-call transcripts from 2015 to 2025 and identified the 426 calls (42.9%, 95% CI 39.9% to 46.0%) where management answered YES to the research hypothesis "Identified near-term contributors not yet in the reported numbers" — in effect, teasing growth not yet visible in reported results. These calls show modestly higher specificity (7.71 vs 7.66), promotion (5.24 vs 5.07), and confidence (7.47 vs 7.36), with slightly lower evasion (2.61 vs 2.66) and stress (2.31 vs 2.35). The standout language pattern is "Volume About to Step Up," which appears 1.58 times more often on these calls. Yet forward returns were not better: the median was -0.0947 vs -0.0697 for the base, and 40.1% beat versus 41.1%.

Key findings
  • 426 of 992 calls (42.9%) flagged near-term contributors not yet in reported numbers.
  • These calls score higher on promotion (5.24 vs 5.07) and confidence (7.47 vs 7.36), with slightly lower evasion (2.61 vs 2.66).
  • The phrase "Volume About to Step Up" is 1.58x more common on these calls than on others.
  • Median forward returns were -0.0947 versus -0.0697 for the base sample, with 40.1% beating versus 41.1%.

1Introduction

Every earnings call contains a quiet genre of optimism: management pointing to revenue contributors that have not yet landed in reported numbers — new volumes ramping, pipelines converting, deals about to close. For listeners, these forward-looking teasers are seductive because they suggest the printed quarter understates the business. But how common is this framing, how does the tone of such calls differ, and does anything show up in subsequent returns? Using 992 transcripts scored by Artul.ai's language models across 2015-2025, this study examines the 426 calls that answered YES to the hypothesis "Identified near-term contributors not yet in the reported numbers."

2Data & methodology

The corpus comprises 992 earnings-call transcripts published between 2015 and 2025, each scored independently by a large language model on an identical 37-field battery: seven categorical business verdicts, eight 0–9 behavioral meters, and twenty yes/no judgments. The study group is defined as calls that answered YES to the research hypothesis "Identified near-term contributors not yet in the reported numbers" (n = 426; 42.9% of the reference set, 95% Wilson interval 39.9%–46.0%). Baseline figures use the set of calls on which this question was tested. Market outcomes join a fixed sample of 22,449 calls with twelve-month total returns in excess of SPY, measured from the first close after each call; this sample skews toward liquid U.S. names and is reported as descriptive history only.

3Results

Calls flagging unbooked contributors make up 42.9% of the corpus (95% CI 39.9%-46.0%). Their language profile tilts promotional: promotion rises to 5.24 from 5.07 and confidence to 7.47 from 7.36, while evasion (2.61 vs 2.66) and stress (2.31 vs 2.35) dip slightly — a picture of smoother, more assertive delivery. The most distinctive marker is "Volume About to Step Up," 1.58x more frequent (39.4% vs 25.0%). Guidance behavior is nearly identical to the base. The returns comparison is sobering: median forward return of -0.0947 vs -0.0697, mean -0.1007, and a 40.1% beat rate vs 41.1% for the base of 382 scored calls.

Table 1. Mean behavioral scores (0–9 scale), study group versus baseline
MeterStudy groupBaselineΔ
Candor6.956.94+0.01
Evasion2.612.66-0.05
Specificity7.717.66+0.06
Stress2.312.35-0.04
Promotion5.245.07+0.17
Confidence7.477.36+0.11
Table 2. Guidance actions, study group versus baseline
ActionStudy groupBaseline
Raised25.8%23.4%
Maintained52.3%51.8%
Lowered12.7%12.7%
Withdrawn0.9%1.1%
Table 3. Co-occurring battery signals ranked by lift (group prevalence ÷ baseline prevalence)
SignalLiftIn groupBaseline
Volume About to Step Up1.58×39.4%25.0%
20150.40%
20160.32%
20170.32%
20180.31%
20190.02%
20200.00%
20210.27%
20220.47%
20230.46%
20240.19%
20250.02%
Figure 1. Share of all analyzed calls matching the study definition, by year.
Table 4. Twelve-month excess total returns versus SPY (descriptive history, not a signal)
StatisticStudy groupReturns sample
Median excess return-9.5%-7.0%
Interquartile range-30.8% to +9.6%
Share beating SPY40.1% (95% CI 33%–48%)41.1%
Observations152382
Table 5. Most recent calls matching the study definition
TickerQuarterCall dateCall grade
FRBAQ2 20242024-07-26A
TPHQ2 20242024-07-25A
MMLPQ2 20242024-07-18B
ASOQ1 20242024-06-11C+
BBYQ1 20252024-05-30C
CRGOQ1 20242024-05-20C+
BIOXQ3 20242024-05-14C+
EROQ1 20242024-05-10A

4Discussion

The honest reading is that this flag measures a rhetorical posture, not an outcome. Calls that tout contributors not yet in the numbers do sound different — more promotion, more confidence, more volume language — but the subsequent returns and beat rates are statistically indistinguishable from, and slightly below, the base. A careful reader should conclude that the signal identifies a style of storytelling, not a predictor. Nothing here establishes that such calls cause better or worse results, nor that the flagged contributors ever materialize; the data simply show how often the pattern appears and what accompanies it.

5Limitations

All fields are AI-read and noisy; tone scores and hypothesis labels inherit model error. The returns sample covers only 152 of the flagged calls against a base of 382, drawn from a broader universe of 22,449 calls skewed toward liquid names, so it is not representative. Our own forward tests falsified directional prediction from these signals. Additionally, LLMs partially remember famous stocks' histories, contaminating any backtest with leaked knowledge of outcomes. See the full methodology, including the C1 pattern’s forward-test failure and the LLM-memorization finding.

Companion page: every company matching this hypothesis is listed at the question’s own page.

Cite this study Artul.ai Research Group (2026). “The Numbers Are Fine; Everything Else Is Pending: Flagging Unbooked Growth on Calls.” Artul.ai Earnings-Call Research Library, Study No. 46. https://artul.ai/research/hypothesis-identified-near-term-contributors-not-yet-in-the-reported-numbers

Related studies

Room to Run and Nowhere to Hide: Calls With UncontesLight at the End of the Tunnel Is Often a Train: RecDon't Stick to the Script: Earnings Calls Where AnswHave Your Cake and Expand the Base Too: Repeat GrowtPressed Harder, Answered Straight: A Profile of 183 Second Front, Funded and Fueled: Who Says Yes on the
Not investment advice. Artul.ai publishes AI-generated earnings-call quality grades and expected-volatility estimates — never buy or sell recommendations. We tested over 1,600 predictive hypotheses against 165,000 transcripts; the honest result, including what failed, is documented in our methodology.