Artul.ai Research LibraryStudy No. 34Hypotheses TestedUpdated 2026-08-28
Hypothesis Tested: Answers go deeper than the script
By Artul.ai Research Group · n = 200 earnings calls · First published 2026-08-28
When an executive stops reading the script and actually answers the question, does the call get better? To find out, we had an AI read 144,500 earnings-call transcripts from the 165,000-call corpus and score each one on candor, specificity, and related measures. The 200 calls where answers went noticeably deeper than the prepared remarks form the group studied here.
These calls scored higher on specificity (7.76 vs 7.65) and confidence (7.57 vs 7.35), with slightly lower evasion (2.42 vs 2.66). They co-occur most with "Volume About to Step Up" at 1.55x the base rate (39.5% vs 25.5%), and less with rehearsed-sounding calls (0.72x). Guidance was raised in 33.5% of cases versus 24.4% in the base. Post-call returns were less negative: median -5.0% versus -6.9% overall.
These AI-read scores are noisy and the pattern is rare, so treat the profile differences cautiously. Return figures are descriptive history from a 22,449-call sample, not a trading signal; our own forward tests falsified directional prediction. Language models also partially remember famous stocks' histories, which contaminates any backtest of this kind.
Matching Calls
200
13.8% of corpus
12m Excess vs SPY (median)
-5.0%
sample baseline -6.9%
Beat SPY
43%
baseline 41% · n=68
Behavior profile
| Meter | This group | All calls |
|---|
| Candor | 7.00/9 | 6.92/9 |
| Evasion | 2.42/9 | 2.66/9 |
| Specificity | 7.76/9 | 7.65/9 |
| Stress | 2.09/9 | 2.33/9 |
| Promotion | 5.08/9 | 5.09/9 |
| Confidence | 7.57/9 | 7.35/9 |
Guidance actions in this group
| Action | This group | All calls |
|---|
| raised | 33.5% | 24.4% |
| maintained | 46.5% | 51.4% |
| lowered | 7.5% | 12.0% |
| withdrawn | 0.5% | 1.0% |
What travels with it
| Signal | Lift | In group | Overall |
|---|
| Volume About to Step Up | 1.55× | 39.5% | 25.5% |
| The Question Left Hanging | 0.66× | 27.5% | 41.7% |
| Calls That Read Rehearsed | 0.72× | 27.1% | 37.8% |
Share of calls by year
| Year | Share | |
|---|
| 2015 | 0.09% | |
| 2016 | 0.18% | |
| 2017 | 0.23% | |
| 2018 | 0.11% | |
| 2019 | 0.01% | |
| 2020 | 0.00% | |
| 2021 | 0.13% | |
| 2022 | 0.15% | |
| 2023 | 0.24% | |
| 2024 | 0.10% | |
| 2025 | 0.00% | |
Recent examples
| Ticker | Call | Date | Grade |
|---|
| TXRH | Q2 2024 | 2024-07-25 | A |
| FTAI | Q2 2024 | 2024-07-24 | A |
| DAO | Q1 2024 | 2024-05-23 | B |
| VITL | Q1 2024 | 2024-05-09 | A |
| TALO | Q1 2024 | 2024-05-07 | B+ |
| FSP | Q1 2024 | 2024-05-01 | D |
| WEC | Q1 2024 | 2024-05-01 | A |
| ROCK | Q1 2024 | 2024-05-01 | B+ |
This hypothesis has its own page with every matching company: see the full question.
Cite this study
Artul.ai Research Group (2026). “Hypothesis Tested: Answers go deeper than the script.” Artul.ai Earnings-Call Research Library, Study No. 34. https://artul.ai/research/hypothesis-answers-go-deeper-than-the-script
Not investment advice. Artul.ai publishes AI-generated earnings-call quality grades and expected-volatility estimates — never buy or sell recommendations. We tested over 1,600 predictive hypotheses against 165,000 transcripts; the honest result, including what failed, is documented in our
methodology.