Research › Business Verdicts
Artul.ai Research LibraryStudy No. 96Business VerdictsUpdated 2026-08-28

Margins Are Expanding, and So Is the Confidence: Reading 44% of Calls

By Artul.ai Research Group · n = 73,217 earnings calls · First published 2026-08-28
Abstract

We examine 73,217 earnings calls from a 165,182-call corpus spanning 1990-2026 where the margins signal was read as expanding, a 44.3% share (95% CI 44.1%-44.6%). Speakers on these calls score higher on confidence (7.57 vs 7.21), specificity (7.72 vs 7.56), and promotion (5.21 vs 5.05), and lower on evasion (2.55 vs 2.70) and stress (2.06 vs 2.43). Guidance behavior differs sharply: 32.1% raised guidance versus 21.1% in the base set, while 6.3% lowered it versus 11.6%. Overrepresented phrases include Pricing Recovering and Deferred Revenue Growing; underrepresented include Results Worse Than Direction. Median forward returns were -6.2% versus -7.2% in the base.

Key findings
  • Margin-expanding calls account for 44.3% of the 165,182-call corpus, with a 95% confidence interval of 44.1% to 44.6%.
  • Confidence on these calls averages 7.57 versus 7.21 in the base set, while stress is lower at 2.06 versus 2.43.
  • Guidance was raised on 32.1% of margin-expanding calls versus 21.1% of base calls, and lowered on 6.3% versus 11.6%.
  • The phrase Pricing Recovering appears at a 1.26x lift on these calls, while Results Worse Than Direction appears at 0.66x.
  • Median forward returns for the 11,186-call returns sample were -0.0618 versus -0.0716 in the base, with 40.7% beating versus 39.5%.

1Introduction

Margins are the line analysts quote first and executives defend hardest, so how management talks about them is a window into tone on the calls that matter. If nearly half of all earnings calls feature margins described as expanding, that framing is not a rarity but a default posture, and its linguistic fingerprints deserve scrutiny. For anyone who follows calls closely, the question is whether the confident, specific, promotion-heavy delivery that accompanies margin talk also travels with guidance behavior and market outcomes. This study profiles 73,217 calls where margins was read as expanding against the full 165,182-call corpus, comparing language, guidance actions, recurring phrases, and forward returns.

2Data & methodology

The corpus comprises 165,182 earnings-call transcripts published between 1990 and 2026, each scored independently by a large language model on an identical 37-field battery: seven categorical business verdicts, eight 0–9 behavioral meters, and twenty yes/no judgments. The study group is defined as calls where margins was read as expanding (n = 73,217; 44.3% of the reference set, 95% Wilson interval 44.1%–44.6%). Baseline figures use all scored calls. Market outcomes join a fixed sample of 22,449 calls with twelve-month total returns in excess of SPY, measured from the first close after each call; this sample skews toward liquid U.S. names and is reported as descriptive history only.

3Results

Margin-expanding calls sound like it: confidence runs 7.57 versus 7.21, specificity 7.72 versus 7.56, and stress sits at 2.06 versus 2.43, with evasion lower at 2.55 versus 2.70. Guidance actions skew positive, with 32.1% raising guidance versus 21.1% in the base and only 6.3% lowering it versus 11.6%. The phrase table matches the mood: Pricing Recovering and Deferred Revenue Growing both show a 1.26x lift, while Results Worse Than Direction (0.66x) and Scale-Dependent Advantage Claims (0.67x) are underrepresented. The annual share is cyclical, dipping to 38.7% in 2020 and peaking at 50.2% in 2024. Median forward returns of -0.0618 versus -0.0716 in the base show only a modest gap, and the beat rate is 40.7% versus 39.5%.

Table 1. Mean behavioral scores (0–9 scale), study group versus baseline
MeterStudy groupBaselineΔ
Candor6.876.86+0.01
Evasion2.552.70-0.15
Specificity7.727.56+0.16
Stress2.062.43-0.37
Promotion5.215.05+0.16
Confidence7.577.21+0.36
Table 2. Guidance actions, study group versus baseline
ActionStudy groupBaseline
Raised32.1%21.1%
Maintained48.7%48.8%
Lowered6.3%11.6%
Withdrawn1.5%2.7%
Table 3. Co-occurring battery signals ranked by lift (group prevalence ÷ baseline prevalence)
SignalLiftIn groupBaseline
Pricing Recovering1.26×27.1%21.5%
Deferred Revenue Growing1.26×11.1%8.9%
Results Worse Than Direction0.66×33.7%51.1%
Scale-Dependent Advantage Claims0.67×7.5%11.1%
201539.19%
201643.44%
201745.96%
201844.18%
201941.07%
202038.71%
202148.39%
202240.72%
202346.46%
202450.24%
202545.86%
Figure 1. Share of all analyzed calls matching the study definition, by year.
Table 4. Twelve-month excess total returns versus SPY (descriptive history, not a signal)
StatisticStudy groupReturns sample
Median excess return-6.2%-7.2%
Interquartile range-24.9% to +12.8%
Share beating SPY40.7% (95% CI 40%–42%)39.5%
Observations11,18622,449
Table 5. Most recent calls matching the study definition
TickerQuarterCall dateCall grade
SBFGQ2 20252025-07-25A
USCBQ2 20252025-07-25B+
HCAQ2 20252025-07-25C
AONQ2 20252025-07-25C
NWGQ2 20252025-07-25B+
FFICQ2 20252025-07-25B+
OMFQ2 20252025-07-25A
GBCIQ2 20252025-07-25A

4Discussion

A careful reader should conclude that calls with expanding margin language carry a distinctive, more confident tone and a guidance profile tilted toward raises. That is a description of co-occurrence, not a mechanism: nothing here shows that the language causes outcomes or that reading margins as expanding predicts returns. The returns gap is small and the beat-rate gap of 40.7% versus 39.5% is close to the base. The annual trend moves with market conditions rather than pointing anywhere forward. Treat the phrase lifts and language deltas as descriptive texture on how management frames good margin news, not as signals to trade on.

5Limitations

The margins reading and all language scores are AI-generated fields and are noisy, so measurement error is baked into every delta. The returns sample covers 11,186 calls from a base of 22,449 and is skewed toward liquid names, limiting generalizability. Our own forward tests falsified directional prediction, so no claimed edge exists here. Additionally, LLMs partially remember famous stocks' histories, which can contaminate any backtest built on these annotations. The confidence intervals are wide enough at the phrase level that small lifts should be read cautiously. See the full methodology, including the C1 pattern’s forward-test failure and the LLM-memorization finding.

Cite this study Artul.ai Research Group (2026). “Margins Are Expanding, and So Is the Confidence: Reading 44% of Calls.” Artul.ai Earnings-Call Research Library, Study No. 96. https://artul.ai/research/when-margins-are-expanding-earnings-calls

Related studies

Steady As She Goes: What Happens When Guidance HoldsCapex Up, Stress Down: Reading 64,183 Calls Where SpGuidance Is a Vibe: Calls Where Demand Reads as AcceBacklog Is Growing, and So Is the Confidence: EarninMargins Are Fine, Thanks for Asking: What ContractioShow Me the Money (Ask): Calls Where Pricing Read as
Not investment advice. Artul.ai publishes AI-generated earnings-call quality grades and expected-volatility estimates — never buy or sell recommendations. We tested over 1,600 predictive hypotheses against 165,000 transcripts; the honest result, including what failed, is documented in our methodology.