BotFlo AI Transformation Score

BATS is a transparent 0–100 benchmark that reveals how deeply and effectively a company is truly transforming with artificial intelligence.

By systematically analyzing full earnings call transcripts — including prepared remarks and Q&A — BATS evaluates strategic intent, revenue innovation, agentic capabilities, management conviction, measurable outcomes, and the critical gap between hype and execution. The result is a clear, comparable score that cuts through marketing language and shows which companies are genuinely building lasting competitive advantage in the AI era.

What makes BATS unique is that it serves a dual purpose. While delivering investor-grade insights into corporate AI maturity, it also functions as a rigorous, real-world benchmark for frontier LLMs. Earnings calls represent one of the harder reasoning challenges for AI systems: dense strategic context, forward-looking claims, financial nuance, internal contradictions, and subtle shifts in management tone. BATS tests how well today’s models can understand and reason across all of these layers.

Every score comes with full transparency — including detailed reasoning traces and direct references to the transcript — so you can see exactly why a company received its rating. No black boxes. No self-reported surveys. Just objective analysis updated periodically over each earnings season.

Whether you’re an investor looking to separate real AI leaders from storytellers, or an AI researcher evaluating model capabilities on complex business communication, BATS gives you a clear, consistent standard.

BotFlo AI Transformation Score

BATS Universal Rubric (Generic AI)

BATS is a transparent, citation-backed benchmark for how seriously a company is pursuing AI transformation, scored from the full earnings call transcript (prepared remarks and Q&A). Each dimension is scored as an integer from 0 up to its maximum; scores are summed for a total out of 100.

15 dimensions 100 points total Cross-sector Citation-backed
# Dimension Max Score bands
1AI Mention Level & Depth6
  • 0 None
  • 1-2 Light / passing mentions
  • 3-4 Moderate / multiple references
  • 5-6 Heavy + detailed throughout
2AI Strategic Centrality9
  • 0 Not mentioned as strategic
  • 1-3 Supportive / peripheral
  • 4-6 Key enabler
  • 7-9 Core pillar / requires strategy evolution
3Management Tone on AI8
  • 0 None / avoidant
  • 1-2 Cautious / measured
  • 3-5 Bullish
  • 6-8 Very bullish + transformative language + urgency
4Revenue & Innovation Focus8
  • 0 No link to revenue
  • 1-3 General mentions
  • 4-6 Specific models (freemium, consumption, AI-first ARR)
  • 7-8 Major business model shift + quantified targets
5Agentic Automation Level8
  • 0 None
  • 1-3 Basic automation / assistants
  • 4-6 Multiple agents + workflows mentioned
  • 7-8 Productized, enterprise-grade agentic systems + orchestration
6Customer Experience Transformation7
  • 0 No CX link
  • 1-3 Generic personalization
  • 4-5 AI-powered CX initiatives
  • 6-7 Full CX orchestration / enterprise transformation
7AI Infrastructure & Platform Investment7
  • 0 None
  • 1-3 Minimal / cloud usage only
  • 4-5 Significant partnerships or platforms
  • 6-7 Major custom infrastructure + acceleration (e.g. NVIDIA Foundry)
8Measurable Impact & Evidence Quality7
  • 0 No metrics
  • 1-3 General claims
  • 4-5 Some quantified metrics
  • 6-7 Detailed, specific KPIs (ARR, MAU, adoption %, multiples)
9Financial Impact, Direction & Trade-offs6
  • 0 Not mentioned
  • 1-2 Neutral / mixed
  • 3-4 Positive but vague
  • 5-6 Explicit positive impact + raised guidance despite trade-offs
10Future Plans Strength & Specificity6
  • 0 None
  • 1-2 Vague
  • 3-4 Moderate guidance / next steps
  • 5-6 Detailed roadmap or clear timing
11Hype vs. Execution Balance6
  • 0 Pure hype, no execution
  • 1-2 Hype heavy
  • 3-4 Balanced
  • 5-6 Strong execution focus with shipped results
12Governance, Risk & Ethics Depth5
  • 0 None
  • 1-2 Minimal mention
  • 3-4 Partial (brand safety, compliance, auditable workflows)
  • 5 Detailed governance framework
13Efficiency & Productivity Focus5
  • 0 None
  • 1-2 Light / vendor only
  • 3-4 Internal productivity + cost savings
  • 5 Disciplined reallocation + quantified gains
14Internal Adoption & Cultural Signals4
  • 0 None
  • 1-2 Low / anecdotal
  • 3 Medium (some metrics or programs)
  • 4 High + cultural integration
15Overall AI Maturity & Coherence8
  • 0-2 Minimal / early
  • 3-4 Developing
  • 5-6 Advanced
  • 7-8 Mature & coherent strategy
Total 100

Each score is supported by direct transcript quotes and a short explanation. Band labels describe the qualitative level assigned within each dimension's allowed range.