Report
The state of AI in production
Forja Studio
·
·
58 pages
Our annual report on language models in live systems, drawn from a survey of 240 teams running them in production. Covers deployment patterns, evaluation practice, cost structure, and the gap between pilot and production. This year the gap widened: more teams ran pilots than last year, and a smaller share of those pilots shipped. The teams that did ship look different in consistent, reportable ways.
The second edition of our annual survey, covering 240 teams across Europe and North America running language models in live systems.
What the report covers:
Deployment patterns by industry and workload
Evaluation practice, from none to continuous
Cost structure across inference, retrieval, monitoring, and review
The pilot-to-production gap, measured for the second consecutive year
Headline numbers: 61 percent of surveyed pilots did not reach production, up from 54 percent last year. Teams of three or fewer shipped at twice the rate of teams of eight or more. Teams with any form of continuous evaluation shipped at nearly three times the rate of teams with none.
Method: structured survey with 240 responses, followed by 31 interviews. Self-reported figures were cross-checked against interview accounts where both existed. The full questionnaire is reproduced in the appendix.
CITATION
Forja Studio. (2025). The state of AI in production. Annual report.
CONFRONTO