Report

The state of AI in production

Forja Studio

·

·

58 pages

Our annual report on language models in live systems, drawn from a survey of 240 teams running them in production. Covers deployment patterns, evaluation practice, cost structure, and the gap between pilot and production. This year the gap widened: more teams ran pilots than last year, and a smaller share of those pilots shipped. The teams that did ship look different in consistent, reportable ways.

The second edition of our annual survey, covering 240 teams across Europe and North America running language models in live systems.

What the report covers:

  • Deployment patterns by industry and workload

  • Evaluation practice, from none to continuous

  • Cost structure across inference, retrieval, monitoring, and review

  • The pilot-to-production gap, measured for the second consecutive year

Headline numbers: 61 percent of surveyed pilots did not reach production, up from 54 percent last year. Teams of three or fewer shipped at twice the rate of teams of eight or more. Teams with any form of continuous evaluation shipped at nearly three times the rate of teams with none.

Method: structured survey with 240 responses, followed by 31 interviews. Self-reported figures were cross-checked against interview accounts where both existed. The full questionnaire is reproduced in the appendix.

CITATION

Forja Studio. (2025). The state of AI in production. Annual report.

CONFRONTO

More research

All

Strategy

Engineering

Process

Opinion

Field notes

FORJA

MENU

Create a free website with Framer, the website builder loved by startups, designers and agencies.