EY refers to the global organization, and may refer to one or more, of the member firms of Ernst & Young Global Limited, each of which is a separate legal entity. Ernst & Young Global Limited, a UK company limited by guarantee, does not provide services to clients.
How EY can help
-
Discover how EY's technology transformation team can help your business fully align technology to your overall purpose and business objectives.
Read more
Software engineering firms are struggling with runaway AI costs, and for many companies, there appears to be no relief in sight. The industry’s situation today is a challenge for leaders because AI was predicted by many to be a cost-saver, not a cost-escalator. The thinking was that AI would transform the very structure of these companies, giving basic code-writing tasks to AI agents overseen by a small group of highly skilled human engineers. But the reality to date has been different.
AI-assisted development changes the cost profile for software engineering because costs are tied to usage shape, not tool access. As AI adoption has grown, so, too, have costs. Token consumption is exploding, usage patterns are inefficient, and consumption-based pricing makes spend unpredictable.
Today, some development teams are spending the equivalent of a junior engineer’s salary every month on tokens.1 And costs continue to spiral upward, to a level where AI spend is becoming a major operational issue at many companies.
It’s true that software development has never been inexpensive. Traditional software development tools have visible license costs. Build services have infrastructure costs. Cloud environments have metered resource costs. But AI development adds a new layer — inference costs created by prompts, context, tool calls, model routes, generated output, retries and agent behavior — all using tokens.
The hidden cost curve appears when normal developer behavior becomes repeated model consumption — ordinary workflow patterns that were never designed as economic decisions. As developers rely more on AI, sessions grow longer. Context grows. The assistant reads more files, tool calls multiply and the agent retrieves or verifies its own work. The final answer is reviewed, corrected and regenerated. None of these steps are unusual or unnecessary. In fact, they are normal components of AI development. That’s where the problem lies, because inference costs are difficult to both predict and manage, especially at companies where AI use is considered mandatory.