
The evidence on AI and software development is more contradictory than most headlines suggest. A rigorous randomized controlled trial found experienced developers were actually 19% slower using AI tools on familiar codebases, while separate controlled studies and large-scale surveys show real speedups on well-defined, unfamiliar-codebase tasks. What's unambiguous: AI has already reshaped who gets hired, not just how fast people work — entry-level software developer employment has dropped roughly 13-20% since late 2022, according to Stanford research using national payroll data, while senior developer employment has stayed stable or grown. For a business planning a software project in 2026, the practical impact isn't "AI makes everything cheaper" — it's a shift toward more senior-heavy teams, faster execution on narrow, well-scoped tasks, and continued human oversight requirements that don't disappear just because a tool wrote the first draft.
Most "AI is changing software development" content either oversells a productivity revolution or dismisses AI coding tools entirely. Neither position survives contact with the actual research. This article works through what's been rigorously measured, what's still genuinely uncertain, and what it means for your project's cost and timeline specifically — not for the industry in the abstract.
The most methodologically rigorous study on this topic found the opposite of what most people expect. In July 2025, the nonprofit research organization METR published a randomized controlled trial — the same experimental design used in clinical drug trials — in which 16 experienced open-source developers completed 246 real tasks on codebases they already knew well.¹ The result: developers using AI tools took 19% longer to complete their tasks than developers working without AI. Before the study, these same developers had predicted AI would make them 24% faster. Afterward, despite the measured slowdown, they still believed AI had made them roughly 20% faster — a nearly 40-percentage-point gap between perception and measured reality.
This is not the only credible data point, and it's important not to over-extend it. Other controlled research points a different direction: GitHub's own controlled experiments found developers completing a defined task 55% faster with AI assistance, particularly on unfamiliar codebases where AI's ability to quickly summarize and navigate code offers a genuine edge over manual exploration.² Google's DORA (DevOps Research and Assessment) research program, based on large-scale industry survey data, found more than 80% of developers self-reporting productivity gains from AI tools.³ Stack Overflow's 2025 Developer Survey found 84% of developers now use or plan to use AI coding tools, up from 76% the year before — near-universal adoption, whatever the actual productivity effect turns out to be.⁴
Why the results disagree so sharply comes down to what's actually being measured. METR's slowdown applied specifically to experienced developers working in codebases they already knew intimately — in that setting, AI-generated suggestions can compete with the developer's own deep institutional knowledge rather than adding to it, and reviewing and correcting AI output can cost more time than writing the code directly. The GitHub and DORA results, by contrast, capture more defined tasks and less-familiar codebases, where AI's speed at pattern-matching and code generation has more room to help rather than compete with existing expertise.
METR itself revisited the question in February 2026 with updated tooling, and reported that newer models showed some evidence of a genuine speedup — but flagged a serious selection-effect problem: 30-50% of developers invited to participate declined to take part unless they could use AI, meaning the pool of developers willing to work without it was skewed toward those who benefit least from it.¹ METR's own conclusion, stated cautiously: "AI likely provides productivity benefits in early 2026" — a real update from their 2025 finding, but not a claim that AI now makes every developer dramatically faster.
The honest takeaway is that AI's timeline impact depends heavily on the type of work, not a flat percentage you can apply to your whole project:
Boilerplate and well-defined tasks (standard CRUD operations, test scaffolding, documentation, simple UI components) see genuine, measurable speedups — this is where most of the GitHub and DORA-reported gains concentrate.
Complex, architecture-level work on established codebases shows more mixed results, and can genuinely be slower when a developer has to review, correct, and integrate AI-generated code that doesn't match existing patterns and conventions.
Novel problem-solving and system design remain largely unaffected by current tools — this is still fundamentally a human judgment task, and no credible study claims otherwise.
If a development company tells you flatly that "AI cuts our timelines by X%" without qualifying what kind of work that applies to, that's a claim worth pressing on. A more credible answer sounds like: "AI speeds up scaffolding and boilerplate meaningfully; architecture and complex integration work isn't meaningfully faster yet, and in some cases the review overhead makes it comparable to writing it directly." That's a more defensible claim, and it's the one the actual research supports.
The clearest, best-evidenced impact of AI on the software industry isn't speed — it's who gets hired. A Stanford Digital Economy Lab study led by economist Erik Brynjolfsson, published in August 2025 and titled "Canaries in the Coal Mine? Six Facts About the Recent Employment Effects of Artificial Intelligence," analyzed anonymized ADP payroll data covering millions of US workers through July 2025.⁵ The finding: employment for workers aged 22-25 in AI-exposed occupations — software development prominent among them — fell 13% relative to trend since late 2022, while employment for workers aged 26-55 in the same occupations stayed stable or grew. In later public remarks, Brynjolfsson cited an even sharper figure specifically for software developers aged 22-26: roughly a 20% relative decline.⁶
The mechanism the researchers describe is straightforward: companies aren't cutting junior salaries — they're simply not hiring for junior roles as they used to, because AI tools now handle a meaningful share of the routine, well-specified coding tasks that used to be a junior developer's primary value. Senior developers, whose value lies in judgment, architecture decisions, and institutional knowledge that AI can't replicate, have seen no comparable decline.
What this means for your project's cost structure specifically: the hourly rate tables most cost guides publish (junior/mid/senior breakdowns) are becoming less representative of how teams are actually staffed. Development companies are increasingly running leaner, more senior-weighted teams — fewer junior developers writing first-draft code reviewed by seniors, more senior developers using AI tools directly and doing the review work themselves. This can mean fewer total billed hours for comparable output, but at a higher average hourly rate, since the team mix has shifted upward in seniority. The net cost effect isn't uniformly "cheaper" — it depends on how your specific vendor has restructured their staffing model, which is worth asking about directly rather than assuming.
One factor that rarely makes it into "AI makes software cheaper" claims: AI-generated code carries a documented, elevated security risk that needs a real review process, not just a good prompt. Application security firm Veracode tested more than 100 large language models across 80 coding tasks and found that 45% of AI-generated code introduced at least one OWASP Top 10 vulnerability — the industry-standard list of the most critical web application security risks.⁷
This matters directly for cost planning: if a vendor is using AI-assisted coding to move faster, that speed gain needs to be paired with an equivalent or greater investment in code review, security scanning, and QA — not less. A development team that treats AI output as "basically done" rather than "a first draft needing the same review rigor as human-written code" is trading a short-term timeline gain for a real, measurable security liability. This is also directly relevant if your project touches any of the compliance areas we cover in our US data privacy and compliance guide — a security vulnerability in AI-generated code that leads to a data breach doesn't get treated more leniently by regulators because "AI wrote that part."
Cutting through the contradictory headline numbers, here's where the evidence most consistently supports real, practical benefit in 2026:
Scaffolding and boilerplate generation — setting up standard project structures, CRUD endpoints, and repetitive UI patterns
Test-case generation — producing a first pass of unit tests that a developer refines, rather than writing from scratch
Documentation — generating and maintaining technical documentation that often gets skipped under time pressure
Code exploration in unfamiliar codebases — this is specifically where GitHub's controlled results showed the strongest gains, since AI can summarize and navigate large unfamiliar systems faster than manual reading
Rapid prototyping — building a rough, disposable version of a feature to validate an idea before committing to production-quality architecture
Where the evidence is weakest or actively negative: deep architectural decisions, debugging in complex or highly specific business-logic contexts, and any work in a codebase the developer already knows intimately — exactly the scenario METR's controlled trial measured.
Given how contested this evidence actually is, a vendor's answer to direct questions about their AI usage tells you a lot about how carefully they're actually thinking about it:
Do you use AI coding tools, and for which parts of a typical project specifically? A vague "we use AI throughout" is less reassuring than a specific breakdown of where it helps and where it doesn't.
What's your code review process for AI-generated code? Given the documented vulnerability rate in AI-generated code, this should be a real, described process — not an assumption that AI output is review-exempt.
Has AI changed your team's staffing mix? A vendor that's thought seriously about this should be able to explain how their junior-to-senior ratio has shifted and why.
Do you disclose AI usage in project documentation or code comments where it materially affected a component's origin? Increasingly relevant for IP and audit-trail purposes on regulated projects.
This is a good addition to the vendor evaluation process covered in our guide to hiring a software development company in the USA — AI usage wasn't a standard vetting question two years ago, and now it should be.
Putting this together for a business planning a project in 2026: don't expect a flat AI discount on your quote, and be skeptical of any vendor promising one. What you can realistically expect is faster turnaround on well-defined, boilerplate-heavy portions of a project (which our software development cost in the USA guide breaks down by project type), continued human-hours-driven cost on architecture and complex integration work, and a development team that's likely somewhat more senior-weighted than it would have been three years ago. The honest question to ask a vendor isn't "how much cheaper is AI making this" — it's "where specifically are you using it, and how are you making sure the output is safe and correct." A vendor with a thoughtful, specific answer to that question is a better bet than one with an enthusiastic, vague one.
Does AI actually make software developers faster?
The evidence is mixed and task-dependent — a rigorous 2025 randomized controlled trial found experienced developers 19% slower on familiar codebases, while other controlled studies show real speedups on well-defined tasks in unfamiliar codebases.
Why do different AI productivity studies show opposite results?
The results depend heavily on what's measured — deep, familiar-codebase work shows little to no benefit or even slowdowns, while defined, boilerplate-heavy tasks and unfamiliar-codebase navigation show measurable speedups.
Has AI reduced software development costs?
Not uniformly — AI has shifted team composition toward more senior developers rather than delivering a flat cost reduction, since junior-level hiring has declined sharply while the review and correction work AI output requires still needs experienced oversight.
Is AI-generated code safe to use in production?
Not without review — testing across 100+ AI models found 45% of AI-generated code introduced at least one OWASP Top 10 security vulnerability, meaning AI output requires the same or greater security review as human-written code.
How has AI affected junior developer hiring?
Significantly — Stanford research using national payroll data found a 13% relative decline in employment for software developers aged 22-25 since late 2022, while employment for developers over 26 remained stable or grew.
Should I ask my development vendor how they use AI?
Yes — given the documented security risks and inconsistent productivity evidence, understanding a vendor's specific AI usage and review process is now a legitimate part of vendor evaluation.
Will AI eventually make software development significantly cheaper?
The current evidence doesn't support that conclusion yet — productivity gains are real but concentrated in specific task types, and the review overhead AI-generated code requires offsets some of the apparent time savings.
Akoode Technologies builds software for clients across the US, UK,Canada,UAE and India, and uses AI tools selectively — in scaffolding, testing, and documentation — while keeping architecture decisions and security review firmly in the hands of senior engineers. If you'd like a specific, honest answer on where AI would and wouldn't speed up your particular project, book a time on our calendar.
Subscribe to the Akoode newsletter for carefully curated insights on AI, digital intelligence, and real-world innovation. Just perspectives that help you think, plan, and build better.