3 September 2026/benchmarking/short papers
Benchmarking AI ML Applications
Generative artificial intelligence (AI) and machine learning (ML) are no longer confined to coding assistants that help developers write software faster — increasingly, they are the functionality being delivered. Recommendation engines, computer vision, fraud detection, forecasting and conversational AI are becoming ordinary line items in development portfolios.
In this short report we look at what it takes to benchmark these AI/ML application projects themselves, why today’s function point standards and the ISBSG repository are not yet equipped to do so, and what the latest industry data on AI outcomes tells us about why that gap matters.
Read this short report
