One peer-reviewed study, and its 5% is not what the headline suggests
Marulanda, Lines, Kassa, Smithwick & Sullivan · Univ. of Kansas, Simplar Foundation, UNC Charlotte, Arizona State · published 2025-12-30 [13][14]
The only peer-reviewed measurement of an AI takeoff tool we could find compares Togal.AI with On-Screen Takeoff on architectural floor plans: a two-story fire station and a multistory hotel (two floor plans each), plus a control case and some initial tests on scanned sets, all performed by a single first-time user. It reports roughly 70% time savings (71% across the two main cases) and that “accuracy remained within a 5% margin compared to On-Screen Takeoff, with minimal discrepancies in smaller quantities.”
Read the method before you quote the number. The 5% is the percentage difference between Togal's values after manual adjustment and the manual On-Screen Takeoff values: the user corrected every quantity whose gap was 5% or more, then compared. The baseline is another tool's takeoff, not a surveyed ground truth, and the figure is post-correction agreement, not the AI's raw hit rate. The authors say so themselves: the single user may influence the results, running the tools in sequence may have favoured one over the other, lower-quality scans reduced accuracy, and relying solely on the AI-automated results “is not advisable.”
The most careful measurement that exists, and it measures human-plus-AI agreement with a manual tool on clean drawings, then recommends review anyway.