Skip to main content

Validation

88% Recall, One Attorney, 18 Hours

A new working paper ran generative AI document review and a managed active-learning workflow head to head on the same 45,004-document corpus. Same review protocol, same reference labels, same scorecard. The GenAI system won on recall, and the paired test backs it up. The rest of the scorecard — precision, the 62-to-1 effort gap, the population extrapolation — needs more qualification than the headline suggests.