[edit]
How Good Are My Predictions? Efficiently Approximating Precision-Recall Curves for Massive Datasets
Proceedings of the 33rd Conference on Uncertainty in Artificial Intelligence, PMLR R15:411-420, 2017.
Abstract
Large scale machine learning produces mas- sive datasets whose items are often associ- ated with a confidence level and can thus be ranked. However, computing the precision of these resources requires human annotation, which is often prohibitively expensive and is therefore skipped. We consider the problem of cost-effectively approximating precision- recall (PR) or ROC curves for such sys- tems. Our novel approach, called PAULA, pro- vides theoretically guaranteed lower and up- per bounds on the underlying precision func- tion while relying on only O(log N) anno- tations for a resource with N items. This contrasts favorably with $\Theta$($\sqrt{}$N log N) anno- tations needed by commonly used sampling based methods. Our key insight is to capital- ize on a natural monotonicity property of the underlying confidence-based ranking. PAULA provides tight bounds for PR curves using, e.g., only 17K annotations for resources with 200K items and 48K annotations for resources with 2B items. We use PAULA to evaluate a subset of the much utilized PPDB paraphrase database and a recent Science knowledge base.