Users appreciate the new PNAS paper analyzing institutional context of legal AI benchmarking, praising its collaborative authorship and calling the findings deeply interesting.
Based on 2 visible X reactions from 2 accounts; directional sample.
Ask a question below.
Published answers will appear here.
Thanks to @NeelGuha for taking the lead on this paper. Great working with him, Andy Zhang, @christinestsang, @JulianNyarko, and Dan Ho. Article is available open access!
@chrmanning @PNASNews This is deeply interesting.
There are 1000s of AI papers on benchmarks. Nearly all are on the design of scientific benchmarks. Our new @PNASNews paper addresses the institutional context. What happens when consumers & regulators use benchmarks? How should the system be designed? https://www.pnas.org/doi/10.1073/pnas.2509757122
Users appreciate the new PNAS paper analyzing institutional context of legal AI benchmarking, praising its collaborative authorship and calling the findings deeply interesting.
Based on 2 visible X reactions from 2 accounts; directional sample.
Ask a question below.
Published answers will appear here.