4 ms·
Pangram links several third party evaluations [1] that all seem to agree with 90+% detection, and very close to 0% false positives. Obviously those could be che
by gibspaulding 22d ago
Pangram links several third party evaluations [1] that all seem to agree with 90+% detection, and very close to 0% false positives. Obviously those could be cherry picked, but I’ve yet to see arguments to the contrary that actually include any supporting data. (E.g here’s a prompt that will get Claude to spit out text that pangram doesn’t detect or here’s an article authored in 2018 that pangram says is AI.) Detection is an arms race, so this could change (though I’d expect in the direction of false negatives), but right now it seems like defense is winning.
[1] https://www.pangram.com/blog/third-party-pangram-evals https://www.pangram.com/blog/third-party-pangram-evals
- CharlesW 22d ago> Obviously those could be cherry picked… The "could" isn't required here, since Pangram calls many of the evals "collaborations", and some received direct support. Results that don't favor Pangram as being among the least-worst of a questionable product category* surely wouldn't be featured. * https://mitsloanedtech.mit.edu/ai/teach/ai-detectors-dont-work/ https://mitsloanedtech.mit.edu/ai/teach/ai-detectors-dont-wo...
- gibspaulding 22d agoThe most recent reference in that article appears to be from 2023.