4 ms·
In the framing of your original comment: A: Mturkers B: ChatGPT C: Experts => ChatGPT better approximates the labelling of D by experts than mTurkers Which
by nmca 3y ago
In the framing of your original comment:
A: Mturkers
B: ChatGPT
C: Experts
=>
ChatGPT better approximates the labelling of D by experts than mTurkers
Which is a coherent and interesting conclusion.
Edit: also, please forgive the snark in my first response
- YeGoblynQueenne 3y ago>> Edit: also, please forgive the snark in my first response No need to apologise! Your snark wasn't overboard, I thought. Anyway, big girl, can take it :) >> ChatGPT better approximates the labelling of D by experts than mTurkers That could be a "coherent and interesting conclusion" but it's not what the article really claims. The article's title is I think hedging its bets, by being very precise about who, exactly, was outperformed by ChatGPT, although it still manages to be vague about how ChatGPT outperformed the Mechanical Turks. I'm also really doubtful that two political science graduate students can be considered as "experts" in the annotation tasks they were called to perform, which were, again if I got that right, about content moderation. "Experts" in this setting would be people with experience in moderating discussion boards etc. I don't see that this was the case with the two students that provided the initial annotation.