7 ms·
"Here is a challenge, designed to be unsolvable or so. We'll give you a bazillion dollars if you complete the challenge, and, in the meantime, we will use your
by ECCME 2y ago
"Here is a challenge, designed to be unsolvable or so. We'll give you a bazillion dollars if you complete the challenge, and, in the meantime, we will use your attempts to train an as AI that will be worth the cost!!"
- skrebbel 2y agoDid you even try the puzzles? They’re not particularly “unsolvable”.
- ECCME 2y agoARC-AGI: "here are some pretty simple puzzles, we'll give you a million dollars to solve them!" Human: "They're quite challenging, this might be a trick to engage activity for the purpose of training models." skrebbel: "You're stupid".
- educasean 2y agoDid you try the puzzles?
- ECCME 2y agoNo. What is the purpose of this competition? Unlikely that the reason for it is to pay out an enormous reward, right? Easy or not easy, the fortune is only rewarded to the system that solves the puzzles. The reward is too valuable to be given away easily. Ipso facto, solving the puzzles is deemed challenging by those who present the competition.
- echoangle 2y agoAre you writing this under every challenge with a monetary reward? The point of the challenge is that it is hard to do for an AI and easy for a human. Of course it is not easy to solve, that’s the point of the challenge. But the puzzle itself is not very hard.
- geor9e 2y agoNo, you missed the point. The striking thing about ARC is the puzzles are super easy, for humans. The average person solves 85% of the tasks, but the worlds best LLMs are only solving 5%. The challenge is to simply make an AI score as well as the average human.
- ECCME 2y ago[flagged]
- hackerlight 2y agoAmazon Mechanical Turk workers, who might not be 100 average IQ but wouldn't be far off that.
- gota 2y agoIn the most charitable interpretation of this comment - I can understand the feeling, when so much of social media interactions are in the form 'It's post a picture of you as a baby, 10 year old, and current age!'. Those and many other instances can bring out excessive skepticism But the people involved in this haven't signaled that they are in that path, either in the message about the challenge (precisely the opposite) or seemingly in their careers so far So I guess I don't share the concern but a better way to phrase your comment could be - "how can we be sure the human-provided solutions won't turn out to be just fodder for training a RL model or something that will later be monetized, closed and proprietary? Do the challenge organizers provide any guarantees on that?"