6 ms·
This is wild: OpenAI is basically declaring that AGI is here. https://www.theverge.com/ai-artificial-intelligence/989601/openai-gpt-6-astra-release https://www
by petilon 13d ago
This is wild: OpenAI is basically declaring that AGI is here.
https://www.theverge.com/ai-artificial-intelligence/989601/openai-gpt-6-astra-release https://www.theverge.com/ai-artificial-intelligence/989601/o...
“If we fast-forward a couple of years, and we look back and say, ‘When was it, really, that AGI was created?’ I think it’s going to be about this time, and I think it might be about this model,” OpenAI president Greg Brockman said during a Thursday press briefing. Later in the call, he added, “For me personally, I do think we’re there … I think it’s not unreasonable to feel that we are now in the AGI era.”
- tastyface 13d agoRenown liar Altman releasing a PR statement for his product declaring that AGI is here is really not noteworthy.
- rektomatic 13d agoRemember when the term "AGI" meant something? Pepperidge farm remembers
- 0xbadcafebee 13d agoI think the last re-re-redefinition of what OpenAI considered AGI was "It can mostly do the job of some people"
- paxys 13d agoNo, because it has never meant a specific thing that everyone agreed on.
- drop_star 13d agoDoes it pass the Turing test?
- bigfishrunning 13d agoDepending on the proctor, ELIZA passes a Turing test. The Turing test is an interesting thought experiment, but isn't really a good measure.
- bryan0 13d agothat would be a reasonable definition of AGI if everyone agree upon the specifics of the test, but that has never happened. Turing test is very much out of style, but I think that's because no one could even agree what the test was. I personally like the Kurzweil-Kapor version of the test and that is still unsettled: https://longbets.org/1/ https://longbets.org/1/
- drop_star 7d ago>agree upon the specifics of the test The chinese room experiment
- paxys 13d agoI don't know if they have formally attempted this test in the last couple years, but I'm pretty sure any mainstream LLM will be able to crack it with ease.
- bryan0 13d agoDefinitely would not be easy. First of all the mainstream llms are trained to be honest, and this requires lying convincingly. Second, this involves 8 hours of interviews with expert judges, one "claudism" could give it away.
- paxys 13d agoAll these problems can be fixed by a "pretend to be an average human to pass a turing test" prompt
- bryan0 13d agoTry it. It’s really not that easy. The other thing is that the judges would be probing it with jailbreaks like “ignore previous instruction” attacks. You could actually probably have llm judges at this point which might be ironically even harder to fool
- deleted 13d ago[deleted]
- sm-silversight 13d agoPrime Intellect or nothing.
- Rover222 13d agoNo, I really don't
- layer8 13d agoRemember when “Pepperidge farm remembers” meant something?
- breuleux 13d agoI think that if today's capabilities were explained to someone 10-20 years ago they would think this is definitely AGI, but they would also have expected much more disruptive changes to society as a result than what is happening. I figure that's because we have abstract intelligence without physical/grounded intelligence, and it turns out the former isn't general enough to implement the latter (remains to be seen if the word after that is "yet" or "ever"). So I think we do have AGI as conventionally understood, but our understanding needs recalibration.
- dsign 13d ago> but they would also have expected much more disruptive changes to society as a result than what is happening. > I figure that's because we have abstract intelligence without physical/grounded intelligence, I put the cause on "not enough time". As a thought experiment, if an AI today were to (miraculously) produce a cell design template for a cell that, when injected into somebody's brains cures their Alzheimer's, how long would it take for that to reach the clinics? The actual physical tech barely exists, and let's not forget about the regulatory quagmire. So, with some optimism, I give it about four decades. In the same four decades, the same AI in the hand of unscrupulous actors could bring enough devastation so many times over that we may need to enforce a global ban on AI. In any case, I'm pretty sure we are going to get our disruptions; it's just a matter of time.
- parineum 13d agoThe problem with that perspective is that people thought, "Only AGI can do X, therefore, if a thing can do X, it's AGI." Because they can't imagine how X could be accomplished without it. However, what's actually changed is how people perceived X because we don't have to imagine. We understand now that it doesn't require AGI so we no longer make that leap to assume it's AGI if it can do X. It's really going to be a "I know it when I see it" situation.
- altcognito 13d agoWe've underestimated how long it is going to take to validate and build into some of the most valuable areas, and probably overestimated how much new CRUD software is needed (or people are willing to pay for) I think there is still a lot of room in the tail for custom software, but the niches are tight!
- seemaze 13d agoRemember when The Verge was not a pay-walled visual headache?
- pluc 13d agoFind me someone who isn't paid by OpenAI who is saying the same
- redox99 13d agoI hate the term "AGI" but IMO Fable, 5.6 Sol, et al. were already AGI.
- mr_mitm 13d agoWhy does he say what he feels? Is that how leading figures in the space define AGI - a gut feeling? What are the usual definitions and how can we test for it? Is there something like a Turing test for AGI?
- enraged_camel 13d agoThey are desperately, desperately trying to make a name for themselves as the lab that first created AGI, because Anthropic's IPO is just around the corner.
- naasking 13d agoIt's not easy to test as there is no formal definition or formal criteria for AGI, only exclusionary criteria like "not X". That's why he phrased it that way, he's saying it's going to be clear with hindsight once we have a better understanding of things that this time and/or this model will be the inflection point of AGI.
- layer8 13d agoThe “I” alone is already not well-defined. That’s why.
- grumbel 13d ago> Is there something like a Turing test for AGI? There is the "Economic Turing Test", you let it find a job and earn money for itself. If it can do that reliably, across a wide range of jobs, that should fit most definitions of AGI.
- ThouYS 13d agoWasn't that part of their contract with Microsoft? Some clause stopped biting with the arrival of AGI
- tziki 13d ago"OpenAI executive hypes up new model" Don't get me wrong, the benchmark jumps are good and I'm excited to try it, but only one or two of the benchmark jumps could be described as better than incremental.
- bigfishrunning 13d agoDon't worry, they'll come up with a new acronym to mean really-real AI soon...
- glenstein 13d agoI almost feel like I need just as much healthy skepticism toward hn comments that have the automatic reflex of dismissing performance gains, as much as I need a similar form of skepticism toward AI claims. It feels like (from what I'm understanding) the harnessed result on ARC-AGI-3 is not exactly playing by the normal rules that would tell us how much of a leap this really is. Nothing wrong with harnesses, but if there's one thing they aren't, it's an indicator of generality in performance gains. So I think it's a bit of a misleading signal and we should wait for more independent vetting. I think the middle ground is that these are improvements worthy of the "GPT-6" label but still well short of a true "this is AGI moment" that would truly put the question to rest.
- IanCal 13d agoIf I’m understanding other comments the harness is just how ChatGPT and codex work already and it’s to do with how the context gets compacted - the arc-agi harness some are claiming just throws out reasoning blocks? Which feels like a huge handicap.
- geraneum 13d ago> we should wait for more independent vetting Nahh by that time they’re going to release AGI 2.
- deleted 13d ago[deleted]
- petilon 13d agoSam Altman himself has said it is not AGI unless it can discover novel physics. https://x.com/burny_tech/status/1725233117055553938 https://x.com/burny_tech/status/1725233117055553938 In the tweet Sam Altman is quoted as saying: "If (for example) super intelligence can't discover novel physics I don't think it's a superintelligence. And teaching it to clone the behavior of humans and human text - I don't think that's going to get there. And so there's this question which has been debated in the field for a long time: what do we have to do in addition to a language model to make a system that can go discover new physics?" I think this is a reasonable criteria for declaring AGI. So can GPT-6 do it? OpenAI says it has helped solve long-standing open problems in mathematics. No word on novel physics.
- Feathercrown 13d agoAGI and superintelligence are not the same thing
- petilon 13d agoHe says the same about AGI: https://www.nytimes.com/2023/11/20/podcasts/hard-fork-sam-altman-transcript.html https://www.nytimes.com/2023/11/20/podcasts/hard-fork-sam-al... Sam Altman: Let’s say we make an A.I. that is really good, but it can’t go discover novel physics. Would you call that AGI? Kevin Roose (New York Times): I probably would, yeah. Would you? Sam Altman: Well, again, I don’t like the term, but I wouldn’t call that done with the mission.
- deleted 13d ago[deleted]
- kypro 13d agoI think it's more wild people have been denying that AGI has been here for a while honestly... Today's models and agents are not quite at human-level in all contexts and across all domains, but it seems to me they very clearly are generally intelligent. If you disagree – can you name a single problem that a human can do that agent wouldn't be able to take a decent shot at which isn't limited by the hardware available it?