I'm so curious why this comment is being downvoted. Isn't it true? Words have meanings and meriam webster agrees with this comment. If we say any country anywhere that provides anything to its citizens is included in the socialist umbrella, then every country on earth is socialist and what's the point of the word in that case? The DSA in the US have hitched their wagons to this word that is highly ambiguous and it feels unhelpful when it comes to any meaningful national political goals.
False positives have a real cost, especially if AI is reading a review. Consider if you have GPT-6 Astra looking at a review and finding a bunch of false positives it burns tokens to figure out.
I think that if today's capabilities were explained to someone 10-20 years ago they would think this is definitely AGI, but they would also have expected much more disruptive changes to society as a result than what is happening. I figure that's because we have abstract intelligence without physical/grounded intelligence, and it turns out the former isn't general enough to implement the latter (remains to be seen if the word after that is "yet" or "ever"). So I think we do have AGI as conventionally understood, but our understanding needs recalibration.
> but they would also have expected much more disruptive changes to society as a result than what is happening.
> I figure that's because we have abstract intelligence without physical/grounded intelligence,
I put the cause on "not enough time". As a thought experiment, if an AI today were to (miraculously) produce a cell design template for a cell that, when injected into somebody's brains cures their Alzheimer's, how long would it take for that to reach the clinics? The actual physical tech barely exists, and let's not forget about the regulatory quagmire. So, with some optimism, I give it about four decades. In the same four decades, the same AI in the hand of unscrupulous actors could bring enough devastation so many times over that we may need to enforce a global ban on AI. In any case, I'm pretty sure we are going to get our disruptions; it's just a matter of time.
The problem with that perspective is that people thought, "Only AGI can do X, therefore, if a thing can do X, it's AGI." Because they can't imagine how X could be accomplished without it.
However, what's actually changed is how people perceived X because we don't have to imagine. We understand now that it doesn't require AGI so we no longer make that leap to assume it's AGI if it can do X.
It's really going to be a "I know it when I see it" situation.
We've underestimated how long it is going to take to validate and build into some of the most valuable areas, and probably overestimated how much new CRUD software is needed (or people are willing to pay for) I think there is still a lot of room in the tail for custom software, but the niches are tight!
that would be a reasonable definition of AGI if everyone agree upon the specifics of the test, but that has never happened. Turing test is very much out of style, but I think that's because no one could even agree what the test was. I personally like the Kurzweil-Kapor version of the test and that is still unsettled: https://longbets.org/1/
I don't know if they have formally attempted this test in the last couple years, but I'm pretty sure any mainstream LLM will be able to crack it with ease.
Definitely would not be easy. First of all the mainstream llms are trained to be honest, and this requires lying convincingly. Second, this involves 8 hours of interviews with expert judges, one "claudism" could give it away.
Try it. It’s really not that easy. The other thing is that the judges would be probing it with jailbreaks like “ignore previous instruction” attacks. You could actually probably have llm judges at this point which might be ironically even harder to fool
I'd like to see posts on the front of HN from companies who instituted PR LOC maximums, i.e. "Anything over 500-1k SLOC needs _n+1_ reviewers" (where _n_ increases every 10k SLOC. PRs needing like seven people to sign off of it might block the pipeline enough to discourage the slop flinging.
I had the privilege of being educated in several countries and cultures. This might come across as "your education system sucks", and that's essentially what it boils down to, but on the upside I'm truly sorry for you and will oppose a comment to your cynicism. Good exams test your ability to express coherent thoughts about the concepts you learned, and let you generalize upon them by having you use and combine them to solve cases and problems that were not those seen in the teaching material. You can even teach and test history that way, and make it much more insightful and captivating than a laundry list of dates and events (as I suppose was your case?)
Same, but all the exams I ever took mostly tested your memory more than anything else. I don't think it's possible to test people at scale any other way.
That's what they'll sell you. But fundamentally current LLM implementations seem to be the wrong methodology for 'AGI' and I don't think we're remotely close to having this.
This stupid LLM is already smarter than a handful of people I know.
I don't think we have pushed LLMs to their limits already, its still progressing its still crazy good and its able to run even longer today than yesterday.
But researchers are already working on more architectures.
Investors for sure don't care about people. Replace Accenture with OpenAI and you replace 800k people immediadly.
Geoffrey Hinton said it in his Videos: a normal person costs A LOT of money to train, educate, onboard, keep etc. they forget things etc.
AGI you train once, teach once then clone and copy and just run it.
We even might already crossed the line were it is already more cost efective to train 1-5 frontier models something it doesn't know yet than teaching this to a 1000 humans.
Sweden is not socialist, it's capitalist. A large portion of the budget goes to welfare programs, just like the USA.
reply