| ▲ | ianjbutler 7 hours ago | |||||||||||||||||||||||||||||||||||||
Sigh, the whole "obviously the turing test is solved" meme is annoying. Like, if we meant that it convincingly masquerades as a shitposter, ok. But everyone still bitches about AI slop, and everyone knows the writing is still bad. How does that even work if the turing test is obviously solved? More to the point though, if you grill SOTA models on counterfactuals, causal world-models etc, you'll trip them up in a way that actually will not work on ESL students and children. Certainly there's no way to find a person that struggles with that and is also capable of cheerful fluent erudite discussion about astrophysics with perfect grammar. Yes, it's getting harder obviously.. but detecting machines with determined, focused and intelligent interrogation remains pretty easy. If nothing else, the models are cooperative where people wouldn't be and that's a signal too. The best progress we've made is that most people do agree that this doesn't practically matter very much, i.e. we generally recognize the stakes were always overstated. But the constant vague appeals to common-sense that "of course it's a solved problem!" always feels naive or fake. | ||||||||||||||||||||||||||||||||||||||
| ▲ | johnsmith1840 3 hours ago | parent | next [-] | |||||||||||||||||||||||||||||||||||||
I was just thinking how anyone still thought AI didn't pass turing already. There's been literal papers proving average people cannot tell reliably. | ||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||
| ▲ | HDThoreaun 2 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
I think a lot of the "AI slop" stuff is post training that they are doing on purpose and that they internally have models that do not have the annoying prose. | ||||||||||||||||||||||||||||||||||||||
| ▲ | ex-aws-dude 3 hours ago | parent | prev | next [-] | |||||||||||||||||||||||||||||||||||||
Well the whole lesson learned was that the Turing Test as it was defined was way too easy, it was a bad criteria for GI because it underestimates how easily humans find meaning/patterns in things. I mean you could show people random markov chain gibberish in 1996 and they would swear they found intelligent meaning in it | ||||||||||||||||||||||||||||||||||||||
| ▲ | CamperBob2 an hour ago | parent | prev [-] | |||||||||||||||||||||||||||||||||||||
How does that even work if the turing test is obviously solved? The answer to the apparent paradox is that these things are deliberately not trained to sound too much like human conversational partners. The labs don't want the bad press that they'd get from people creating deceptively-convincing bots, or from people forming emotional bonds with them like they did with GPT-4o. If you actually RLHF'ed a frontier-grade LLM to pass a Turing test, rest assured, it could do it. "It's not x, it's y" and other goofy superficial tells do not have to be part of an LLM's response. But the last thing OpenAI wants to release is a GPT-4o with twice the IQ, so we have to put up with a lot of stupid clanker clichés. For evidence, just look back at the best conversational models from a couple of years ago, and you will probably agree that they are better at fooling humans than their newer counterparts are. | ||||||||||||||||||||||||||||||||||||||
| ||||||||||||||||||||||||||||||||||||||