| ▲ | jgalt212 6 hours ago |
| when was the turing test beaten? |
|
| ▲ | simonh 5 hours ago | parent | next [-] |
| 1966 https://en.wikipedia.org/wiki/ELIZA_effect It turns out the limiting factor isn't how sophisticated algorithms are, it's how gullible humans are. |
| |
| ▲ | ashetr 5 hours ago | parent | next [-] | | What does that have to do with the Turing Test? The TT has clear rules: There are judges that have a dialogue with anonymized AI/humans. The humans cannot cheat and impersonate a machine, they have to act normally. The AI obviously should try to sound human. No AI would pass this test with experienced judges. | | |
| ▲ | pibaker 2 hours ago | parent | next [-] | | Colloquially the Turing test is just a stand in for "can a human mistake a computer for a person." No need to overcomplicate it. | |
| ▲ | stavros 5 hours ago | parent | prev | next [-] | | You think current frontier models couldn't pass for a human on an online chat? You and I have very different perceptions of reality. | | |
| ▲ | stickfigure 3 hours ago | parent | next [-] | | I think that if you have a long enough chat, yeah, I think you can figure out who's meat. The original rules for the TT specified a short interaction, but I can probably accelerate it by pasting in large code snippets to force early compactions. | |
| ▲ | frollogaston 4 hours ago | parent | prev [-] | | Sounds like something we could settle right here and now. | | |
| ▲ | stavros 4 hours ago | parent [-] | | Ha ha! Fool! You've been talking to an LLM all this time! Your wife is actually Haiku 4.5. | | |
| ▲ | frollogaston 4 hours ago | parent [-] | | Huh, should've known it was odd for my wife to always say I'm right (as the boomers would say) |
|
|
| |
| ▲ | mingus88 4 hours ago | parent | prev | next [-] | | You’ve moved the goalposts. You can always say “oh well these judges don’t have the experience to catch this type of AI. The fact that you have to insert this qualifier, to ensure you always have a way to discredit the test, pretty much shows to me that we’re beyond it. | |
| ▲ | frollogaston 5 hours ago | parent | prev [-] | | The test doesn't say that the judge has to be experienced. But I also don't care if some random gullible person can't tell the difference. Nothing passes the Turing Test for me yet. Edit: Also doesn't say anything about who the human test subject is | | |
| ▲ | simonh 5 hours ago | parent | next [-] | | Of course, and I'm sure OP wouldn't disagree with you, it was clearly a joke for emphasis. Some people round here need to clean and calibrate their humour detectors more often. | | | |
| ▲ | cortesoft 2 hours ago | parent | prev [-] | | How can you be certain you haven’t failed a Turing test? | | |
| ▲ | frollogaston 2 hours ago | parent | next [-] | | I've never done a test. That means 1. you know it's a test 2. you get like 5 back-and-forths or 5 minutes 3. you have a human subject to compare to. But pretty sure I'd pass anyway, as the judge or subject. | |
| ▲ | 2 hours ago | parent | prev [-] | | [deleted] |
|
|
| |
| ▲ | beardedetim 5 hours ago | parent | prev | next [-] | | I think of this and the book the author wrote Computer Power and Human Reason every time I try to talk to product about the short comings of LLMs | |
| ▲ | applicative 3 hours ago | parent | prev | next [-] | | It was always a bad test, despite the greatness of Turing. The human organism is built to 'project' humanity onto anything available; apart from this none of the peculiar phenomena of the so-called 'modern human' is even intelligible, even the possibility of science. I bring all that is in me onto you as soon as you seem to be saying something, and reciprocally. We do this at the drop of a hat, and all specifically human life depends on it. But this power shows its 'gullibility' with 'gods' as also with Eliza. I am not snide about it because it is overreach by something the significance of which is overwhelming , but one is indeed amazed by the failure to reflect on the part of the ones eg giving LLMs rights - to take extreme case of a very widespread cultus - as if /they/ were the rational party, not ancients placating the storm god. | |
| ▲ | bell-cot 5 hours ago | parent | prev [-] | | Yeah...but in context, "gullible" seem a bit pejorative. Humans are also hopelessly incapable of sensing radioactivity, methanol in their alcoholic drinks, carbon monoxide, and a great many other things that our ancestors just didn't encounter much. Though we're pretty good at sizing up a person's emotional balance/maturity and competence at familiar tasks. So maybe have an old blacksmith watch the AI/robot interact with horse owners for a while, then shoe their horses, and see how well it does. |
|
|
| ▲ | Someone1234 5 hours ago | parent | prev | next [-] |
| The EU just had to pass a law to force companies to disclose if a customer service agent is AI or Human. It is beaten. https://commission.europa.eu/news-and-media/news/safer-and-m... |
| |
| ▲ | AnotherGoodName 3 hours ago | parent | next [-] | | Also the endless online debates of ‘is this post made by ai? What about those images, that video or that music?’. | |
| ▲ | stavros 5 hours ago | parent | prev | next [-] | | Nowadays, I prefer AI CS agents to humans. I just had a chat with an AI yesterday, it understood me perfectly even when I made mistakes, I was impressed. In contrast, humans tend to paste me the same barely-relevant macro over and over, no matter how much time I spend explaining my issue. | | |
| ▲ | rogerrogerr 3 hours ago | parent [-] | | Yeah, at least LLMs read everything you write (for now). Human first level support agents are incredibly frustrating if you have to explain anything with more than one logical step. |
| |
| ▲ | maximilianthe1 5 hours ago | parent | prev [-] | | Customer service is very different. Crappiest audio quality possible & scripted answers all the way down.
Almost like humans are forced to behave like machines. |
|
|
| ▲ | bnchrch 5 hours ago | parent | prev | next [-] |
| According to Psychology today, April this year by GPT 4.5 https://www.psychologytoday.com/ca/blog/the-digital-self/202... |
| |
|
| ▲ | goda90 5 hours ago | parent | prev | next [-] |
| I think it was determined that the Turing test is too easy because humans are too easily fooled. |
| |
|
| ▲ | mdp2021 5 hours ago | parent | prev | next [-] |
| The Turing test is more complex than what gets suggested. And the "popularized" version is faulty also since it uses an ideal, abstract human judge (like the "spheroidal economic agent"). But if you want to add declinations to the said popularized image of the Turing test, you may add Maxim Lott's IQ tests at trackingai.org . Between the end of 2024 and the beginning of 2025 LLMs reached an equivalent IQ of 100, for example. |
|
| ▲ | 3836293648 5 hours ago | parent | prev | next [-] |
| ~1960 ELIZA beat the Turing test and then everyone forgot about it. Humans are just really terrible at recognising robots. |
|
| ▲ | mdp2021 5 hours ago | parent | prev | next [-] |
| Or can we reframe it: when did humans start losing the (so-called) "Turing test". I think there are elements showing lowering of performance and expectation. |
|
| ▲ | olmo23 5 hours ago | parent | prev | next [-] |
| 2001 according to Wikipedia https://en.wikipedia.org/wiki/Turing_test |
|
| ▲ | shric 5 hours ago | parent | prev | next [-] |
| It depends who takes the test. I am not yet, to my knowledge, fooled by AI. I've tried [1] and I almost 100% detect which is the AI. I really want to convince myself I have failed, does anyone know of a better site/resource for this? I know it might be moving goalposts but I would consider AI to have passed in a well and truly undisputed manner when [2] is resolved. But in a more practical sense, if AI can impersonate humans so well today then why are state of the art frontier models so obviously AI when they create PRs, commit messages, documentation, etc. Are the companies deliberately making them unnatural? [1] https://turingtest.live/ [2] https://www.metaculus.com/questions/11861/date-when-ai-passe... |
| |
| ▲ | jayGlow 5 hours ago | parent | next [-] | | we might need to bring back the Voight-Kampff test. anthropic at the very least is introducing a water making system to Claude which might make them more identifiable to humans as well as much easier to detect for machines. | |
| ▲ | dmd 5 hours ago | parent | prev [-] | | “advanced LLMs like GPT-4” | | |
| ▲ | shric 5 hours ago | parent [-] | | > "advanced LLMs like GPT-4" Not sure where you're quoting from but if it's the metaculus question comments, many of them are from 2023. The consensus is it will resolve in 2029. I believe it will not resolve before 2035. | | |
|
|
|
| ▲ | intrasight 5 hours ago | parent | prev | next [-] |
| 2050 I'm guessing |
|
| ▲ | Rover222 3 hours ago | parent | prev [-] |
| are you living under a rock? |