Remix.run Logo
▲ tptacek 2 hours ago

When Timnit Gebru described LLMs as "stochastic parrots", she was making a deep point about what she believed to be the limitations of the technology. Her argument claimed that while we perceived legible text coming from ChatGPT, we were actually being fooled by our innate pattern-matching instincts. Notably, at the time she wrote this, it wasn't uncommon to see LLMs collapse into piles of gibberish, lending credence the the idea that there was never any actual meaning in the output other than what we readers brought to it.

It's safe to say that argument has collapsed utterly. Modern LLMs virtually never collapse into endless babbling loops. Far more importantly, they solve objectively hard problems. An LLM didn't merely inspire a mathematician to spot a counterexample to the Jacobian Conjecture that had been in front of his nose the whole time; it found the counterexample. These things are unquestionably generating meaningful outputs.

(That obviously doesn't mean that you can simply trust LLM outputs!)

At this point, trying to dunk on frontier models by calling them "stochastic parrots" is saying more about you than about the models. Moreover, it suggests you're not only out on a limb about the capabilities of LLMs, but also on what even the skeptic literature about it was saying.

You should know what "stochastic parrot" actually meant; you shouldn't be using the term just because you think it sounds snazzy.

▲MisterMunchkin an hour ago | parent | next [-]

I see them collapse when faced unusual problems. They’ll go down a weird rabbit hole and then just keep going, even if their idea is completely irrational and not working.

They’re incapable of changing their mind, because they’re a Markov chain. They’re cool and they’re good at doing things, but they’re not sentient in the slightest.

▲tptacek an hour ago | parent | next [-]

Not at all the same idea. People also rathole into doomed approaches. What we're talking about here is literally LLMs producing line noise, or English words that don't fit into sentences.

(I don't think LLMs are "sentient".)

▲mr_mitm an hour ago | parent | prev | next [-]

I've seen this reflex quite a few times now. To diminish the impressive capabilities of the latest LLMs, one resorts to the assertion that they're not sentient, or not truely intelligent, or various other semantic diversions. No one claimed that. GP was all about usefulness.

▲card_zero 37 minutes ago | parent [-]

Then GP misrepresented the "On the Dangers of Stochastic Parrots" paper, which spoke about meaning. But GP also said "meaning" several times, so no.

There's some haziness here about the difference between comprehensible output and comprehending. To put it mildly. I mean that's been the whole question for five or six years.

▲ an hour ago | parent | prev [-]
[deleted]
▲girvo 2 hours ago | parent | prev | next [-]

I really wish the article in question hadn’t used that phrase, because you’re right. But the rest of the article is also correct in the dangers facing us from this technology as an accelerant regardless, and I think it’s much more interesting to talk about that: but this entire comment thread will dunk on the stochastic parrot paragraphs instead.

The funny thing is though, I think it’s an in group signalling mechanism. You need the anti-AI crowd to know that you’re anti-AI, and this is a rapid way of doing so, to get it upvoted here.

▲hbcdbff an hour ago | parent | next [-]

I’m glad the article used the phrase because it was a useful indicator of the author’s underlying biases

▲user43928 an hour ago | parent | prev [-]

The article is extremist nonsense where the author argues billions are going to die due to climate change in the next decades, and that Altman and Musk are hoping AI can until then replace human labor in order to keep their living standard in such a future.

It's not like the "stochastic parrot" thing was the only problem here.

▲cluckindan an hour ago | parent | prev | next [-]

An LLM did not do those things! An agentic harness around an LLM did those things.

▲tptacek an hour ago | parent [-]

This is a distinction that only matters if you're having a philosophical argument, but the claim I'm addressing from this article isn't philosophical.

▲cluckindan an hour ago | parent [-]

It is a practical distinction. An engine is not a car.

▲AnimalMuppet an hour ago | parent | next [-]

For a mechanic, that's a practical distinction. For most people, it isn't.

▲cluckindan 3 minutes ago | parent [-]

I don’t see people driving engines around, since that is impossible. You actually need the rest of the car around the engine to have it function as a drivable vehicle.

It is a pretty practical distinction.

▲verdverm 29 minutes ago | parent | prev [-]

the most widely used API surfaces are already shaped around chat, it's different and most analogies will fall short in some way

▲verdverm 30 minutes ago | parent | prev | next [-]

Sherry Turkle is far more articulate on these things in her latest book Artificial Intimacy (which is about much more than the relationship chatbots)

▲card_zero 2 hours ago | parent | prev | next [-]

Well, that may be what it was coined to mean, what does it mean now? Stochastic = statistical, parrot = using training data.

https://en.wikipedia.org/wiki/Stochastic_parrot

"a metaphor that frames large language models as systems that statistically mimic text without real understanding". I suppose you're defending them as having understanding. It's not "safe to say that argument has collapsed utterly", you're part of a campaign to reject the meme. On the other side, the meme is popular as a way to be rude about AI, because it's a cutting insult.

▲caaqil an hour ago | parent [-]

> what does it mean now?

It means whatever you want it to mean, that's the best part. To me, it's best used as it was intended, but primarily in the contexts where the stochastic parrots are plowing through the ivory towers we built mostly for idiot savants.

▲card_zero an hour ago | parent [-]

I don't even think the claim (about what she was saying) is true. It's still used with the original meaning, which was not about being fooled into perceiving legible text, but fooled into perceiving understanding.

▲bigstrat2003 2 hours ago | parent | prev [-]

> It's safe to say that argument has collapsed utterly. Modern LLMs virtually never collapse into endless babbling loops. Far more importantly, they solve objectively hard problems.

And yet, they still output gibberish. They also still fail to solve objectively easy problems on a regular basis. That is because they are, indeed, still stochastic parrots without any actual understanding of anything or ability to reason about stuff. The argument hasn't collapsed at all.

▲hbcdbff an hour ago | parent | next [-]

> They also still fail to solve objectively easy problems on a regular basis

Which?

▲holoduke an hour ago | parent | prev | next [-]

What is exactly "actual understanding" . If it tells 1+1 = 2 it can perfectly explain why the result is 2. Or do you mean the actual transformer work during the 1+1 question?

▲CamperBob2 an hour ago | parent | prev [-]

How'd your parrot do at IMO this year?

LLMs are certainly stochastic, but they no longer meet any sane person's definition of "parrot." I can't imagine Gebru disagreeing (has anyone bothered to ask her?)