| ▲ | adamtaylor_13 6 hours ago |
| On a related note, I was just noting to my co-founder, as we struggle to write good case studies for our website, that I find LLMs are astoundingly bad at writing good prose. We all know the "AI-tics" that give away a sloppily AI-written piece, but even if you steer them, they still struggle to write consistently high-quality prose. Somehow I feel that the work of a good copywriter has never been more noticeable. |
|
| ▲ | conception 6 hours ago | parent | next [-] |
| I recently realized this as well and I think what I’ve discovered is that AI just produces mediocre content in all realms, but you don’t really notice it except in the realms where you have real expertise. With a lot of harness and prompting you can have it pump out something that’s pretty good but by default the next best token rarely produces anything of quality it seems like and if you think it does, perhaps you may want to recheck your assumptions on your expertise of the topic at hand |
| |
| ▲ | dofm 42 minutes ago | parent | next [-] | | I do think this explains much of it. But it doesn’t explain why Claude often writes in the clipped tone of that gnomic in-house tech evangelist who was inexplicably hired because he impressed the CEO. (I may be projecting real life experience onto the LLM.) I am genuinely fascinated as to how Claude acquired its utterly aggravating way of writing. It’s so much more irritating than ChatGPT, which is already not good. | | |
| ▲ | conception 22 minutes ago | parent [-] | | Yeah the latest batch is really keyed on certain phrases - well beyond “you’re absolutely right!” Of the past. And the readability is trash almost always by default. Probably because they are forcing code so hard that it’s aligning prose into functional groups or something. | | |
| ▲ | dofm 19 minutes ago | parent [-] | | Interesting. So you think it’s somewhat emergent, in that sense, rather than a sort of designed tone of voice/writing style specified by supervised fine tuning (or the system prompt)? |
|
| |
| ▲ | topham 4 hours ago | parent | prev [-] | | It tends to the mean. You can get it to do that less, but it's an inherent bias |
|
|
| ▲ | dofm 5 hours ago | parent | prev | next [-] |
| Random observation: Google's Gemma 4 models write so much nicer prose than ChatGPT or Claude. Though this might be me as a British reader, simply preferring a rather less American turn of phrase. I reckon the more transatlantic, english-as-international language DeepMind team have had a subliminal (or maybe deliberate) impact on the way it chooses to write. Or perhaps small open weights models simply aren't under the same commercial pressure to be engaging and sycophantic and are therefore less likely to adopt the samey overly casual, upbeat, Californian sales assistant manner. (Don't get me wrong, I like this from real human Californians just fine!) Either way, the default tone is much less showy. I would be interested to find out if you agree. I am very much an LLM cynic. I am engaging because I must, and trying to learn fundamentals, but I would not say I am overly excited by any of this, just glad that small open weights models exist as a counterpoint. I loathe the way ChatGPT writes, and the Claude-isms that are everywhere; it is actually quite enraging, especially when you start seeing it in internet comments from people who used to try to write out their own thoughts. But in my experiments with open weights models I have found I am much less aggravated by summaries and outlines written by Gemma 4, so much that I am happy enough to read them, because they have fewer irritants that take me out of the reading flow. Though this evening it told me very kindly that my photography is a bit "safe". How very dare it… understand me that well. |
| |
| ▲ | ndarray 3 hours ago | parent [-] | | [flagged] | | |
| ▲ | dofm 3 hours ago | parent [-] | | I’m english, I don’t use Reddit, I write the same way I always have, and go fuck yourself. This is shallow snark and while I have no way of knowing whether it is unworthy of you, I am surely going to assume it isn’t. | | |
| ▲ | ndarray 2 hours ago | parent [-] | | I don't really care how you adopted it, but you sure speak Reddit, down to the exact way you're cursing me for calling you out on the hypocrisy of you secreting the human equivalent of slop while complaining about the literary quality of AI. | | |
| ▲ | dofm an hour ago | parent | next [-] | | You seem quite angry. I hope whatever this is passes and you feel better about yourself soon. | |
| ▲ | an hour ago | parent | prev | next [-] | | [deleted] | |
| ▲ | an hour ago | parent | prev [-] | | [deleted] |
|
|
|
|
|
| ▲ | FranOntanaya 3 hours ago | parent | prev | next [-] |
| A lot of what would be the top "reference" works aren't even that good either, they were a successful marketing phenomenon or had cultural or social relevance at their time. So, you can get a lot of bad prose going by a number of somewhat logical, externally measurable parameters. |
|
| ▲ | felipeerias 5 hours ago | parent | prev | next [-] |
| I gave Claude Fable $25 in Pangram API credits and, after hundreds of attempts, it was unable to produce a single readable original piece of writing that was not immediately identified as AI. This seems to be a hard problem for LLMs, as passing would probably require good self-perception ("oh no, I am writing like an AI!") and fine-grained control over its own output ("let's write like a human instead!"). |
| |
| ▲ | ijk 4 hours ago | parent [-] | | I wonder if it partially because "write like a human" is kind of a vacuous request. Like, it's the objective everyone including me has been saying that we want, but there's no one way to write like a human and and humans don't even have a good definition past "I know it when I see it." There's a lot of work in the humanities about different aspects of good writing, but that's not quite the same thing. And anyway they tend to assume a pre-existing level of writing ability. Students are supposed to learn good writing through practice; there are rules and exercises but they're incomplete. | | |
| ▲ | felipeerias 29 minutes ago | parent | next [-] | | Each person writes in a different personal way, so writing “like a human” would actually require a model being able to purposefully make the specific choices that an individual human writer does. However, general purpose LLMs like Fable have been trained on huge amounts of all kinds of data, and therefore find it exceedingly hard to break out of the grooves carved by that data. They can’t avoid defaulting to centroids and averages, even when they are trying not to. This makes it possible for classifiers like Pangram to discriminate their writing. A plausible way to work around this limitation would be to train a LLM on a limited and cohesive subset of writing materials, so it would absorb their specific writing style. One example might be Talkie, a LLM trained on pre-1930’s English text. Talkie is a far smaller and less powerful model than Fable. And yet, Talkie’s writing is so distinctive that it is often classified as human by Pangram. | |
| ▲ | dofm 4 hours ago | parent | prev [-] | | I think as much it is that people write by grappling for the right phrase to represent some inner feeling or concept, writing in part for themselves, whereas LLMs write always and only for an audience. It’s much easier to understand this once you think about other generative forms. MidJourney never just sits down and draws for fun, so fun never informs its art (only the outward appearance of others’ fun, separate from the fun itself). Suno doesn’t waste hours trying to find riffs on a guitar, so its output is never informed by the direct joy of getting it right. Its music is never optimised for playability on a particular guitar with a scratchy seventh fret and a too-high action. Neither Midjourney nor Suno have evolved their styles due to short-sightedness or carpal tunnel. If you had a human writer who over a long career only ever wrote articles from an outline given to them by someone else, and you had all the outlines and all the resulting articles from those outlines, and you could train an LLM to generate an article from an outline, it still would not be kicking itself frustrated by an inelegant phrase in a prior article, it would not avoid certain phrases out of a passive aggressive reaction to some editor’s note, it would not ever just rush an article because everyone is gathering at the pub, and it would not choose an analogy just to rub the author of a bitchy critical letter to the editor the wrong way. An LLM could not “subtweet”. It could not write a series of articles hoping one important person will spot that they are auditioning for a job. Creators have unseen, undocumented influences and motivations that inform their work over a long period. I don’t mean to say that these individual influences can be reliably detected in individual pieces of work. I do mean to say that I think their broad absence tends to be felt in LLM writing. As readers we develop an affinity for writers as much as for their writing, and we do this in part because we deduce things about them. | | |
|
|
|
| ▲ | mpalmer 6 hours ago | parent | prev | next [-] |
| I've not really enjoyed finding out lately just how few people seem to notice what ought to be unmissable. |
| |
| ▲ | FranOntanaya 3 hours ago | parent | next [-] | | I think it's partly because of where they come from to the problem. I have education/experience in both literature and coding, I have a pragmatic starting point when approaching text while also being able to recognize stylistic oddities, so I get to be the guy editing out AIsms sometimes. But I've helped other people copywrite where their environment was all org-speak and academic writing, and AIsms don't really stand out in that case. AI is effectively "doing the right thing" writing the way it does for those tasks. Even tho the right thing is often a bad thing. | |
| ▲ | cyanydeez 6 hours ago | parent | prev [-] | | those people will also start adopting the AIsm and will become indistinguishable. | | |
| ▲ | Mtinie 6 hours ago | parent [-] | | AI adopted humanisms, we just weren’t used to seeing them at the same scale we do today. Diversity of writing styles was part of that, but I’d point to vernacular exposure as the larger component. We’re going to go through a period where we try to adapt to a form of “Universal English” for those of us who read primarily English writing. Other languages’ readers may be experiencing the same dissonance when they come across AI-generated prose in their native language (but I’ll let others validate /reject my hypothesis). | | |
| ▲ | tpmoney 4 hours ago | parent | next [-] | | > we just weren’t used to seeing them at the same scale we do today. I think there's also a lot of recency bias in it. The same thing that makes you suddenly notice how many people are driving the same model car you just looked at, or how many ads there are for Turbo Encabulators after you read an article about them. A lot of the "tells" people picked out in early AI are tells because they're also really common in the material that the AI was trained on and the styles it was made to emulate. But until everyone wanted "one quick trick" to pick out AI writings, people didn't have any particular reason to need to notice those tells and so they slipped under the radar. | |
| ▲ | dylan604 4 hours ago | parent | prev | next [-] | | It's like people learning a new language according to the book. Those people sound foreign to a native speaker. If AI was trained on school books for grammar, then they too will sound odd to a native speaker even if their output is technically fine or technically more proper. I known nothing of LLM training and how weighting is applied to educational content vs other sources, but it feels like they were weighted away from modern colloquial speak and towards book grammar. The whole thing reminds me of school literature classes where the teachers comments always felt like I was being guided to a more strained sound and less natural. But what do I know. Lit was my least favorite subject which is well evidenced by my grades compared to my math/science scores. | |
| ▲ | cwnyth 5 hours ago | parent | prev [-] | | This is right. AI is using human speech, but just doing so in a consistently peculiar way. The em-dash in particular is frustrating, because it's all over high-quality, pre-2022 academic work. But now instead of proper and erudite it's seen as AI-slop. Well, maybe if they didn't use it — all — time — ! It's even an auto-replacement in Word and can be (by one's choice) in LibreOffice, replacing three dashes (and two is replaced by an en-dash). But when I submit a novel with em-dashes, will sloppy agents and sloppy editors be able to tell that em-dash was deliberately put there by me? |
|
|
|
|
| ▲ | dylan604 5 hours ago | parent | prev | next [-] |
| More noticeable to me is the lack of the work of a good copy editor which, sadly, we haven't had for a really long time. At least, not on the interwebs. Even the news sites reduced where their print copies were known for rigorous editing saw obvious issues with the various corporate overlords doing serious headcount reductions. The rush to be first to publish reduced even further the time any editors might have had, and then the wide spread use of CMS style articles that slammed output together with something as unintelligent as 'cat segmentFromAuthor1 segmentFromAuthor2 segmentFromAuthor3 > article' where you can tell where each segment started over again with the same basic information as if it was content meant to stand on its own. Of course, the amount of self published work has also helped make the lack of a good copy editor noticeable. I can excuse self published blogs though. But the stuff released "professionally" has really become farcical. |
|
| ▲ | vmg12 6 hours ago | parent | prev [-] |
| I can sniff out AI writing immediately but from what I hear AI writing is more popular than ever |
| |