Remix.run Logo
calmoo 21 hours ago

I’ve never actually seen someone give an example of what they say pangram get wrong, so please go ahead and share.

jedberg 19 hours ago | parent | next [-]

You can read it here: https://gist.github.com/jedberg/e24124e577c0da1c8445c22798eb...

I posted here on HN but it got removed.

calmoo 16 hours ago | parent [-]

I'll take your word that it's human written, but reading that gist, it absolutely reads like Claudeslop / GPT slop, it doesn't surprise me in the slightest that Pangram flagged it as AI when it reads identically to LLM output - this feels like a very acceptable edge case to me (assuming you are telling the truth).

Are you absolutely sure you wrote this by hand? If so it's kind of remarkable how close to an LLM you write like.

jedberg 15 hours ago | parent [-]

It’s important to remember that LLMs were trained on well written human text. People who write well are going to sound like an LLM. Especially if it’s a marketing message for a website.

I’ve been accused of being an LLM multiple times here on HN too. I know you have no way to know for sure other than trusting that I’m not using an LLM to write. But it’s pretty frustrating that people jump right to LLM accusations.

calmoo 14 hours ago | parent | next [-]

I’ll be honest, i’ve never read text that is human written that reads as close to an LLM as your sample sounds.

I think discounting Pangram’s accuracy based on that sample isn’t very reasonable. Really nobody writes like that other than LLMs!

jedberg 12 hours ago | parent [-]

Here is another example: https://www.reddit.com/r/ExperiencedDevs/comments/1pyjkuf/i_...

Today Pangram says it is 100% human, which is correct. But yet I got multiple DMs when I posted it 8 months ago saying "stop posting AI slop!" in response to that comment. At the time, Pangram marked it as 50% AI.

So to Pangram's credit, they got better.

calmoo 6 hours ago | parent [-]

That comment doesn’t read like slop in the slightest to me, so the people DMing you have a bad eye for it. Regardless, I think your ‘human’ sample is not a good indicator of the quality of pangram. I would update your priors a bit.

ekelsen 11 hours ago | parent | prev [-]

Do you have any published dateable text from pre 2023? I'm super curious if you always wrote like this, because it is exactly in the style of AI slop.

People say it's in the training data, but I haven't seen any good examples of clearly dated pre 2023 text that sounds like this.

jedberg 11 hours ago | parent [-]

I have 1000s of reddit and hacker news comments from before 2023, and lots of long form writing too. But as you point out, those are all in the training set, and in pangram's "definitely human" training set too.

The ones that sound like LLMs tend to be the ones that were well researched and spent more time on, not the off the cuff stuff, which is most of what I write. So it would take me a while to find something like that.

But you're welcome to dive into my reddit and HN history, or all my blog posts on the wayback machine if you want to look for one. :)

21 hours ago | parent | prev [-]
[deleted]