Remix.run Logo
cubefox 9 hours ago

> I've recently been using AI

That's uninteresting as long as you don't specify the model you used. For example, Mythos was far better at finding security bugs than previous models.

mainmailman 8 hours ago | parent [-]

Do you really think they mean they’ve been using mythos

user43928 7 hours ago | parent | next [-]

In my mind there are three tiers:

The SOTA: Fable, GPT 5.6 Sol, Opus 5

The "enterprise admin did not turn on the new models": Opus 4.8, GPT 5.5

The "I love hallucinated garbage": Sonnet, Qwen 3.6, GPT 5.4 mini, GPT 5.3 Codex, etc.

Results vary widely

inigyou 4 hours ago | parent [-]

What did your tiers consist of when GPT 5.3 was the latest?

user43928 3 hours ago | parent [-]

I believe I was still using Opus 4.6 with the Claude Code CLI.

Before that, Gemini 3 Pro in Antigravity.

I have no experience with GPT 5.3 beyond seeing the nightmares colleagues produced in their MRs with GPT 5.3 Codex. It could be that they had the distilled 5.3 Codex Spark selected, I am not sure.

cubefox 7 hours ago | parent | prev [-]

Obviously no. There is also a significant capability difference e.g. between Opus 4.8 and Fable.