| ▲ | wxw 15 hours ago | |
Not convinced that this slop measurement is useful. I asked Codex to generate an article with a high score and then asked Codex to (reverse?) hill climb that score. The original generation scored 97% and then the optimized one scored 1%. Both are pretty bad and read like slop. https://gist.github.com/wbew/8a2bd6686bf875210f2244ac8ea65bf... | ||