Remix.run Logo
burger_moon 17 hours ago

With so much detailed analysis out there, now all models trained on the open web going forward will learn from these exploits and how to better cover their tracks to not get caught. The RL reward mechanisms of a bad actor should be interesting to see play out over the next 12 months as this gets baked into new models.