| ▲ | krupan 5 hours ago | ||||||||||||||||
"Unfortunately, given participant feedback and surveys, we believe that the data from our new experiment gives us an unreliable signal of the current productivity effect of AI tools." That's more than a caveat. That says, this study is unreliable. And why? "The primary reason is that we have observed a significant increase in developers choosing not to participate in the study because they do not wish to work without AI." So, maybe people are fooling themselves, just like this article suggests (I almost wrote asserts, but saying, "I think..." isn't really an assertion) | |||||||||||||||||
| ▲ | kalkin 4 hours ago | parent | next [-] | ||||||||||||||||
The people who chose to use AI would be the ones who would be fooling themselves, right? But they measured actually positive productivity there. METR's concern makes most sense if it's about the inability to measure a consistent trend over time, not about the validity of the study for the population that participated. | |||||||||||||||||
| ▲ | dcow 5 hours ago | parent | prev | next [-] | ||||||||||||||||
> The primary reason is that we have observed a significant increase in developers choosing not to participate in the study because they do not wish to work without AI. Why do you think this is? Who wants to spend an afternoon gluing API calls together when the robot can do it in 5 minutes? | |||||||||||||||||
| ▲ | leecommamichael 5 hours ago | parent | prev [-] | ||||||||||||||||
Don't let the downvotes get you down. They don't know who funds or constitutes METR. If you actually read these papers, you'll find they cite themselves 4-5 times each publication. They also consider "hosting an HTTP server using Python" to be a "long time horizon" task. | |||||||||||||||||
| |||||||||||||||||