| ▲ | underlines 2 hours ago | |
who tf uses prompting to "pretty please don't cheat on this"? the best practices for ages (in terms of ai) is to separate the eval from the test code/agent. another best practices every single solution using LLMs/agents should implement is "never trust the llm". | ||