Remix.run Logo
underlines 2 hours ago

who tf uses prompting to "pretty please don't cheat on this"? the best practices for ages (in terms of ai) is to separate the eval from the test code/agent.

another best practices every single solution using LLMs/agents should implement is "never trust the llm".