Funding better evaluations of AI's impact on wellbeing
anthropic.com
1 thread
> In brief, we’re seeking evaluations that:
> ...
>* Test both precautions and harms (i.e., evaluate the risk of both overcompliance and overrefusal);
So, no overuse, overreliance etc.
Why are we not surprised.