We’re putting too much faith in AI’s ability to say no
Ask a chatbot something dangerous and it will refuse the request. But this critical tool of AI safety is far from foolproof and could become an instrument of repression.
MIT Technology Review
Topics: Policy
Entities: Policy