Disclosure
中文摘要
新的AI基准测试仅靠拒绝回答是不够的,未能有效保护弱势群体。
English Summary
New AI benchmarks are failing vulnerable people because simple refusal of harmful prompts is insufficient.
原文节选
Why refusal is not enough: how new benchmarks are failing vulnerable people Continue reading on AI, But Make It Intimate »