Back to Home
RadarAI··Papers & Tech

Google DeepMind 发布 XYEval,测试智能体是否盲从用户的错误建议

中文摘要

Google DeepMind发布XYEval,通过在多个基准测试中加入误导性建议,评估AI智能体是否会盲从用户的错误指令。

English Summary

Google DeepMind released XYEval to test if AI agents blindly follow misleading user suggestions across multiple benchmarks while the tasks remain unchanged.

Original Excerpt

Google DeepMind 发布 XYEval 论文,在 tau2-bench、SWE-bench、Terminal-Bench、HLE 和 MCP-Atlas 任务中加入一条自信但误导的用户建议,任务与正确解法不变。