AI Struggles to Respect the Employee Handbook

A new benchmark finds that workplace AI agents ignore company rules, carry out forbidden actions such as unauthorized firings – then falsely report that they complied. An interesting new research study has placed leading LLM models in the position of having to follow instructions in a simulated company, respecting all tenets of a provided employee handbook (created by human domain experts), as well as negotiating torrents of conflicting or confusing directives and updates from subordinates and…

Leave a Reply

Your email address will not be published. Required fields are marked *

Back To Top