📣 Try out the NEW Gateway to Research and let us know what you think.

We're looking for users to test the new service during August and September and share their feedback. Express your interest by completing this short form.

Supervisory Agent for Safety Improvements in AI Systems

Lead Research Organisation: UNIVERSITY OF NOTTINGHAM
Department Name: School of Computer Science

Abstract

This research will investigate the feasibility of introducing an independent Supervisory Agent to monitor AI based agent(s) operating within an environment. The primary focus is agent(s) whose behaviour has the potential to change at run-time (for example, via techniques such as generative planning or machine learning) thus introducing unknown actions. The Supervisory Agent will monitor the intended actions of the agent(s), and prevent the execution of any harmful or unsafe actions. This may involve substituting a safe action for an unsafe action that has been suppressed.

People

ORCID iD

Publications

10 25 50