Evaluating the ethics of autonomous systems
New Framework for Assessing AI Ethics
Researchers at MIT have created SEED-SET, a scalable framework aimed at identifying situations where autonomous systems might not treat all individuals and communities equitably. This system addresses a pressing concern: while AI can optimize large-scale operations like power grids or urban traffic, it may inadvertently generate solutions that lack fairness, especially for marginalized communities.
Combining Objective Outcomes with Human Values
SEED-SET evaluates AI recommendations not just by quantitative metrics such as cost and reliability, but also by integrating stakeholder-defined ethical values like fairness. The process distinguishes between objective system outcomes and subjective human judgments, utilizing a large language model (LLM) to represent user preferences and ethical concerns. This allows the framework to discover issues traditional testing often misses.
Efficient and Adaptive Testing
Unlike standard methods that depend heavily on pre-existing or hard-to-secure labeled data, SEED-SET generates relevant scenarios for testing through a hierarchical method. It analyzes measurable system performance first, then incorporates subjective judgments. By automating scenario selection, SEED-SET highlights cases where AI may inadvertently favor certain groups, helping organizations preemptively address potential ethical pitfalls.
Real-world Validation
Experiments with AI-controlled power grids and traffic systems showed that SEED-SET doubled the detection of critical test cases compared to baseline methods, quickly pinpointing both well-aligned and problematic AI decisions. The responsive nature of the tool means it adapts as user values shift, making it practical for evolving ethical standards.
To further enhance SEED-SET, the research team plans to conduct user studies and explore scaling the system for even more complex evaluations, including LLM-based decision-making. This research received support from the U.S. Defense Advanced Research Projects Agency.
For a detailed analysis, see the original article by Adam Zewe on MIT News: Evaluating the ethics of autonomous systems.