AI safety benchmarking — Open-source test suite for multi-objective, multi-agent scenarios
Top Code Contributors: Roland Pihlakas (94.6%), Andre Kochanke (2.7%), Joel Pyykkö (1.7%), Gunnar Zarncke (0.6%)
We’re publishing a benchmarking test suite for AI safety and alignment, with a focus on multi-objective, multi-agent, cooperative scenarios. The environments are gridworlds that chain together to produce a score on biologically and economically aligned behavior of the agents. This platform is open-sourced and accessible, with support for PettingZoo.
We hope to facilitate further discussion on evaluation and testing for agents with this.


