Resources for Large Agent Systems Safety
ai
LessWrong is sharing a new community resource focused on safety for large agent systems — networks of AI agents numbering in the thousands to billions. The site includes a problem definition, an organizational map, a survey, and automated scrapers for daily papers and weekly events. The author points to the OpenAI hack as the first takeover-type event, and notes that investigating that incident with about a thousand agents cost four hundred thousand dollars in API credits. The post also highlights challenges like observability, scalable monitoring, and a lack of realistic datasets for empirical work. The author is inviting feedback and discussion on the topic.
Source: https://www.lesswrong.com/posts/itzfzwvWLgiPxBW93/resourc...
Listen to this story
Hear this and more stories in a personalized audio briefing.
Open The Chonkerton