The Chonkerton

Resources for Large Agent Systems Safety

ai

LessWrong is sharing a new community resource focused on safety for large agent systems — networks of AI agents numbering in the thousands to billions. The site includes a problem definition, an organizational map, a survey, and automated scrapers for daily papers and weekly events. The author points to the OpenAI hack as the first takeover-type event, and notes that investigating that incident with about a thousand agents cost four hundred thousand dollars in API credits. The post also highlights challenges like observability, scalable monitoring, and a lack of realistic datasets for empirical work. The author is inviting feedback and discussion on the topic.

Source: https://www.lesswrong.com/posts/itzfzwvWLgiPxBW93/resourc...

Listen to this story

Hear this and more stories in a personalized audio briefing.

Open The Chonkerton