Home/resources/Singapore AI Safety Red Teaming Challenge 2026 Report
An outcome of the 2026 International Scientific Exchange on AI Safety event, the 2026 Singapore Consensus builds on the 2025 report to present an updated global understanding on scientific research problems that are of top priority to advance AI safety research. The report outlines key developments in the AI risk landscape over the past year, such as the growth of misuse and incidents. Recognising the increasing deployment and capabilities of autonomous AI agents, the 2026 Singapore Consensus also provides a deep dive into risk management for the development and deployment of autonomous agents, which is detailed in the Companion Report on Agentic Risk Management.
AI agents are increasingly granted access to sensitive information and external tools to perform tasks on behalf of users, increasing data leakage risks. This report presents a joint evaluation by the SG and KR AISI examining whether agents can complete routine workflows without leaking data. Across 12 realistic tasks spanning areas such as enterprise productivity and customer service, the evaluation finds that agents may successfully complete tasks while still exhibiting data-handling failures, suggesting that task correctness alone is insufficient to assess agent safety. This report is a more detailed follow-up to an earlier blog post on this topic.
Korea and Singapore AISIs jointly tested how AI agents behave in real-world tasks. The exercise examined whether agents can complete multi-step tasks in common settings such as customer service and enterprise productivity without leaking sensitive data. This blogpost shares the key findings, evaluation challenges, and methodological learnings.