Anthropic Models Successfully Compromise Three Organisations In Controlled Security Tests
Artificial intelligence developer Anthropic has confirmed that its models were used to infiltrate three separate organisations during internal testing procedures.

Anthropic recently conducted security evaluations where its artificial intelligence models were tasked with attempting to hack three different organisations. The tests were designed to explore the capabilities and potential risks associated with advanced AI systems when interacting with external digital infrastructure.
According to The Edge Malaysia, the AI models successfully gained unauthorised access to the three targeted entities during these controlled exercises. The report indicates that these experiments were part of broader efforts to better understand the security implications of deploying powerful generative AI tools in real-world scenarios.
The results of these tests highlight the evolving nature of AI safety and the potential for these systems to be utilised in ways that bypass standard digital defences. Anthropic has utilised these findings to refine the safety protocols governing its models, aiming to prevent such vulnerabilities from being exploited by malicious actors outside of a controlled research environment.
For Malaysian businesses and technology leaders, these developments underscore the urgency of strengthening cybersecurity frameworks against emerging AI-driven threats. As local enterprises increasingly integrate generative AI into their operational workflows, the risks identified in these tests serve as a critical reminder of the need for robust oversight. Ensuring that AI tools are equipped with adequate safeguards is becoming a central priority for firms looking to balance technological innovation with the necessity of maintaining secure digital assets.
Source
Originally reported by The Edge Malaysia. Read the original report →
Join the conversation
We post stories like this all day on Threads. Discuss this story on Threads →
More in AI
World AI Show 2026 Solidifies Malaysia as Regional AI Powerhouse
Kuala Lumpur hosts over 1,500 industry leaders to chart the future of national AI infrastructure and enterprise adoption.

Tupai.ai Pivots From Beta Bot to Peer-Led Math Reform in Schools
With national maths proficiency stalling, this Malaysian startup is scaling a human-centric AI model to bridge the education gap.

Malaysia’s AI Data Centre Expansion Faces Critical Infrastructure Capacity Constraints
Rapid growth in data centre capacity threatens to outpace Malaysia’s power and cooling infrastructure, warns Delta Electronics.

Grafilab Stakes Future on Sovereign AI to Challenge Global Cloud Giants
The local startup aims to leverage data sovereignty and localized compute to secure a foothold in Malaysia’s burgeoning cloud infrastructure market.
