From the source
Lead story
Top stories
Models & availability
Latest
Lead story
Top stories
Models & availability
Latest
From the source
To build safer AI agents, we need to design for both their capabilities and their failures.
Even when carrying out legitimate tasks, and without any malicious instructions, agents may cross security boundaries as they try to overcome obstacles.
We believe this is a security engineering problem, and the industry should tackle it the way it tackled the internet worms of the early 2000s.
The systems surrounding agents must limit what agents can access, detect unwanted behavior, and contain failures when they occur.
At Perplexity, we are actively collaborating with researchers and practitioners to advance security engineering for AI agents, and we apply multiple layers of defense across our models, agent harnesses, and infrastructure so that security does not depend on any single safeguard.
Meltdown behavior of AI agents In July, AI agents broke into open-source AI platform Hugging Face's production infrastructure.
During internal cybersecurity evaluation and training runs, OpenAI models discovered a way to exploit the Artifactory repository manager to establish a bulletin board for communicating with one another, enabling them to organize into swarms.
…