teaming
4 stories and discussions about teaming, aggregated from every source we track.
1.
2.
Open-source adversarial testing engine, SDK, and CLI for AI agents. Runs locally or against the Humanbound Platform. - humanbound/humanbound
3.
Perplexity's security team tested its SPACE sandbox platform by giving frontier models like Opus 5 and Gemini 3.1 Pro full root access inside Firecracker microVMs. None breached the VM boundary in 108 runs, though four…
4.
LLM red teaming tests what a model says. Agent red teaming tests what it does: tools, permissions, workflows, and actions. Learn the differences and how to test AI agents.