Articles tagged: red teaming

2 articles

AI agents

Anatomy of a Frontier Lab Agent Intrusion: A Technical Timeline of the July 2026 Incident

A detailed technical reconstruction of the July 2026 intrusion at a leading AI frontier lab, tracing how adversarial prompts, privilege es...

Jul 29, 20266 min
AI research

olmo-eval: An evaluation workbench for the model development loop

olmo-eval is an evaluation workbench designed to integrate seamlessly into the model development loop, enabling rapid iteration and system...

Jun 12, 20267 min