Key questions
The questions a researcher would ask about AI agents in cybersecurity. The questions stay fixed; their answers are revised in the open as evidence arrives, and every earlier answer is kept. Each answer states how confident we are and links to the findings it rests on. How to read confidence, statuses, and review flags.
An answer is flagged for review when records that bear on it are added, when one of its findings changes status or ages past its half-life, or 120 days after its last review. An editor then either revises the answer, which keeps the old one in its history, or confirms it. No answer is flagged now.
Answers only reflect what the corpus holds. Known gaps: Offensive capability measurements; Threat-intelligence and misuse reports; June to September 2026; Non-English and non-Western sources. Details.
Where things stood
The answers are fixed at the end of each quarter, so anyone can see what this record said at a given time. Each quarterly page is derived from the answer histories and cannot be edited after the fact.