Test whether AI agents retain critical instructions as conversations grow.
Verification confirms publisher identity (repo ownership), not code safety. The security scan covers known CVEs and suspicious install scripts — it cannot prove the absence of malicious code.
Test whether AI agents retain critical instructions as conversations grow.