Recent AI evaluation incidents expose gaps in containment, configuration and evidence
Read the recent AI evaluation disclosures involving four labs as one repeated failure and you get it wrong. A novel exploit, a misconfiguration and a leaky allowlist are different problems,…
























