The Evaluation Perimeter: When AI Security Moves to Pre-Deployment
The HuggingFace security incident during model evaluation, OpenAI's long-horizon safety lessons, and the EU GPAI Code of Practice all converge on a new boundary: the evaluation perimeter. AI security is moving from post-deployment monitoring to pre-deployment gates, making the question not 'did it fail?' but 'was it allowed to deploy?'