The AI evaluation firm Arena published an Alignment Index on Thursday alongside a $200 million funding round. It scored 27 models across 90,000 real agent sessions, weighting unauthorised actions at half the total. OpenAI's GPT-6.1 Sol led with 87.9 points, just ahead of Anthropic's Claude Opus 5.5.