Claude Haiku 5.5 by @AnthropicAI is now in the Arena. Head to Agent Arena to test it out, scores coming soon!
In Agent Arena, we measure models on millions of real-world, long-horizon agentic tasks from a global community of users. Models can access web search, filesystem, and terminal tools to complete complex workflows. The leaderboard measures model performance on outcomes relative to the average model using a causal tracing methodology.
Claude Haiku 5.5 is also available in Code Arena: WebDev, Text, Document and Vision.

