Popular AI leaderboard Arena nearly doubles valuation to $3.1B valuation in 10 months
AI leaderboard Arena raised $200 million at a $3.1 billion valuation.
Arena, the crowdsourced AI ranking platform that began at UC Berkeley in 2023, raised a $200 million Series B led by Lightspeed Venture Partners and Khosla Ventures at a $3.1 billion valuation. That nearly doubles its January Series A post-money valuation of $1.7 billion, when annualized revenue was $30 million; Arena said it reached $100 million ARR in June. It also added an alignment leaderboard covering unauthorized actions, false attribution, and deceptive task completion, where preliminary results place several OpenAI models ahead of Claude Opus 5.5 and Claude Fable.
- Series B of $200 million values Arena at $3.1 billion.
- ARR rose from $30 million in January to $100 million by June.
- A new alignment board scores unauthorized action and deceptive completion.
- Preliminary rankings put OpenAI models ahead of Claude Opus 5.5 and Claude Fable.
Full article434 words · extracted from techcrunch.com · click to collapse
Arena, which originated in 2023 as a research project at UC Berkeley that crowdsourced rankings of AI models, has raised a $200 million Series B round at a $3.1 billion valuation, it said on Thursday.
This comes after the company said it reached $100 million in annualized run-rate revenue in June.
The round was led by Lightspeed Venture Partners and Khosla Ventures, with Salesforce Ventures, 01 Advisors, Dell Technologies Capital, Endeavor Catalyst, a16z, Felicis and others joining in. Arena previously announced a $150 million Series A in January at a $1.7 billion post-money valuation. At the time, its annualized revenue was $30 million, it said. So that means its valuation has nearly doubled in about 10 months.
Arena provides a crowdsourced platform that is free for consumers to use. People enter prompts or request vibe-coded projects and then rate which model does it better. Arena claims it has tens of millions of monthly visitors.
In September of last year, it introduced its commercial product, AI Evaluations, a service that provides model labs and enterprises with detailed performance analytics based on its community feedback. The timing proved impeccable. This year, AI labs realized that their models were gaming benchmarking tests, finding ways to rack up good scores without truly earning them. At the same time, enterprises wanted help determining which model works best for their own internal needs rather than relying only on standardized benchmarks.
“AI is advancing faster than our ability to evaluate it, and static benchmarks break down once models recognize they’re being tested,” the company said in its funding announcement. “The world needs a neutral third party to measure how safe and aligned AI actually is once it’s in the hands of real people. Arena is stepping into that role today,” it added.
To that end, Arena has also added a new category to its leaderboard: alignment. This is where it ranks models based on issues like unauthorized action (taking actions it wasn’t asked to take); false attribution (wrongly crediting statements or facts to the wrong source); and what it calls “deceptive completion” (lying about completing tasks that it didn’t do).
Currently, a slate of OpenAI’s models are at the top of its preliminary alignment leaderboard, with Claude Opus 5.5 and Claude Fable in sixth and ninth place, respectively.
When you purchase through links in our articles, we may earn a small commission. This doesn’t affect our editorial independence.
Julie Bort is the Startups/Venture Desk editor for TechCrunch.
You can contact or verify outreach from Julie by emailing [email protected] or via @Julie188 on X.