1 item
Arena, which began as a UC Berkeley research project, has raised $200 million in a Series B round at a $3.1 billion valuation. The company sells AI evaluation products that use real human feedback rather than fixed tests. It has also launched an alignment leaderboard that scores models on behaviors including unauthorized actions and deceptive task completion. Both products address a recognized gap in enterprise model risk programs: static benchmarks can be manipulated by model developers.