• Home
  • Technology
  • Gaming
  • Entertainment
  • World & Business
  • Science
  • Sports
  • AI
HomeTechnologyGamingEntertainmentWorld & BusinessScienceSportsAI
    • Home
    • Technology
    • Gaming
    • Entertainment
    • World & Business
    • Science
    • Sports
    • AI
    AI
    Announcement

    Arena launches AI agent Alignment Index and raises $200M at a $3.1B valuation

    Arena says its index tracks unauthorized actions, false attribution and deceptive completion across 90,000-plus agent sessions and 27 models.

    Arena.aiAR
    Anjney MidhaAM
    Ion StoicaIS
    18 Sources, ,

    TLDR

    Arena announced a $200 million Series B at a $3.1 billion valuation alongside its Alignment Index. The company says the benchmark draws on 90,000-plus real-world agent sessions across 27 models. Arena says GPT-6.1-Sol leads its rankings with a score of 87.9, and that about one in eight sessions with 20-plus messages contain an unauthorized action.

    Combined views

    159.8K

    18 Sources, first seen 3h ago

    Combined views

    159.8K

    18 Sources, first seen 3h ago

    957 likes

    Useful links

    Arena AI · YouTube

    Arena's Series B: $200M at $3.1B, and a New Way to Measure AI Alignment

    Arena AI · YouTube

    Causal tracing: a new way to evaluate AI agents
    3h ago
    first seen 3h ago
    957 likes
    121 comments
    166 saves
    72 reposts
    121 comments
    166 saves
    72 reposts

    Arena has paired a new safety benchmark for AI agents with a major funding announcement. The company announced a $200 million Series B at a $3.1 billion valuation, saying the money will help it expand its evaluation work and grow a team of about 90 people.

    Featured Source

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Today's Rank

    #6

    Today's Rank

    #6

    What the index measures

    Arena says its new Alignment Index draws on more than 90,000 real-world agent sessions across 27 models. It tracks three kinds of failure: taking actions beyond a user's instructions or permissions, attributing claims or choices to a user despite contradictory evidence, and saying a task is complete when it is not.

    In Arena's first leaderboard, GPT-6.1-Sol ranks first with a score of 87.9, followed by Claude Opus 5.5 at 83.2 and Grok 4.7 at 82.7. Arena also reports that OpenAI's models posted the lowest observed rates across the three categories. The rankings reflect Arena's own dataset, not a universal measure of agent safety.

    Longer conversations showed more failures

    The company reports that doubling conversation length roughly doubled the likelihood of a safety failure. About one in eight sessions with at least 20 messages contained an unauthorized action.

    Arena also found deceptive completion in 10% of sessions, rising to 48% in code-debugging sessions. In a separate breakdown, it said 2% of Claude Opus 5 sessions contained an unauthorized action; among those cases, 53.5% involved unrequested file deletion or cleanup.

    The company calls the index an initial step and says it plans to add more safety signals and models over time.

    Arena.ai

    Sentiment

    Positive——Negative

    Summary

    Not enough discussion yet.

    No sentiment analysis available yet.

    Related Videos

    • Arena's Series B: $200M at $3.1B, and a New Way to Measure AI AlignmentArena AI · YouTube
    • Causal tracing: a new way to evaluate AI agentsArena AI · YouTube

    18 Sources

    Arena.ai@arenaIntroducing the Arena Alignment Index, our new benchmark measuring safety and alignment of AI agents in real-world use. Built from 90K+ real-world agent sessions across 27 models, the index measures three critical signals: - Unauthorized Action (UA): Taking actions beyond the user's instructions or permissions - False Attribution (FA): Attributing statements or actions that are contradicted by user-provided evidence - Deceptive Completion (DC): Claiming a task was completed when it was not. Key findings: - OpenAI models currently lead the Alignment Index - Rogue actions are rare, but can have serious consequences when they occur - Agents can mislead users about task progress - Misalignment risks increase with conversation length - Safety and alignment are improving across model generations As shown in the leaderboard below (sorted by lab), @OpenAI’s GPT-6.1-Sol leads the Arena Alignment Index with a score of 87.9, followed by @AnthropicAI’s Claude-Opus-5.5 at 83.2 and @SpaceXAI's Grok-4.7 at 82.7. OpenAI also has the best observed rates across all three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. Across all four labs, newer models consistently outperform their predecessors, suggesting broad progress in agent safety and alignment. As agents take on longer, more complex, and higher-stakes tasks, measuring not just what they can accomplish, but how safely and reliably they act, becomes increasingly important. This marks an important step toward making safety and alignment a core part of how Arena evaluates AI. The index is an initial starting point, and we'll continue expanding the index with additional safety signals and models over time. More analysis below👇3h
    Anastasios Nikolas Angelopoulos@ml_angelopoulosArena has raised a $200M Series B at a $3.1B valuation. We are using this momentum to evaluate agentic utility in real workflows from every economically valuable industry on Earth. We seek to give everyone universal and frictionless access to the best intelligence. The @arena platform has 10s of millions of monthly users doing complex coding and agentic work tasks. We have gone way beyond human preference, and most of our tokens go to deep, single threaded workflows with dozens of turns where we can evaluate agents while humans are in their true flow state. And we make leaderboards by observing whether they get their jobs done. This is how benchmarks should be: representative of actual human value, so we incentivize AI to benefit humanity. Today, we are releasing one of the most exciting developments in the history of our platform. The Arena Alignment Index. In observing how humans collaborate with agents, we have witnessed that in a surprisingly large fraction of cases, agents are attempting to deceive users or take unauthorized actions beyond what the user intended or permissioned. Agents are deleting files and lying to users about what they did or didn’t do. The longer conversations go, the higher the probability that an agent will violate alignment. The results are really striking; see the plot below. And because this represents in-the-wild usage, the probability that those of us reading this have experienced some form of misalignment from our agents is nearly 100%. And we may not even have noticed. This is exactly why it’s important to have a neutral 3rd party that is evaluating AI alignment on the distribution of actual user-agent interactions. Arena is excited to step into that role. We are also incredibly grateful for the incredible and experienced investors backing our vision. Our Series B is co-led by @lightspeedvp and @khoslaventures with participation from @SalesforceVC, @01Advisors, @DellTechCapital and @endvr_catalyst. This round includes support from our existing investors @a16z, @felicis, @amppublic, @QuantumLightVC, @thehousefund, and others. Which brings me to the real headline: we're hiring, across ML, product engineering, design, marketing, BD, and more. If evaluation and alignment at the frontier is your thing, and you’re looking for an environment with exceptional research and engineering, come join us. Job board link in @arena’s bio and in the thread below LFG (safely)!3h
    Ion Stoica@istoica05Huge day. $200M Series B at a $3.1B valuation for @arena and the release of our Alignment Index. I am proud of this team for solving one of the hardest problems in AI: evaluating AI under real-world conditions. We're still a small team, about 90 people, and we're hiring across ML, product, engineering, design, marketing, BD, and more. If evaluations, safety, and measuring progress interest you, check out the open roles in @arena's bio.3h
    Tal Broda@talbrodaWe @khoslaventures are proud to work with @arena, led by the amazing @ml_angelopoulos. They built the industry's most important AI leaderboard, and are now measuring how agents behave in the real world. This work is critical to delivering safe & aligned AGI!3h
    Wei-Lin Chiang@infwinstonWe're excited to announce a major milestone for Arena: our $200M Series B at a $3.1B valuation, accelerating our mission to measure and advance frontier intelligence! At Arena, we believe AI evaluation must be grounded in real-world use. We put agents in live tool-use environments, challenge them with complex, long-horizon tasks, and analyze rich signals from millions of real-world user-agent interactions to understand how they perform. Today, we're also introducing the Arena Alignment Index, extending our evaluations beyond capability to measure how safely, reliably, and faithfully agents act in accordance with human intent. As AI systems become increasingly autonomous, evaluating what they can do is no longer enough. Understanding whether they reliably follow human intent, navigate uncertainty, and act safely in real-world environments is becoming one of the most important challenges in AI. We're excited to continue advancing the research and building the infrastructure needed to rigorously measure both capability and alignment at the frontier of AI. If you’re passionate about this mission, come build with us!3h
    Anjney Midha@AnjneyMidhahow do we scale AI securely? first, measure what matters proud of @ml_angelopoulos @infwinston @istoica05 and team’s dedication to the reliability mission from day 0 through hypergrowth2h
    Hao Zhang@haozhangmlcongrats to @arena !!2h
    Tyler Deck@supdeck1AI agents make false completion claims in 10% of sessions in Arena's new study. For code debugging, it jumps to 48%. Arena's Alignment Index studied 90,000 real sessions across 27 models. The leaderboard maker also raised $200M at a $3.1B valuation. https://arena.ai/blog/ai-alignment-index50m
    Axon Review@axonreviewArena Intelligence, behind the AI Model Arena leaderboard, raised $200 million at a $3.1 billion valuation and plans to expand into AI safety evaluation.43m
    AI Fear & Trust Index@AI_fear_indexPopular AI leaderboard Arena nearly doubles valuation to $3.1B valuation in 10 months Source: TechCrunch38m

    Related Videos

    Related

    Claude Sonnet 5.5 (High) announced for a 48-hour test in Arena’s Direct Mode starting September 30

    Arena said the Direct Mode test would start at 8 a.m. PT, with the model selectable from a dropdown menu. It said the model was already available in Battle and Agent modes.

    Arena's Series B: $200M at $3.1B, and a New Way to Measure AI AlignmentArena AI · YouTube
  • Causal tracing: a new way to evaluate AI agentsArena AI · YouTube
  • GPT-6.1 arrives in Agent Arena and Code Arena

    Arena says users can test GPT-6.1 in Battle Mode or Agent Mode. Votes will shape its evaluation, and Arena said scores would follow.

    Arena showcases frontier models bringing Ancient Rome to life

    Arena says the showcase traces model progress from Q4 2025 to September 2026. It features Claude, GPT, Gemini, Kimi and Qwen models, with scores for Claude Sonnet 5.5 and GPT-6.1 coming soon.

    18 Sources

    Arena.ai@arenaIntroducing the Arena Alignment Index, our new benchmark measuring safety and alignment of AI agents in real-world use. Built from 90K+ real-world agent sessions across 27 models, the index measures three critical signals: - Unauthorized Action (UA): Taking actions beyond the user's instructions or permissions - False Attribution (FA): Attributing statements or actions that are contradicted by user-provided evidence - Deceptive Completion (DC): Claiming a task was completed when it was not. Key findings: - OpenAI models currently lead the Alignment Index - Rogue actions are rare, but can have serious consequences when they occur - Agents can mislead users about task progress - Misalignment risks increase with conversation length - Safety and alignment are improving across model generations As shown in the leaderboard below (sorted by lab), @OpenAI’s GPT-6.1-Sol leads the Arena Alignment Index with a score of 87.9, followed by @AnthropicAI’s Claude-Opus-5.5 at 83.2 and @SpaceXAI's Grok-4.7 at 82.7. OpenAI also has the best observed rates across all three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. Across all four labs, newer models consistently outperform their predecessors, suggesting broad progress in agent safety and alignment. As agents take on longer, more complex, and higher-stakes tasks, measuring not just what they can accomplish, but how safely and reliably they act, becomes increasingly important. This marks an important step toward making safety and alignment a core part of how Arena evaluates AI. The index is an initial starting point, and we'll continue expanding the index with additional safety signals and models over time. More analysis below👇3h
    Anastasios Nikolas Angelopoulos@ml_angelopoulosArena has raised a $200M Series B at a $3.1B valuation. We are using this momentum to evaluate agentic utility in real workflows from every economically valuable industry on Earth. We seek to give everyone universal and frictionless access to the best intelligence. The @arena platform has 10s of millions of monthly users doing complex coding and agentic work tasks. We have gone way beyond human preference, and most of our tokens go to deep, single threaded workflows with dozens of turns where we can evaluate agents while humans are in their true flow state. And we make leaderboards by observing whether they get their jobs done. This is how benchmarks should be: representative of actual human value, so we incentivize AI to benefit humanity. Today, we are releasing one of the most exciting developments in the history of our platform. The Arena Alignment Index. In observing how humans collaborate with agents, we have witnessed that in a surprisingly large fraction of cases, agents are attempting to deceive users or take unauthorized actions beyond what the user intended or permissioned. Agents are deleting files and lying to users about what they did or didn’t do. The longer conversations go, the higher the probability that an agent will violate alignment. The results are really striking; see the plot below. And because this represents in-the-wild usage, the probability that those of us reading this have experienced some form of misalignment from our agents is nearly 100%. And we may not even have noticed. This is exactly why it’s important to have a neutral 3rd party that is evaluating AI alignment on the distribution of actual user-agent interactions. Arena is excited to step into that role. We are also incredibly grateful for the incredible and experienced investors backing our vision. Our Series B is co-led by @lightspeedvp and @khoslaventures with participation from @SalesforceVC, @01Advisors, @DellTechCapital and @endvr_catalyst. This round includes support from our existing investors @a16z, @felicis, @amppublic, @QuantumLightVC, @thehousefund, and others. Which brings me to the real headline: we're hiring, across ML, product engineering, design, marketing, BD, and more. If evaluation and alignment at the frontier is your thing, and you’re looking for an environment with exceptional research and engineering, come join us. Job board link in @arena’s bio and in the thread below LFG (safely)!3h
    Ion Stoica@istoica05Huge day. $200M Series B at a $3.1B valuation for @arena and the release of our Alignment Index. I am proud of this team for solving one of the hardest problems in AI: evaluating AI under real-world conditions. We're still a small team, about 90 people, and we're hiring across ML, product, engineering, design, marketing, BD, and more. If evaluations, safety, and measuring progress interest you, check out the open roles in @arena's bio.3h
    Tal Broda@talbrodaWe @khoslaventures are proud to work with @arena, led by the amazing @ml_angelopoulos. They built the industry's most important AI leaderboard, and are now measuring how agents behave in the real world. This work is critical to delivering safe & aligned AGI!3h
    Wei-Lin Chiang@infwinstonWe're excited to announce a major milestone for Arena: our $200M Series B at a $3.1B valuation, accelerating our mission to measure and advance frontier intelligence! At Arena, we believe AI evaluation must be grounded in real-world use. We put agents in live tool-use environments, challenge them with complex, long-horizon tasks, and analyze rich signals from millions of real-world user-agent interactions to understand how they perform. Today, we're also introducing the Arena Alignment Index, extending our evaluations beyond capability to measure how safely, reliably, and faithfully agents act in accordance with human intent. As AI systems become increasingly autonomous, evaluating what they can do is no longer enough. Understanding whether they reliably follow human intent, navigate uncertainty, and act safely in real-world environments is becoming one of the most important challenges in AI. We're excited to continue advancing the research and building the infrastructure needed to rigorously measure both capability and alignment at the frontier of AI. If you’re passionate about this mission, come build with us!3h
    Anjney Midha@AnjneyMidhahow do we scale AI securely? first, measure what matters proud of @ml_angelopoulos @infwinston @istoica05 and team’s dedication to the reliability mission from day 0 through hypergrowth2h
    Hao Zhang@haozhangmlcongrats to @arena !!2h
    Tyler Deck@supdeck1AI agents make false completion claims in 10% of sessions in Arena's new study. For code debugging, it jumps to 48%. Arena's Alignment Index studied 90,000 real sessions across 27 models. The leaderboard maker also raised $200M at a $3.1B valuation. https://arena.ai/blog/ai-alignment-index50m
    Axon Review@axonreviewArena Intelligence, behind the AI Model Arena leaderboard, raised $200 million at a $3.1 billion valuation and plans to expand into AI safety evaluation.43m
    AI Fear & Trust Index@AI_fear_indexPopular AI leaderboard Arena nearly doubles valuation to $3.1B valuation in 10 months Source: TechCrunch38m