EveryDev.ai
Subscribe
Home
Tools

3,431+ AI tools

  • New
  • Trending
  • Featured
  • Compare
  • Arena
Categories
  • Agents2189
  • Coding1574
  • Infrastructure698
  • Marketing534
  • Projects498
  • Research456
  • Design416
  • Analytics389
  • Testing296
  • MCP290
  • Security286
  • Data262
  • Integration197
  • Prompts189
  • Communication183
  • Extensions173
  • Learning170
  • Voice151
  • Commerce135
  • DevOps123
  • Web86
  • Finance26
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Tools
    3. Arena (LMArena)
    Arena (LMArena) icon

    Arena (LMArena)

    Academic Research
    Featured

    A community-powered platform for evaluating and comparing frontier AI models through real-world human feedback, featuring a public leaderboard and battle mode.

    Visit Website

    At a Glance

    Pricing
    Free tier available

    Access to Arena's battle mode and public leaderboard at no cost.

    AI Evaluations: Custom/contact

    Engagement

    Available On

    Web
    API

    Resources

    WebsiteDocsGitHubllms.txt

    Topics

    Academic ResearchConversational AgentsLLM Evaluations

    Alternatives

    Artificial AnalysisDesign ArenaTracking AI
    Developer
    LM ArenaSan Francisco, CAEst. 2023$250M raised

    Updated Jul 2026

    About Arena (LMArena)

    Arena (formerly LMArena) is a community-powered platform created by researchers from UC Berkeley for understanding AI model performance in real-world conditions. The platform lets users interact with frontier AI models, compare their outputs side-by-side in "Battle Mode," and submit feedback that shapes a public leaderboard grounded in actual human preference data. According to the About page, tens of millions of builders, researchers, and creative professionals use Arena to engage with frontier models.

    What It Is

    Arena is an AI model evaluation and benchmarking platform that collects human preference signals at scale. Users can prompt multiple AI models simultaneously, vote on which response is better, and contribute to a continuously updated public leaderboard. The platform's core thesis is that real-world human feedback is a more reliable signal of model quality than synthetic benchmarks alone. The project originated from UC Berkeley research and has since grown into a broader community and commercial offering under the "Arena" brand.

    Battle Mode and Core Workflow

    The flagship feature is Battle Mode, where two AI models respond to the same prompt and users vote on the better answer. This blind comparison format is designed to reduce bias and generate preference data that reflects genuine user needs. The homepage also highlights practical use cases such as creating landing pages, building dashboards, making browser games, converting designs to code, building full-stack apps, and launching storefronts — suggesting the platform targets developers and builders as a primary audience.

    Leaderboard and Evaluation Services

    Arena publishes a public leaderboard ranking AI models based on aggregated human preference votes. The platform also offers an "AI Evaluations" service aimed at enterprises, model labs, and developers, providing comprehensive evaluation grounded in real-world human feedback. This commercial arm is positioned as a way for organizations to benchmark their models against real user preferences rather than static test sets.

    Community and Research Roots

    The platform was built by UC Berkeley researchers and maintains a strong community orientation. Arena operates a Discord server, X/Twitter account, LinkedIn page, and YouTube channel to connect researchers, developers, and AI enthusiasts. The About page frames the mission as measuring and advancing the frontier of AI for real-world use, with a vision of building the foundation for everyone to understand, shape, and benefit from AI.

    Rebranding: LMArena to Arena

    The platform was previously known as LMArena and has since rebranded to Arena (arena.ai). The lmarena.ai domain now redirects to the new brand. The GitHub organization remains under the lmarena handle, and the blog repository (lmarena.github.io) is open-source under the MIT License. The core platform itself is a web-based service, not an open-source product.

    Arena (LMArena) - 1

    Community Discussions

    Start a new discussion about Arena (LMArena)
    Joe Seifi's avatar
    Joe Seifi
    January 23, 2026·Apple, Disney, Adobe, Eventbrite,…

    Video Arena is out on LM Arena and it's pretty wild

    So LM Arena shipped their Video Arena feature and I've been messing around with it. You can now compare AI video models side by side, things like Sora, Hailuo, Veo 3.1, and a bunch of others. The cool part is it runs through their Discord server and you get to vote on which model output you like bet…

    0
    news

    Pricing

    FREE

    Free

    Access to Arena's battle mode and public leaderboard at no cost.

    • Battle Mode AI comparisons
    • Public leaderboard access
    • Frontier model access
    • Community participation

    AI Evaluations

    Enterprise evaluation service for model labs, enterprises, and developers grounded in real-world human feedback.

    Custom
    contact sales
    • Comprehensive AI model evaluation
    • Real-world human feedback grounding
    • Custom evaluation for enterprises and model labs
    • Dedicated team support
    View official pricing

    Capabilities

    Key Features

    • Battle Mode for side-by-side AI model comparison
    • Public AI model leaderboard based on human preference votes
    • Support for frontier AI models
    • Real-world human feedback collection
    • AI Evaluations service for enterprises and model labs
    • Practical use-case prompts (landing pages, dashboards, games, apps)
    • Community-driven model ranking
    • Search conversation history

    Integrations

    Multiple frontier AI model providers
    API Available
    View Docs

    Ratings & Reviews

    No ratings yet

    Be the first to rate Arena (LMArena) and help others make informed decisions.

    Developer

    LM Arena

    LM Arena builds a web platform for evaluating and deploying large language models. The team builds tools for model comparison, hosted inference, and API integration. They focus on simplifying model workflows for engineers and researchers. The company emphasizes accessible deployment and monitoring features.

    Founded 2023
    San Francisco, CA
    $250M raised
    40 employees

    Used by

    OpenAI
    Google DeepMind
    Anthropic
    Meta
    +9 more
    Read more about LM Arena
    WebsiteGitHubX / Twitter
    1 tool in directory

    Similar Tools

    Artificial Analysis icon

    Artificial Analysis

    Independent AI benchmarking platform that evaluates and compares AI models across intelligence, speed, cost, and capabilities to help users choose the best model and provider for their use case.

    Design Arena icon

    Design Arena

    A crowdsourced benchmark platform that pits top AI models against each other on design tasks and lets users vote to power live leaderboards.

    Tracking AI icon

    Tracking AI

    A free web tool that quizzes 17+ AI models weekly on IQ tests and political compass questions to monitor and compare AI biases and capabilities over time.

    Browse all tools

    Related Topics

    Academic Research

    AI tools designed specifically for academic and scientific research.

    56 tools

    Conversational Agents

    AI chatbots and virtual assistants that can engage in natural dialogue.

    285 tools

    LLM Evaluations

    Platforms and frameworks for evaluating, testing, and benchmarking LLM systems and AI applications. These tools provide evaluators and evaluation models to score AI outputs, measure hallucinations, assess RAG quality, detect failures, and optimize model performance. Features include automated testing with LLM-as-a-judge metrics, component-level evaluation with tracing, regression testing in CI/CD pipelines, custom evaluator creation, dataset curation, and real-time monitoring of production systems. Teams use these solutions to validate prompt effectiveness, compare models side-by-side, ensure answer correctness and relevance, identify bias and toxicity, prevent PII leakage, and continuously improve AI product quality through experiments, benchmarks, and performance analytics.

    112 tools
    Browse all topics
    Back to all toolsSuggest an edit
    ratings
    1discussion
    367views
    1upvote