EveryDev.ai
Subscribe
Home
Developers

3,289+ AI companies

  • Radar
  • Trending
AI Tools by Topic
  • AI Coding Assistants
  • Agent Frameworks
  • MCP Servers
  • AI Prompt Tools
  • Vibe Coding Tools
  • AI Design Tools
  • AI Database Tools
  • AI Website Builders
  • AI Testing Tools
  • LLM Evaluations
Follow Us
  • X / Twitter
  • LinkedIn
  • Reddit
  • Discord
  • Threads
  • Bluesky
  • Mastodon
  • YouTube
  • GitHub
  • Instagram
Get Started
  • About
  • Editorial Standards
  • Corrections & Disclosures
  • Community Guidelines
  • Advertise
  • Contact Us
  • Newsletter
  • Submit a Tool
  • Start a Discussion
  • Write A Blog
  • Share A Build
  • Terms of Service
  • Privacy Policy
Explore with AI
  • ChatGPT
  • Gemini
  • Claude
  • Grok
  • Perplexity
Agent Experience
  • llms.txt
Theme
With AI, Everyone is a Dev. EveryDev.ai © 2026
    1. Home
    2. Developers
    3. Agent Memory Leaderboard

    Agent Memory Leaderboard

    Agent Memory Leaderboard (AML) is an open, unified and reproducible evaluation platform for long-term memory systems and memory-enabled agents. It standardizes Add/Search interfaces and holds the Answer, evaluation, orchestration, scoring and publication workflow constant so textual, coding and multimodal memory systems can be compared fairly.

    Visit Website

    At a Glance

    1Tool Listed
    3Products
    9Capabilities
    Discussions
    2026Est.
    Focus Areas
    Agent Memory
    LLM Evaluations
    Academic Research
    Connect
    Latest News
    Cycle 2 Agent Memory Challenge scheduled to open with three tracks and RMB 150,000 prize poolSep 20, 2026
    Official site confirmed the first cycle had closed and Cycle 2 was expected to open September 20Aug 24, 2026
    Markets
    • Universities and academic researchers
    • Research institutes and independent research teams
    • Open-source memory-system maintainers
    • Commercial AI memory products and API providers
    • +1 more

    AI Tools by Agent Memory Leaderboard

    (1)
    View Agent Memory Leaderboard
    Agent Memory Leaderboard tool icon

    Agent Memory Leaderboard

    AI Agent Memory Benchmark

    Agent MemoryLLM EvaluationsAcademic Research

    Discussions

    No discussions yet

    Be the first to start a discussion about Agent Memory Leaderboard

    Latest News

    09/20/2026

    Cycle 2 Agent Memory Challenge scheduled to open with three tracks and RMB 150,000 prize pool

    agentmemoryleaderboard.ai
    08/24/2026

    Official site confirmed the first cycle had closed and Cycle 2 was expected to open September 20

    github.com
    08/22/2026

    CSIG announced the second Agent Memory Challenge and its host/organizing institutions

    m.csig.org.cn
    08/17/2026

    MemoraX AI ranked first in AML's inaugural commercial ranking, bringing attention to the new benchmark

    globenewswire.com

    Products & Services

    3
    Agent Memory Leaderboard platform
    2026-08-12

    Public ranking and evaluation platform comparing textual, multimodal and coding-agent memory systems under versioned datasets, fixed answer/evaluation settings and capability-level metrics.

    Agent Memory Challenge
    2026-07-29

    Recurring public evaluation challenge for researchers, open-source maintainers and commercial product teams. Participants host Add and Search APIs; AML runs the standardized Answer, Eval, orchestration, review and publication process.

    Agent Memory Leaderboard evaluation repository
    2026

    Public GitHub repository containing per-benchmark evaluation contracts, public answer/scoring behavior and runtime configuration for transparency and methodological review; it is not the production leaderboard service and excludes protected benchmark data and participant artifacts.

    Market Position

    AML positions itself as a neutral measurement and comparison layer rather than as a memory database or agent-memory product. It differentiates through a common Add/Search contract, fixed downstream Answer/Eval conditions, private evaluation and review, version traceability, and separate comparable boards; the inaugural commercial board compared systems including MemoraX, MemOS, Mem0, Vectorize, SuperMemory, TencentDB and NetEase submissions.

    Founding Story

    AML was created to address the lack of directly comparable agent-memory results: memory systems had been evaluated on different datasets, with different answer models, retrieval settings, judges and aggregation rules. The organizers' initial vision was an open benchmark that would provide broad coverage, controlled comparison and capability-level diagnosis; the repository says AML launched on July 29, 2026, with researchers from more than 20 universities and research organizations.

    Business Model

    Revenue Model

    The public challenge is free to enter. Participants pay their own API, database, bandwidth and compute costs, while AML covers the unified Answer, Eval and evaluation-orchestration costs. No subscription, API usage price or other commercial revenue model is described in the reviewed sources.

    Pricing Tiers

    Public challenge participation
    Free

    No registration fee; participant teams provide and operate their own Add/Search service and infrastructure.

    Target Markets

    Industries & Segments
    • Universities and academic researchers
    • Research institutes and independent research teams
    • Open-source memory-system maintainers
    • Commercial AI memory products and API providers
    • Enterprise systems and teams building memory-enabled agents
    Use Cases
    • Comparing long-term memory systems across research methods and commercial APIs
    • Evaluating long conversations, cross-session history, personal preferences, temporal events and continuous narratives
    • Testing coding agents' retrieval and reuse of prior debugging, architecture and development experience
    • Testing memory over image-rich or other multimodal tasks
    • Diagnosing memory-system strengths, failure modes, governance and privacy behavior
    • Submitting reproducible open-source methods or verifiable hosted memory products

    Quick Facts

    Founded
    2026

    History & Milestones

    2026-07-29

    Registration opened for the first Agent Memory Challenge; the project describes this as its launch date.

    2026-08-12

    The first public leaderboard release was published, with separate open-source/academic-method and commercial/industry-system rankings and textual, coding and multimodal tracks.

    2026-08-22

    CSIG announced the second Agent Memory Challenge, hosted by the China Society of Image and Graphics and organized with Nanjing University, Zhejiang University, Datawhale and the CSIG Enterprise Liaison and Standardization Committee.

    2026-09-20

    Cycle 2 is scheduled to open at 00:00 UTC+8; it has a RMB 150,000 open-method prize pool across textual, coding and multimodal tracks.

    2026-11-04

    Cycle 2 evaluation is scheduled to close at 23:59 UTC+8, with official results planned for mid-November 2026.

    Key Capabilities

    9
    Unified participant-hosted Add and Search API protocol
    Fixed Answer, evaluation, judge, aggregation and orchestration pipeline
    Versioned evaluation contracts with benchmark, pipeline, model and scoring configuration traceability
    Textual memory evaluation covering fact recall, multi-hop reasoning, temporal/event understanding, governance, personalization, process execution and safety/privacy
    Coding-agent memory evaluation covering debugging and development memory
    Multimodal memory evaluation covering cross-modal retrieval and evidence-grounded answers

    Integrations & Partnerships

    Platform Integrations

    • Participant-hosted HTTP Add/Search APIs
    • GitHub repositories and fixed commits for open-source-method disclosure and verification
    • Hugging Face organization and leaderboard Space for public releases and community updates
    • Official website leaderboard, evaluation console, documentation and API Guide

    Key Partnerships

    Hosted by the China Society of Image and Graphics (CSIG)
    Organized by Nanjing University, Zhejiang University, Datawhale and the CSIG Enterprise Liaison and Standardization Committee
    Public evaluation/release presence on Hugging Face

    Connect

    Website
    agentmemoryleaderboard.ai
    GitHub
    AML-memory
    X / Twitter
    AgentMemoryL

    AI Topics

    3

    Agent Memory Leaderboard focuses on these topics:

    Agent Memory(1)
    LLM Evaluations(1)
    Academic Research(1)
    Back to all developersSuggest an edit