τ-bench is an open-source benchmarking framework developed by Sierra AI designed to evaluate the performance of AI agents in collaborative, real-world enterprise scenarios. It challenges agents to coordinate, guide, and assist users in achieving shared objectives across complex domains such as banking and voice interactions. The platform provides standardized tasks and leaderboards to measure agent capabilities in multi-turn, goal-oriented workflows.

LLM mention score The LLM mention score is the total number of mentions of this brand in different LLM chatbots, normalized to the scale from 0 to 100. You can get actual, non-normalized numbers via the LLM Mention API from DataForSEO.

Normalized 0–100 · last 8 weeks

DataForSEO API

Get LLM mention data of any company via DataForSEO API

Get access to the structured data on keyword, brand, and website mentions in LLMs, including metrics like AI search volume, impressions, and mentions count. 

How to get LLM mention data →
// Fetch τ-bench mentions
POST v3/ai_optimization/llm_mentions/search/live
[
    {
        "target": [
            {
                "keyword": "τ-bench",
                "search_scope": ["any"]
            }
        ],
        "platform": "chat_gpt",
        "order_by" : ["ai_search_volume,desc"]
    }
]