# AI Agent Evaluation Platforms > Platforms designed for testing, evaluating, and validating the performance, reliability, and safety of AI agents, particularly through simulation of multi-turn conversations and complex interactions. - Brands: 15 - URL: /brand/category/ai-agent-evaluation-platforms ## Brands - Collinear AI - Radiant Labs - Coval - Bifrost - Coxwave - Roark (YC W25) - Cekura - ContextQA - Converra - Botcopy - Noveum.ai - Adaline - Futureagi - Cleanlab - Arklex Ai --- ## Full Details / RAG Data ### Category Metadata | Field | Value | |----------------|-------| | Name | AI Agent Evaluation Platforms | | Slug | ai-agent-evaluation-platforms | | URL | /brand/category/ai-agent-evaluation-platforms | | Brand Count | 15 | | Last Updated | 2026-09-24T15:48:47.876Z | ### Brand Listing | Name | Slug | BAI Score | |------|------|-----------| | Collinear AI | collinear-ai | — | | Radiant Labs | radiant-labs | — | | Coval | coval | — | | Bifrost | bifrost | — | | Coxwave | coxwave | — | | Roark (YC W25) | roark-yc-w25 | — | | Cekura | cekura | — | | ContextQA | contextqa | — | | Converra | converra | — | | Botcopy | botcopy | — | | Noveum.ai | noveum-ai | — | | Adaline | adaline | — | | Futureagi | futureagi | — | | Cleanlab | cleanlab | — | | Arklex Ai | arklex-ai | — | ### Links - Canonical page: /brand/category/ai-agent-evaluation-platforms - JSON endpoint: /brand/category/ai-agent-evaluation-platforms.json - LLMs.txt: /brand/category/ai-agent-evaluation-platforms/llms.txt