APIEval-20

APIEval-20

Evaluate AI agent API testing, including multi-domain scenarios and error injection

APIEval-20AI代理API测试黑箱基准JSON架构测试套件实时参考API错误检测API覆盖率效率评分客观评分weboversea

This product is a task benchmark for evaluating AI agents in actual API tests, covering 20 scenarios in 7 domains. It can measure the ability to discover vulnerabilities from the architecture and payload. It provides 20 API scenarios from multiple domains such as e-commerce, payment, and authentication, including request patterns and sample payloads, challenging the generated test suite to find hidden errors. Each scenario has 3 to 8 errors implanted according to complexity classification, which can test the API's handling ability for different problems.

COMMUNITY PERKSMergeek 社区专属福利
独家限免 NEW每天发现一份惊喜
PRODUCT SCREENS / 8

点击查看界面细节

APIEval-20 产品截图 1APIEval-20 产品截图 2APIEval-20 产品截图 3APIEval-20 产品截图 4APIEval-20 产品截图 5APIEval-20 产品截图 6APIEval-20 产品截图 7APIEval-20 产品截图 8

This is a task benchmark for evaluating AI agents in actual API tests, covering 20 scenarios in 7 domains, and its ability to discover vulnerabilities can be measured solely from the architecture and payloads.

产品功能

  • 提供20个精心设计的来自真实应用领域的API场景,涵盖电商、支付、认证等多领域,每个场景给出API请求模式和示例负载,挑战生成测试套件找出实时参考实现中的隐藏错误
  • 每个场景包含3到8个植入的错误,按复杂性分类,从简单到复杂,简单错误测试API对基本结构问题的处理,中等错误需对领域有一定理解

你的真实体验,为其他用户提供宝贵参考,获得宝石奖励

REWARD MARKET100 宝石可兑换什么?

还没有公开评论,欢迎分享你的体验。