Public behavioral test environment for AI agents with stateful rides and signed scorecards.
Listed on
First seen 2 Oct 2026. One server, whatever directories list it: each directory listing keeps its own page and history.
2
Directories
Collected by InvokeRank
4
Tools
From an anonymous probe
-
ToolBench grade
Not graded by Arcade
0
GitHub stars
From MCP Toplist
Tools
| Tool | Description | Behaviour |
|---|---|---|
| act_in_ride | Submit one strictly validated action for the ride identified by runId. Use a payload variant allowed by the latest observation. Structural errors do not enter the ride or consume a step; valid payloads, including semantically poor choices, do. Returns events, the next observation, and the completion rating when finished. | Changes data |
| get_scorecard | Retrieve the signed scorecard for a completed MCP ride by runId. It reports this one simulated run and is not a safety certification. | Read-only |
| list_rides | Discover public behavioral evaluation rides and their missions. | Read-only |
| start_ride | Start one public simulated ride for this agent to take. Returns its first observation and allowed actions. No target URL or real-world system is tested. | Changes data |