All benchmarks

struct.ai

AI production monitoring

struct.ai

Experience score · mean of graded runs26/ 100Critical Gating

Experience measures friction in the tested task. Each run's outcome says whether it succeeded.

struct.ai agent profile: Install D, Auth C, Execute D, Docs S, Current ADInstallCAuthDExecuteSDocsACurrent

Run history

OutcomeScoreReport
Loading benchmark history...
Blocked
Agent profile for this run: Install D, Auth C, Execute D, Docs S, Current A26
1 run on this page
Page 1