Agent Readiness: Notion vs Asana vs ClickUp
Standardized task: create a project and invite a teammate across three provider-compatible demo adapters and repeated runs.
Equivalent task success, median actions, recovery and provider consistency explain the ranking.
Structured labels identify navigation, hidden-control, modal and unclear-labeling friction.
Six normalized component scores and published weights reproduce the final score.
Recommendations include evidence-based confidence labels and validation steps.
Evidence boundary: This example uses deterministic simulated runs to demonstrate the reporting system. It is not live-agent performance evidence.