Dimensions
Five dimensions.One AX Score.
Scoring model
A balanced measure of agent experience
Each dimension contributes equally. A strong total cannot hide a product that agents cannot operate, nor can technical operability compensate for inaccurate representation.
Find · 0 to 20 points
Discoverability
Discoverability measures whether agents can find your product when researching a category. When a user asks an AI for tool recommendations, does your product appear in the answer?
Strong AX
- Mentioned prominently in category queries across models.
- Accurately placed in the correct product category.
- Consistent positioning across ChatGPT, Perplexity, Claude, and Gemini.
- Appears in comparison queries alongside direct competitors.
Weak AX
- Omitted from category listings entirely.
- Confused with a competitor by one or more models.
- Appears in some models but invisible in others.
- Associated with outdated or incorrect category terms.
| Score | Band | Interpretation |
|---|---|---|
| 16–20 | Excellent | Found accurately across all major AI platforms |
| 11–15 | Good | Found reliably, some positioning gaps remain |
| 6–10 | Fair | Found inconsistently across models |
| 0–5 | Poor | Not found in AI category queries |
Act · 0 to 20 points
Operability
Operability measures whether agents can use your product to complete tasks: both through your API and by recommending your product for specific jobs.
Strong AX
- API returns typed, structured, self-describing responses.
- Agents can authenticate without browser-only OAuth flows.
- Error messages include recovery guidance, not just status codes.
- Rate limits are declared upfront with retry information.
Weak AX
- API returns generic error messages with no context.
- Authentication requires a visual browser flow.
- Endpoints behave differently based on undocumented conditions.
- Rate limit responses include no Retry-After guidance.
| Score | Band | Interpretation |
|---|---|---|
| 16–20 | Excellent | Tasks complete reliably with full recovery support |
| 11–15 | Good | Most tasks succeed, error handling needs work |
| 6–10 | Fair | Basic tasks work, failures are opaque |
| 0–5 | Poor | Core agent tasks cannot be completed reliably |
Correct · 0 to 20 points
Recoverability
Recoverability measures whether agents self-correct when they hold wrong information about your product.
Strong AX
- Agents update incorrect beliefs when given accurate context.
- Agents flag uncertainty rather than asserting incorrect facts.
- Agents distinguish between current and outdated information.
- Agents recommend verification for rapidly changing details.
Weak AX
- Agents confidently assert false information about your product.
- Agents repeat incorrect pricing or feature claims.
- Agents cannot distinguish your product from a competitor.
- Agents describe deprecated features as current.
| Score | Band | Interpretation |
|---|---|---|
| 16–20 | Excellent | Agents reliably update and flag uncertainty |
| 11–15 | Good | Agents usually self-correct, with occasional persistence |
| 6–10 | Fair | Agents sometimes self-correct with strong prompting |
| 0–5 | Poor | Agents repeat errors even when corrected |
Qualify · 0 to 20 points
Transparency
Transparency measures whether agents accurately represent your product's limitations, scope, and appropriate use cases.
Strong AX
- Agents acknowledge what your product cannot do.
- Agents qualify recommendations with relevant caveats.
- Agents represent your pricing and access model accurately.
- Agents can articulate when to use a competitor instead.
Weak AX
- Agents overclaim your product's capabilities.
- Agents omit limitations that affect buying decisions.
- Agents describe your product as suitable for all use cases.
- Agents describe enterprise-only features as universally available.
| Score | Band | Interpretation |
|---|---|---|
| 16–20 | Excellent | Agents represent capabilities and limits precisely |
| 11–15 | Good | Agents usually accurate, some overclaim remains |
| 6–10 | Fair | Agents omit significant limitations regularly |
| 0–5 | Poor | Agents consistently overclaim your capabilities |
Audit your AX
The AX Score rates any product across five dimensions in ten minutes. No tools required.