Stop guessing which trust problem to fix first.
The seven-axis score shows exactly where confidence is leaking, so your next sprint targets the dimension that's losing users, not the loudest opinion in the room.
The System Usability Scale was built for software that behaves the same way every time. AI doesn’t. AI-UX Score is a respondent-backed assessment of how people experience your AI across seven dimensions of trust and usability, the things SUS was never designed to see.
Classic UX research measures how fast someone finishes a task, how many clicks it took, and whether the interface was clear. Those questions still matter, but they assume a system that responds the same way every time. AI is probabilistic, adaptive, and occasionally wrong in ways a user can't predict.
When your product's behaviour is probabilistic, you need a metric built for probability: one that measures perceived trust, control, and accountability, not just whether the button was easy to find.
Not a spreadsheet of averages, but a living report your team can read at a glance and take straight into planning.
Acceptable overall, but real friction is showing. Focus on the weakest dimensions.
24 responses
The shape tells you where trust leaks before you read a word: one headline score, and a seven-axis radar anyone can act on.
"People don't trust it" becomes "Accountability 2.9, Privacy 3.5." Every dimension gets its own score and grade; open any row to read the questions behind the number.
Run it again after every fine-tune or redesign. Ten rounds in, the number that used to be a gut feeling is a line you can defend.
Take one clean PDF into a design review or leadership deck, or pull the raw responses into your own analysis. Both come on any plan; the reliability stats that back every number, Cronbach's α and 95% confidence intervals, come with Pro.
A score is the start, not the point. Here's what the assessment turns into: a clear diagnosis, and a backlog you can ship.
The seven-axis score shows exactly where confidence is leaking, so your next sprint targets the dimension that's losing users, not the loudest opinion in the room.
Every weak dimension becomes ranked, concrete design work, grounded in what that dimension measures. A plan to start from, not a verdict to sit with.
Re-run the assessment after you ship and watch the score move across rounds: the evidence you bring to the next roadmap conversation, not a hunch.
The action plan gives AI-generated suggestions: a starting point to review with your team, not a substitute for talking to users. A Pro feature.
AI-UX Score is respondent-backed: real users rate their experience right after using your AI feature. Their answers roll up into seven dimensions, each one a place where trust is either earned or lost.
Generate a lightweight assessment link, or embed the survey right inside your product flow. Users answer in under five minutes.
Responses synthesise automatically into your seven-axis radar, with a sub-score for every dimension and a read on how much confidence your sample supports.
Get a prioritised gap report: exactly where trust is leaking, and which design changes to make next.
Two clocks, and we keep them honest. Setup takes minutes. Collecting responses takes days, because it depends on real people replying, and that wait is the point.
Plan for about 25 responses in total (a total, not per dimension) to reach a reliable read, and gather them from your users, not your team. A group rating its own work is exactly the self-assessment this instrument exists to replace, which is what makes the number worth defending.
Your first project, forever.
For teams measuring over time.
They're companions, not rivals. SUS gives one usability number, validated on conventional software where behaviour is fixed. AI-UX Score is built for probabilistic systems: seven dimensions SUS never covers (trust, transparency, agency, accountability, fairness, data safety, quality). Use SUS for the interface, and AI-UX Score for the AI. Many teams run both. Meet the SUS Calculator →
Around 25 gives a reliable read for most teams, and the tool tells you how much confidence your sample supports. Fewer gives a directional signal; 50–100+ is better for benchmarking over time.
Yes. The seven dimensions are fixed (that keeps scores comparable), but each item has an [AI system] placeholder you set to your context ("this assistant", "the recommendation engine").
Yes. Your first project is free: one project, up to 25 responses, no card, enough to run a real assessment and get your score and grade across all seven dimensions. Pro lifts the cap to unlimited responses, projects, and rounds.
Set up your first AI-UX survey in under two minutes. No credit card required.
Free first project · up to 25 responses · no credit card