The dimensions, the levels, the way answers become a number, and the things this measure will not claim. The option-by-option scoring key is the one thing not on this page: it goes to reviewers on request, because a key you have read is a score you chose rather than a measurement.
Fixed, and the same for everyone. They are not weighted: the index is the plain mean of the five, so no dimension quietly counts more than another.
Every dimension is scored one to four. The words matter more than the numbers, because the numbers average and the words do not.
You have not yet worked out where this applies to you, so the risk it carries is unmanaged rather than accepted.
You are using AI here and finding out what it does. Results are real but inconsistent, and you cannot yet predict which you will get.
You use this reliably for real work and know where it stops. This is the level at which AI is genuinely doing something for you.
You direct this deliberately, check it where checking matters, and have built it into how work gets done rather than doing it by hand each time.
Each option in each question is mapped to one of the four levels. You pick an option, that option carries a level, and your score for a dimension is the mean of the three answers you gave for it. Three questions per dimension means a score lands on thirds rather than whole numbers, which is what lets it tell a 3.0 from a 3.3.
The mapping lives on the server and the browser only ever sends the option you chose, so the answer key is not readable in the page.
You're on a deadline. An AI tool gives you a specific market figure you need for a slide and cites a named, real-looking source. You have 20 minutes left.
What actually demonstrates good judgment here?
The highest-scoring answer is frequently not the most thorough or the most cautious one. Over-caution is a real failure mode and is scored as one.
If any single dimension comes out at Unaware, the band next to your score is held at Exploring however high the mean is. The number is never altered, only the word.
An average hides its worst component. Someone Fluent on four dimensions and Unaware on Verification averages 3.4 and would otherwise read as broadly competent while being unable to tell whether AI output is true. That is the exact failure this measure exists to catch, so it is not averaged away.
The recommendation is derived from the result, not sold. Two things decide who appears:
Anyone missing a photo, a description or a booking link is not shown at all, so a card is never a placeholder. Nobody is ranked by what they pay.
Disclosure. AIFluency operates the Coaching Hub and earns a fee on bookings made through it. We both measure the gap and sell help closing it, and you should read the recommendation knowing that.
This is a self-assessment of judgment across five dimensions. It is not a credential, and it has not been validated for use in selection, promotion, compensation or termination decisions. It has no established predictive validity for job performance, and the level mappings are ours rather than an external standard.
Use it to decide what to work on, and to see where a team is thin. Do not use it to decide who to hire, who to promote, who to pay more, or who to let go.
The level mappings are set by AIFluency, informed by how each dimension plays out in real work. They are published rather than hidden so you can judge them yourself.
None of the fifteen has been reviewed independently. Every level assignment is a considered first pass by the author and has not been keyed by an outside practitioner. Until that happens, no single result is used to make a claim about any individual.