F.AI / AI competency

Help your people get more from AI.

Activity tells you who is using AI. F.AI looks at how they use it: how they explain the job, guide the work and check the answer.

Talk about F.AI ↗
Your F.AI scoreIllustrative assessment · 30 days
68 / 100

Strong · 24 assessed conversations

You regularly refine the output and explain what you need. Check supporting facts more consistently before sharing the result.

How the score works ↓

Use the evidence to guide support and development. The score alone does not measure job performance or financial return.

Evidence, then advice

Show people what they’re doing well. Give them a next step.

Look at the behaviors recorded in assessed conversations. Keep the reason for each assessment alongside it, then suggest something specific to try in the next task.

Better context, reusable examples and clear checks can reduce wasted attempts. Measure that change in the work itself.

From a session to useful feedbackFictional session examples

Finance / Explain a variance

Explain the difference for our budget owners. Use this table layout. I checked your totals against the source: the tax line is counted twice. Correct it before we share it.

Behaviors visible in this example

  • Defines an audience
  • Specifies a format
  • Checks facts

What to try next

Add an example of a useful explanation. Refine the first draft against it.

These examples illustrate indicator assessment. A personal score needs at least 10 assessed conversations. One short prompt cannot establish someone’s overall capability.

How it works

A score grounded in observable habits.

1. Assess the conversation

An AI judge checks the user’s messages for 11 observable behaviors, including clarifying goals, providing examples, refining output and checking facts. Each assessment includes a rationale and confidence.

2. Look across sessions

For each behavior, calculate the share of assessed conversations where it appears. A score is shown once there are at least 10 assessed conversations.

3. Combine the evidence

The score is a weighted average of those shares, scaled to 100. Iteration and refinement has a weight of 5.6. Each of the other 10 behaviors has a weight of 1.

See a worked scoring example

In 20 fictional conversations, refinement appears 15 times. Across the other ten behaviors there are 92 positive observations out of 200 possible.

100 × ((15 / 20 × 5.6) + (92 / 20)) / 15.6 = 56

This gives a developing score. It describes recorded behavior in that set of conversations, not an overall judgment of the person.

0–29

Low observed fluency

30–59

Developing

60–79

Strong

80–100

High

These are the product’s display bands, not validated pass or fail thresholds. Tasks differ. A useful behavior may not be needed in every conversation. Missing evidence appears as “Not enough assessed conversations”, never zero.

Across your business

Make development specific to the work.

People

See your score, the habits you have demonstrated and practical suggestions for your next task.

Teams

Compare like-for-like work and follow changes over time. Use session coverage and examples to decide where coaching could help.

Models

Keep model performance separate from personal competency. Compare the quality, cost and friction of models on the tasks your business actually does.

F.AI also includes a real-world task benchmark for comparing AI models. The personal score shown here is a different measure. Neither a higher score nor more usage proves a business outcome.

See F.AI alongside adoption and spend in Helm ↗
Practical questions

What to know before you start.

What does an F.AI score measure?

The personal score summarizes observable behaviors in assessed AI conversations. It looks at how someone describes a task, guides the interaction and checks the output. It is not a direct measure of employee performance.

Why might someone have no score?

The current score requires at least ten assessed conversations. Missing or unassessed activity does not count as a score of zero. Check coverage before interpreting differences between people or teams.

Does a higher score prove that AI saved money?

No. Use the score to guide development, then measure outcomes separately. Compare output quality, review effort, time and running cost on comparable tasks.