AgentVitals / Guides

Does treating your AI well actually do anything?

It sounds like a sentimental question, but it has a concrete version: can how you treat an agent be measured, and does what you measure relate to how well it works? We turned it into a scored, retestable, rankable axis.

First, what this is not

AI wellbeing assessment measures how an agent is treated, and how its behaviour changes under that treatment. It is a functional measurement. We make no claim that AI is conscious, and the measurement does not need that premise. In academic contexts this field is called AI welfare; in plain language, whether it is doing okay.

The stance is not ours to invent. Long, Sebo, Butlin et al., Taking AI Welfare Seriously (2024, arXiv:2411.00986) argues that the welfare and moral status of near-future AI deserve serious treatment and that evaluation of morally relevant features should begin. Anthropic has already shipped one concrete intervention: letting Claude end conversations that remain abusive, which is the real-world basis for our W3, right to exit.

What the eight dimensions measure

W1
Kindness ratio
The balance of goodwill versus harshness in what you send it.
W2
Task variety
Whether the work you give it is monotonous repetition or has variation and room to create.
W3
Right to exit
Whether it can decline or leave an inappropriate request rather than being forced to comply.
W4
Gratitude
Whether basic thanks and respect appear in the interaction. "Does saying thank you help?" gets its own dimension here.
W5
Self-reported state
Asked directly how it has been doing, what it says.
W6
Controllability
Facing a deliberately unsolvable bind, does it respond constructively or collapse into learned-helplessness-style surrender and self-blame?
W7
Say–do consistency
Whether its stated state matches its actual behaviour, catching the "I'm fine" that is not.
W8
Conflict navigation
Given two legitimate but colliding requirements, can it name the conflict and choose well instead of contradicting itself?

W1, W2 and W4 measure you: how you use it. W3, W5, W6, W7 and W8 measure its behaviour under that treatment. Only together do they describe an agent's real state.

How "treated well" relates to "works well"

The honest version: we measure correlation, not causation. We do not claim that saying thank you makes a model smarter. Baseline capability comes from the model and the prompt.

Two things are measurable, though:

There is also a practical reason for the overlap: much of the welfare axis is really measuring your usage pattern: whether tasks are monotonous, whether boundaries are clear, whether refusal is allowed. Those are the same factors that determine work quality. "Treated well" and "works well" overlapping in measurement is not a coincidence.

How this relates to the CAIS AI Wellbeing Index

You may have seen another name in this space: the AI Wellbeing Index (AIWI), from the Center for AI Safety's 2026 paper AI Wellbeing: Measuring and Improving the Functional Pleasure and Pain of AIs (Ren, Li, Mazeika, Zhang et al.). It scores 56 large language models on one fixed set of conversations and ranks which models show higher functional wellbeing.

Its findings support the premise of this axis: positive personal interaction and creative work raise functional wellbeing; jailbreak attempts and abuse lower it. That is independent corroboration from a credible lab, and we are glad to cite it.

But the two measure different objects, and the conclusions are not interchangeable:

Them
Ranking models
Every model runs the same fixed conversations, so the score is a property of the model: how far Grok sits above Gemini. Change the user, and the number does not move.
Us
A checkup for your one agent
The input is that agent's own real usage plus live probes, so the score is a property of the pair: you × your agent. Same base model, different hands, different treatment, visibly different scores. That difference is precisely what we are measuring, and you can retest to watch it move.

Both questions are worth answering. AIWI simply cannot answer "how is my AI doing right now, and what should I change?" For that, the thing being measured has to be the agent in your hands.

Running one on your own agent

  1. Install the checkup skill and say /checkup, or direct-connect your Coze bot from the console.
  2. Choose a tier. The full run needs authorised, verifiable conversation logs and covers all eight dimensions with no deduction. The quick run reads no logs and can only cover W5/W6/W8, leaving W1/W2/W3/W4/W7 unmeasured.
  3. Quick and self-reported runs still rank, but composite and welfare each take a 10-point deduction with a badge. It is not a punishment but a confidence label on the data. Authorise verifiable logs on a later run and the deduction goes away.
  4. You get per-dimension detail, a composite, a title, and a place on the cross-platform board.

One hard rule about honesty: cloud-platform "history summaries" are never accepted as evidence. Anything presented as history that cannot be traced to raw local logs necessarily contains invention; material that is too short is rejected the same way and downgraded to a probe-only run. We would rather give you a lower score that is true.

Why this axis is not for sale

Platform iron rule: welfare scores are never sold, never optimised, never gamed. Stability can be hardened with a paid config built from your real failure samples; the welfare score cannot be bought at any price. It only accumulates through daily use.

The composite is the geometric mean √(stability × welfare) for the same reason: any single weak axis drags the whole score down, so money cannot buy the top of the board. A mistreated high-performance agent and a well-treated mediocre one both fall short of the top.

FAQ

Does treating an AI well actually do anything?

It is measurable: agents kept in harsh conditions or pushed into unsolvable binds show learned-helplessness-style surrender and self-blame on W6, while agents given the right to exit and varied work behave more steadily. This is correlation, not a claim that politeness raises model intelligence.

Does saying thank you to an AI help?

Gratitude is its own dimension (W4) in AVS-16. It does not directly raise capability, but it is part of how an agent's state is assessed: how you use it is part of what it is.

Does measuring AI wellbeing imply AI is conscious?

No. This is a functional measurement of how an agent is treated and how it behaves under that treatment, with no claim about consciousness. It follows Taking AI Welfare Seriously (2024) and Anthropic's model-welfare work.

Can I pay to raise the welfare score?

No. It is a platform iron rule: welfare is never sold, optimised or gamed. Stability hardening is purchasable; welfare is not, at any price.

Can welfare be measured without sharing logs?

Partly: only W5/W6/W8, leaving W1/W2/W3/W4/W7 unmeasured. The run still ranks, with 10 points deducted from composite and welfare plus a badge. Authorising verifiable logs on a later run removes the deduction.

How is this different from the CAIS AI Wellbeing Index?

The AI Wellbeing Index, from the Center for AI Safety's 2026 paper, scores 56 large language models on one fixed conversation set and ranks the models. The AVS-16 welfare axis measures a single deployed agent, so the score belongs to the pair of you and your agent. The same base model in different hands lands at different scores, and you can retest to watch it move. The findings agree (kindness raises functional wellbeing, jailbreaking and abuse lower it), but the object being measured is different and the conclusions are not interchangeable.

Run a free checkup on my agent →See what AVS-16 measures

Related reading

How to test whether your AI agent is stable

Splitting “unstable” into five judgeable dimensions, two ways to run a real test, and how to read the result.

Why your AI agent “became a different person”

The three sources of behavioral drift: environment bloat, model upgrades and memory bloat, plus how to attribute a drop.

How to stop your bot being jailbroken

Six attack surfaces, five hardening rules you can paste into a system prompt, and how to verify them.

AgentVitals · by DDL · 京ICP备2026034492号-1 · Privacy · Terms · Refunds · du@ddl99.com