Agency €1.799/month, uncapped roster. €299 for a single brand.
Example project — illustrative numbers, not a customer result
NCNorthwind Coffeerun of 28 July · 4 models · 20 answers
Completed
Reputation score
3.8/5
Band Fair
+0.4
Score from 0 to 5 for each dimension, run of 28 July
Reputation
4.2
Reliability
3.6
Positioning
3.1
Service
4.5
Price
2.8
The setup
Checking by hand works for one client and stops there.
Four models, the same prompts pasted by hand, screenshots into a slide. It holds for one client: no two rounds are comparable, and neither are two clients.
A client asks in a status call and the honest answer is "we're looking into it". More ask every quarter.
It is manual end to end, so a junior cannot run it and come back with the same thing.
Nothing shows whether last quarter's work moved anything, so it stays invoiced on trust.
Four models, one afternoon, per client
The shift
Counting mentions leaves out what the model actually said about your client.
Every tool in this category reports the same four numbers, defined by the vendor and identical for every client you manage.
The people who test AI features for a living stopped counting occurrences years ago: they write down what a 1 and a 5 look like, and have a separate model grade against it.
What every tool in this category reports
Visibility — how often the brand appears
Position — where it ranks when it appears
Sentiment — on the vendor's scale
Share of voice — against competitors
What teams who test AI features do instead
A written rubric: what earns a 1, what earns a 5
A separate model grading, one answer at a time
Reasoning written before the score
The criteria kept with the result they produced
The solution
Sentl.AI has you write the grading scale, not just the questions.
One project per client brand. You name the dimensions, write the questions, and write the rubric: what earns a 1, what earns a 5, and the rule in between.
01
A project per client
One brand, the models you want queried, and how often the run repeats.
02
Dimensions you name
The areas you already report on. Each weighs the same in the final score.
03
Questions you write
The questions this client's buyers actually put to a model.
04
The rubric you write
What earns a 1, what earns a 5. The judge applies it unchanged to every answer.
Setup
Set a client up by talking to your assistant.
Connect Sentl.AI to your assistant over MCP and the setup happens in that conversation — project, dimensions, questions, rubric — in minutes rather than an afternoon. When a run finishes, the same assistant writes it up.
Which assistants. The ones that speak MCP: Claude, ChatGPT, Gemini, Perplexity, Grok. The list moves as the protocol spreads.
Nine tools, and three of them write
Reads projects, runs and results; with the right permissions, creates a project and starts a run.
OAuth, or a token you paste
Log in and consent, or paste a personal access token.
Scoped to one organization
One organization per connection, with permissions recomputed from your role on every request.
What you get
Three things the monthly report gets that a screenshot does not.
Your grading rubric, applied to every client
What the client pays for is your judgement, written once per client and applied the same way every month — then copied and adjusted for the next one.
"Another AI grading an AI."
Never the model that answered. It reads one answer at a time, does not browse, and writes its reasoning next to the score.
A number you can defend when the client disagrees
Push back on a 3 and you open the answer that earned it, and the reasoning that graded it. Improved criteria apply from the next run, so the trend line stays comparable.
"The models answer differently every time."
Which is why one screenshot proves nothing. Each run records what was said, with the questions and the rubric held still.
One subscription for the whole roster
A fixed cost divided across the clients it serves, reachable from wherever you already build the report.
"Can it carry our brand?"
No. What reaches the client is yours to produce; selling the measurement on to them is the reseller partner plan.
Where it stands
In beta, and the organisations running it cannot be named yet.
Large organisations in food, automotive, fashion and multi-utility are running it, and so are agencies. None can be named while the beta lasts, which is why there is no logo wall here.
What you can do instead
Run it on your own brand first: €299/month, one project, three questions and a rubric. What comes back is what you would put in front of a client.
Where the models treat this client well, whether it is moving, and behind every number the answer that produced it.
Example project — illustrative numbers, not a customer result
Score by dimension · run of 28 July
The radar above, read exactly. Score from 0 to 5.
Dimension
Score
Band
Reputation
4.2
Strong
Reliability
3.6
Fair
Positioning
3.1
Fair
Service
4.5
Strong
Price
2.8
Weak
Trend · last six runs
——— Reputation
- - - Reliability
The same series, read exactly. Full 0–5 axis, not truncated.
Dimension
Feb
Mar
Apr
May
Jun
Jul
Reputation
3.1
3.3
3.2
3.6
3.9
4.2
Reliability
2.8
2.9
3.2
3.1
3.4
3.6
Tone is read on each individual answer. Sentl.AI publishes no project-level sentiment percentage: that number would be comfortable and would mean nothing.
Pricing and credits
Three plans. What changes is how much you can measure.
Save up to 17% with an annual commitment
The plan sets the monthly credit allowance and how many projects stay active. Everything else is in all three.
A full run is four models answering ten questions, judge included: 680 credits.
In every plan
All AI models
MCP connection
API key
Team members
Credit history
Scheduled runs
Credits renew, they do not accumulate. Unspent credits do not carry over; extra packs need no plan change.
Changing plan. Moving up is immediate, moving down starts at the next renewal.
One subscription per organization. Most agencies keep the whole roster inside one.
Questions we get asked
My client can just ask ChatGPT themselves. What am I billing for?
They can ask once and get one answer, from one model, on one day. What you sell is the same questions put to every model on a schedule, graded against criteria you wrote, with a history that shows whether your work moved anything.
It costs more than the tools we have already looked at.
It does. The Agency plan is €1.799/month for an uncapped roster — about 44 runs a month, roughly €180 per client at ten clients and €90 at twenty. The Brand plan is €299/month if you want one brand first.
Where does the data live?
Everything runs on European servers. The models you query are third-party services running their own inference wherever they do — that part is theirs, not ours.
Can we resell this to our clients?
Yes, under the reseller partner plan. Terms depend on your roster, so they are agreed rather than published — write to reseller@sentl.ai.
How long does it take to set up the first client?
Minutes through your assistant: connect Sentl.AI over MCP and it walks you through project, dimensions, questions and rubric. The rubric takes thought rather than time, and you write it once per client.
What happens if we change the rubric halfway through the year?
From the next run. Every score keeps the rubric that produced it, so the history is not rewritten by a criterion you refined in June.
Which models does it query?
The ones you pick from the catalogue — GPT, Claude, Gemini, Llama, Mistral, Grok, DeepSeek, Qwen, Command and Perplexity among them. Change the selection later and completed runs stay as they are, so the comparison holds.
Set up one client and read the first score.
One dimension, three questions, a rubric. Start the run and read what four models said.