AI answer monitoring, for agencies

Your clients are asking what the AI models say about them.

Sentl.AI asks them your questions, for every client on your roster, and has an independent judge grade each answer 1 to 5 against your rubric.

Agency €1.799/month, uncapped roster. €299 for a single brand.

Example project — illustrative numbers, not a customer result

Northwind Coffee run of 28 July · 4 models · 20 answers
Completed

Reputation score

3.8/5

Band Fair +0.4

Score from 0 to 5 for each dimension, run of 28 July
Reputation4.2
Reliability3.6
Positioning3.1
Service4.5
Price2.8
The setup

Checking by hand works for one client and stops there.

Four models, the same prompts pasted by hand, screenshots into a slide. It holds for one client: no two rounds are comparable, and neither are two clients.

  • A client asks in a status call and the honest answer is "we're looking into it". More ask every quarter.
  • It is manual end to end, so a junior cannot run it and come back with the same thing.
  • Nothing shows whether last quarter's work moved anything, so it stays invoiced on trust.
Four models, one afternoon, per client
The shift

Counting mentions leaves out what the model actually said about your client.

Every tool in this category reports the same four numbers, defined by the vendor and identical for every client you manage.

The people who test AI features for a living stopped counting occurrences years ago: they write down what a 1 and a 5 look like, and have a separate model grade against it.

What every tool in this category reports

  • Visibility — how often the brand appears
  • Position — where it ranks when it appears
  • Sentiment — on the vendor's scale
  • Share of voice — against competitors

What teams who test AI features do instead

  • A written rubric: what earns a 1, what earns a 5
  • A separate model grading, one answer at a time
  • Reasoning written before the score
  • The criteria kept with the result they produced
The solution

Sentl.AI has you write the grading scale, not just the questions.

One project per client brand. You name the dimensions, write the questions, and write the rubric: what earns a 1, what earns a 5, and the rule in between.

  1. 01

    A project per client

    One brand, the models you want queried, and how often the run repeats.

  2. 02

    Dimensions you name

    The areas you already report on. Each weighs the same in the final score.

  3. 03

    Questions you write

    The questions this client's buyers actually put to a model.

  4. 04

    The rubric you write

    What earns a 1, what earns a 5. The judge applies it unchanged to every answer.

Setup

Set a client up by talking to your assistant.

Connect Sentl.AI to your assistant over MCP and the setup happens in that conversation — project, dimensions, questions, rubric — in minutes rather than an afternoon. When a run finishes, the same assistant writes it up.

Which assistants. The ones that speak MCP: Claude, ChatGPT, Gemini, Perplexity, Grok. The list moves as the protocol spreads.

  • Nine tools, and three of them write

    Reads projects, runs and results; with the right permissions, creates a project and starts a run.

  • OAuth, or a token you paste

    Log in and consent, or paste a personal access token.

  • Scoped to one organization

    One organization per connection, with permissions recomputed from your role on every request.

What you get

Three things the monthly report gets that a screenshot does not.

Your grading rubric, applied to every client

What the client pays for is your judgement, written once per client and applied the same way every month — then copied and adjusted for the next one.

"Another AI grading an AI." Never the model that answered. It reads one answer at a time, does not browse, and writes its reasoning next to the score.

A number you can defend when the client disagrees

Push back on a 3 and you open the answer that earned it, and the reasoning that graded it. Improved criteria apply from the next run, so the trend line stays comparable.

"The models answer differently every time." Which is why one screenshot proves nothing. Each run records what was said, with the questions and the rubric held still.

One subscription for the whole roster

A fixed cost divided across the clients it serves, reachable from wherever you already build the report.

"Can it carry our brand?" No. What reaches the client is yours to produce; selling the measurement on to them is the reseller partner plan.

Where it stands

In beta, and the organisations running it cannot be named yet.

Large organisations in food, automotive, fashion and multi-utility are running it, and so are agencies. None can be named while the beta lasts, which is why there is no logo wall here.

What you can do instead

Run it on your own brand first: €299/month, one project, three questions and a rubric. What comes back is what you would put in front of a client.

Sign up
Results

What comes back from a run.

Where the models treat this client well, whether it is moving, and behind every number the answer that produced it.

Example project — illustrative numbers, not a customer result

Score by dimension · run of 28 July
The radar above, read exactly. Score from 0 to 5.
DimensionScoreBand
Reputation4.2Strong
Reliability3.6Fair
Positioning3.1Fair
Service4.5Strong
Price2.8Weak
Trend · last six runs
  • Reputation
  • Reliability
The same series, read exactly. Full 0–5 axis, not truncated.
DimensionFebMarAprMayJunJul
Reputation3.13.33.23.63.94.2
Reliability2.82.93.23.13.43.6

Tone is read on each individual answer. Sentl.AI publishes no project-level sentiment percentage: that number would be comfortable and would mean nothing.

Pricing and credits

Three plans. What changes is how much you can measure.

Save up to 17% with an annual commitment

The plan sets the monthly credit allowance and how many projects stay active. Everything else is in all three.

Brand

One client brand, or your own.

€247 /month

Saving €624 a year

  • 4.200 credits a month
  • up to 3 active projects
  • about 6 full runs a month
Start with Brand

Team

Several brands or markets.

€497 /month

Saving €1.224 a year

  • 9.200 credits a month
  • up to 8 active projects
  • about 13 full runs a month
Start with Team

A full run is four models answering ten questions, judge included: 680 credits.

In every plan
  • All AI models
  • MCP connection
  • API key
  • Team members
  • Credit history
  • Scheduled runs
  • Credits renew, they do not accumulate. Unspent credits do not carry over; extra packs need no plan change.
  • Changing plan. Moving up is immediate, moving down starts at the next renewal.
  • One subscription per organization. Most agencies keep the whole roster inside one.

Questions we get asked

My client can just ask ChatGPT themselves. What am I billing for?

They can ask once and get one answer, from one model, on one day. What you sell is the same questions put to every model on a schedule, graded against criteria you wrote, with a history that shows whether your work moved anything.

It costs more than the tools we have already looked at.

It does. The Agency plan is €1.799/month for an uncapped roster — about 44 runs a month, roughly €180 per client at ten clients and €90 at twenty. The Brand plan is €299/month if you want one brand first.

Where does the data live?

Everything runs on European servers. The models you query are third-party services running their own inference wherever they do — that part is theirs, not ours.

Can we resell this to our clients?

Yes, under the reseller partner plan. Terms depend on your roster, so they are agreed rather than published — write to reseller@sentl.ai.

How long does it take to set up the first client?

Minutes through your assistant: connect Sentl.AI over MCP and it walks you through project, dimensions, questions and rubric. The rubric takes thought rather than time, and you write it once per client.

What happens if we change the rubric halfway through the year?

From the next run. Every score keeps the rubric that produced it, so the history is not rewritten by a criterion you refined in June.

Which models does it query?

The ones you pick from the catalogue — GPT, Claude, Gemini, Llama, Mistral, Grok, DeepSeek, Qwen, Command and Perplexity among them. Change the selection later and completed runs stay as they are, so the comparison holds.

Set up one client and read the first score.

One dimension, three questions, a rubric. Start the run and read what four models said.

Set up your first client

Agency €1.799/month, €299 for one brand · European servers · You decide when each run starts · Already have an account? Log in