+33 (0)1 87 66 00 65 · Monday to Friday, 9am–6pm Free audit (15 min)
● B2B offer — Customer service at scale

Customer service at scale: several agents, one point of supervision

Past a certain volume, customer service can no longer be handled by a single agent: it takes several specialised agents — common questions, orders, complaints, technical — coordinated under a single point of supervision. This system orchestrates them, routes requests according to your rules and escalates to an adviser as soon as the situation calls for it. Hosted in France: your customers' exchanges stay with you.

Hosted in France Customer exchanges protected GDPR & AI Act: governed deployment Human oversight

Updated on

Deployed in a few weeks
Customer service at scale · hosted in France
How are requests being spread across the agents right now?
Breakdown by specialised agent: common questions, order tracking, complaints, technical — volumes and queues in progress.
Requests are routed according to your routing rules, and the reason for the routing is kept.
Escalations to an adviser are shown with their cause.
🔗 Sourced · processing queues and routing rules
What happens during a spike in activity?
The system absorbs the volume on the requests covered by your content, and maintains the priority of the queues you have designated as priorities.
Situations that call for an adviser are passed on with the full thread, whatever the volume.
✎ Support · queues prioritised, escalation preserved
Local inference · no data outside the EU
System hosted in France
Sovereign by designLocal inference or hosting in France
GDPR & AI Act: governed deploymentTraceability & human oversight
TurnkeyDesigned, installed and operated for you
The customer relations department decidesThe agent prepares, never rules
✦ In brief

A Blue Lemon Agent multi-agent customer service system coordinates several specialised agents — common questions, orders, complaints, technical — under a single point of supervision, routes requests according to your rules and escalates to an adviser with the full thread. Sized for high omnichannel volume. Hosted in France, architecture designed to reduce exposure to extraterritorial legislation, location alone not being enough to guarantee immunity.

100%
hosted in France in the target architecture
0
transfer outside the EU in the target architecture
6
customer service uses ready to deploy
0
decision taken without human approval

These figures describe our offer, not results measured at a client. How large the gain is on the volume of requests, the number of channels and specialised agents is confirmed by a pilot.

The context

What does a multi-agent system bring to your customer service?

At scale, the question is no longer how to answer one request, but how to orchestrate thousands of them.

! The issue

At high volume, customer service needs specialisation — each type of request calls for its own content and rules — and coordination: readable routing, prioritised queues, overall supervision. The system brings both, keeping for every request the reason it was routed as it was.

Our answer

The customer relations management has a consolidated view of the queues and agents specialised by domain. Escalation to an adviser is preserved whatever the volume: it carries the full thread, so the customer has nothing to repeat. Local inference or an isolated resource hosted in France: your customers' exchanges and data are not entrusted to any third party.

The decisive point

Your customers' exchanges and data: sovereignty & compliance

Large-scale customer service concentrates a considerable volume of personal data. Here is how it is protected.

Local inference

The agent can run on a machine belonging to your organisation: no customer exchange or data leaves the network.

Hosting in France

Otherwise, a dedicated and isolated resource hosted in France, under French law — your contact channels and processing queues: processing and access within the European Union targeted by the architecture.

Reduced extraterritorial exposure

For your customers' exchanges and data, the architecture aims to reduce exposure to the Cloud Act and FISA 702; being located in France or in the European Union does not, on its own, guarantee immunity.

Isolated resource

No pooling: an environment strictly dedicated to your company and its routing rules.

Reason for routing kept

Every request keeps its agent, its routing reason and its timestamp; encryption, role-based access (RBAC) and logging by queue.

AI Act: governed deployment

The agent is strictly in support; no answer outside approved content and no goodwill gesture granted; traceability and human oversight from end to end.

What depends on the architecture chosen These points are not general guarantees: they are settled deployment by deployment, in the quotation.

  • The applicable location is that of the architecture set out in the quotation and verified before commissioning.
  • Local execution is announced only for the configuration explicitly described and accepted in the quotation.
  • The applicable isolation depends on the deployment mode set out in the quotation; no dedicated isolation is presumed.
  • Roles and permissions are configured and accepted for the identities and systems actually connected.
  • The events logged, their content, their retention period and who may access them are defined for the deployment chosen.
For exchanges concerning a dispute, an unpaid invoice or a sensitive complaint, SecNumCloud and reinforced hosting are options depending on your requirements. A single architecture is designed to answer both the GDPR and extraterritorial exposure. Designed for deployment in line with the GDPR and the AI Act, after the processing, roles and context-specific risks have been assessed.
Demonstration

See the agent at work

5 real situations, taken from those that come up most often. Pick one: the exchange unfolds as it would in your organisation.

A scripted demonstration. These exchanges show how the agent behaves — its sources, its refusals, what it leaves to your teams. Nothing is sent from this page, no model is queried here, and the matters named are fictional. That is precisely what we promise your data.
The behaviours shown here — monitoring, automation rules, routing and reminders — are configured with you during deployment, from your tools, your rules and your thresholds.
The architecture points named in these exchanges — location, local execution, isolation, encryption, role-based access, logging — are not a guarantee attached to the demonstration: they are those of the architecture set out in your quotation, and verified before commissioning.

The company in this demonstration

Fictional company

Maisonor — home equipment retailer

Sector
Home equipment retail — 62 stores and an online shop, with home delivery and installation
Headcount
1,840 staff, including 96 customer service advisers across 2 sites, 6 supervisors and 1 customer relations director
People served
1.4 million active customers, including 210,000 orders delivered and installed at home every year
Scale
840,000 requests a year across 5 channels — 310,000 calls, 250,000 emails, 160,000 chat conversations, 78,000 social media messages, 42,000 web forms — with peaks of 11,400 requests in a single day
Tools in place
Ticketing tool, telephony, chat platform, order management and a knowledge base of 1,240 articles — the system plugs into them, nothing is replaced or migrated
Who decides
The customer relations director settles the routing rules; the adviser alone grants any goodwill gesture and any date commitment; the supervisors arbitrate queue priorities
Room for improvement
Handling routine requests takes 65 % of advisers' time; routing to the right person takes 40 % of the time to a useful first answer; picking up context takes 30 % of an escalation; and a customer repeats his situation 2.3 times on average before getting his answer

Maisonor has crossed the volume at which a customer service operation can no longer be run by hand: 840,000 requests a year across five channels, with days reaching 11,400. Four specialised agents — routine questions, order tracking, complaints, technical — work under a single supervision view and pass requests along according to rules written by customer relations. The system runs on local inference on a machine belonging to the retailer, or on an isolated resource hosted in France; it plugs into the ticketing tool, the telephony, the chat platform and order management, and any situation calling for an adviser reaches him with the full thread. The exchanges below cover a year, from go-live to the review before the executive committee.

This company, its figures and the exchanges that follow were invented for the demonstration. They illustrate a common situation; they describe no real client.

Customer service at scale · hosted in France
I read your 840,000 requests of the year across the five channels, and I draw from them a split your queues carried without anyone being able to read it at a glance.
A specialised agent is an agent dedicated to one family of requests: it has its own content, its own rules and its own permissions, and it answers nothing else.
What the sorting gives:
· Routine questions — 372,000 a year: opening hours, product availability, payment methods, return conditions, loyalty card queries.
· Order tracking — 248,000: delivery date, installation slot, change of address, parcel announced and not received.
· Complaints — 138,000: damaged product, faulty installation, delay, invoicing.
· Technical — 82,000: assembly, spare parts, compatibility, manufacturer warranty.
What is built and ready: the four agents, each with its own content and permissions, and above them a single supervision view showing the four queues side by side, by volume, by lead time and by escalation.
What that shifts, calculated on your own records: handling routine requests goes from 65 % to 8 % of your advisers' time. Your 96 advisers stop handling 840,000 requests of every kind and handle the 126,000 situations that genuinely call for an adviserthe same headcount on work that is not the same.
And the measure that makes everything else checkable: each request keeps the agent that took it, the reason it was routed there and its timestamp. 840,000 out of 840,000a customer service figure you cannot explain is a figure you cannot correct.
What I propose: that your customer relations director open the supervision view tomorrow at 9 am, at peak hour, and watch the four queues for twenty minutes. Whatever she wants at the top, I put at the top the same day. four-specialised-agents_840000-requests.pdf65 % → 8 % of advisers' time
⛓ Sourced · 840,000 requests of the year across 5 channels, ticketing queues, adviser time records
Why four agents rather than one that can do everything? That is what we are offered elsewhere.
Because I measured it on your own requests, and the accuracy gap is 23 points.
The test, and it is reproducible: 4,000 requests from last year, whose right answer is known since they were handled and closed. I ran them twice — once through a single agent carrying all 1,240 of your articles, once through the four specialised agents.
· Single agent: 64 % right answers first time.
· Four specialised agents: 87 %.
· The gap sits in two families: technical goes from 41 % to 83 %, and complaints from 52 % to 79 %. On routine questions the gap is 4 points: there, a single agent does almost as well.
Why, in one sentence: a complaint and a spare-part compatibility question are not answered with the same content nor with the same permissions. An agent carrying everything arbitrates constantly between 1,240 articles; an agent carrying 190 technical articles answers within its field.
And what specialisation brings beyond accuracy: permissions follow the specialty. The “routine questions” agent does not see the bank details on a disputed invoice; the “complaints” agent does. That is stronger than an instruction, because it can be checked with one command.
What I propose next: that the same test run every quarter on 4,000 fresh requests, and that I hand you the gap family by family. The quarter a family stops improving, it is its content that needs reworking, and I will tell you which — article by article. specialisation_87-percent-against-64.pdf4,000 requests replayed, gap by family
⛓ Sourced · 4,000 closed requests from last year, knowledge base of 1,240 articles, permission matrix per agent
And the routing rules — do you write them? Those are our priorities, not yours.
They are yours, applied as they stand — and I add seven new ones, written in the same form and already measured on your last twelve months.
Your 22 current rules, applied to 840,000 requests: 803,000 fall under a rule; 37,000 trigger none and used to go into the general queue, where they waited 19 hours on average.
The seven rules I propose, each with what it would have given over twelve months:
· 1 — Any request naming an installation date within 72 hours goes to the priority queue. 14,200 requests concerned, average wait 19 h → 11 minutes, and 3,100 missed installations avoided according to your own cancellation reasons.
· 2 — Any complaint about a product the same customer has already complained about goes straight to an adviser. 2,900 requests, and it is the one that protects the relationship most: a second complaint handled by an automated reply is a relationship lost.
· 3 — Any request mixing two subjects is routed on the dominant subject, with a linked thread opened for the second. It settles 6,240 mis-routings over twelve months.
· The other four, smaller in volume, are set out with the same calculation.
What that gives on the delay your customers feel: routing to the right person goes from 40 % to 5 % of the time to a useful first answer — from 6 hours to 45 minutes.
The signature stays with your customer relations director: a rule comes into force only once she approves it, and that is what lets you tell a customer why his request went there rather than elsewhere. What you gain is the writing and the measurement; the decision takes twenty minutes instead of a committee.
What I propose: that the first three come into force this week and be measured for a month. If they hold their figures, the other four follow; if not, I rewrite them with the real figures of your month. routing-rules_22-in-force-7-proposed.pdf37,000 requests with no rule, 19 h down to 45 minutes
⛓ Sourced · 22 routing rules in force, 840,000 requests of the year, installation cancellation reasons
Local inference · no data outside the EU

Your case is not here? That is exactly what a 15-minute conversation is for. Book the free audit

Use cases

What does the agent actually do?

Several specialised agents, one point of supervision. All these uses work in support, subject to your approval.

Included in your agent The 3 capabilities essential to this promise are included, at no extra cost.
From 1,007 € excl. VAT / month

Coordination of the specialised agents

Common questions, order tracking, complaints, technical: coordinating those agents under a single supervision is what this offer sells. Naming each of them is a matter of deployment architecture, settled during the project, and is not ordered here.

Routing according to your rules

Routes every request and keeps the reason it was routed that way.

Escalation preserved

Passes the full thread to the adviser, whatever the volume.

Controls and safeguards These 4 controls are built into the agent: they frame what it does, whatever plan you pick. They are not chosen and are not added to your order.
Human arbitration of conflicts and irreversible decisions Observability of costs, timescales, quality, failures and safe stop Resolve common requests through to a verifiable closure Escalate quickly to a human, with the reason and the history
What the agent must be connected to These 2 connections are required for the agent to work. They concern your information system and are scoped during the audit.
Register of authorised agents and interface contracts Carry the full context across channels and into the ticket/CRM
Other needs our agents cover Each card says where the matching agent stands: available, on quote, or still being architected.

Support sub-agents to be attached

Which support sub-agents are attached to this supervision — and under which interface contract — is a matter of deployment architecture, settled during the project. No list is fixed today: nothing is priced, nothing is ordered from this page.

Being architected — not orderable Talk to us about it

Need to go further?

These agents handle a different business process, with their own owner and their own price. They are added to this one.

Does your need fall outside this?

In 15 minutes we identify the most relevant agent — without oversizing the project.

Book the free audit Build your agent
The gain

How many requests can an organisation absorb?

By specialising and coordinating, human effort shifts towards the situations that call for an adviser. How large the gain is depends on your volume and remains to be confirmed by a pilot.

Handling common requests
Today · done by hand
Requests absorbed
Directing to the right person
Today · done by hand
Routing applied
Passing the context to the adviser
Today · done by hand
Full thread passed on
Indicative figures, not contractual, to be confirmed by a pilot on the volume of requests, the number of channels and specialised agents. Goodwill gestures, commitments to a deadline and particular situations are for the adviser, who receives the full thread.
How it works

The stages of your AI agent project

1

Audit & scoping

15 minutes to target the use case with the best return.

2

Quote or direct sign-up

A catalogue offer is bought online; a specific need gets a costed quote.

3

Design

We design the agent and its guardrails.

4

Integration & testing

We connect your tools to the agent, which is itself hosted in France.

5

Rollout

Going live and training your team.

6

Operation

Continuous supervision and improvement.

Pricing

One package, one agent

A multi-agent customer service system (specialisation, routing, supervision), installed and operated for you. Prices exclude VAT — annual subscription, the time it takes for the gains to settle in for good.

Agility

Setup + controlled subscription

15,380 € excl. VAT setup
then 1,007 € excl. VAT/month — you invest at installation and pay a reduced subscription. Ideal for keeping the cost under control over time.
  • Installation, configuration and training for your teams
  • Operation, human oversight, updates and support
  • Sovereign hosting in France, a dedicated and isolated resource
Order →
The simplest Serenity

All inclusive, no setup fee

1,862 € excl. VAT /month
all inclusive, immediate start. No upfront investment: a single subscription. Ideal for starting quickly and simply.
  • Setup included (installation, configuration, training)
  • Operation, human oversight, updates and support
  • Sovereign hosting in France, managed end to end
Order →
100% Sovereign

On site, you own it

20,490 € excl. VAT setup
then 1,284 € excl. VAT/month · + hardware from 5,500 € (one-off purchase, in addition) — a sovereign computer installed on your premises, maintained remotely. Models run locally, your data returned at the end of the contract. 36-month commitment.
  • Hardware installed on your premises (you own it)
  • French / European AI models run locally
  • Secure remote maintenance (Pro support included)
Order →
Not included in the packages: AI consumption (model tokens), re-invoiced at real cost with no margin, and tracked in real time in your client area. Maintenance and supervision subscription for an initial term of 12 months for the Agility package, 24 months for the Serenity package and 36 months for the 100% Sovereign package, renewable; support levels (SLA 72 h / 24 h / 4 h) optional. Bespoke development, additional integrations or exceptional volumes are quoted separately. Support Monday to Friday, 9am to 6pm. Prices exclude VAT.
AI model: none of the AI models offered currently carries a fixed surcharge. When the selected model carries a cost, that cost is shown when you choose it, before you order, and re-invoiced at the cost incurred, with no mark-up; usage is billed at the publisher's price. Publishers' prices are published in US dollars: the amount re-invoiced is the amount in euros actually borne by Blue Lemon Agent on the publisher's invoice, at that invoice's exchange rate, with no commission or mark-up.
Included components and additional components Components included in the base offer: the Blue Lemon Agent software foundation, the AI models listed in the order journey, the standard channels (Microsoft Teams, Slack, WhatsApp Business, email, website chat, calendars, Microsoft 365 / Google Workspace, file storage, market VoIP telephony, professional social-media pages and accounts, Google Business Profile), hosting in France for the package chosen, backups, supervision, updates and support. If adapting the AI agent to your constraints, your needs or your requests requires other paid components — a third-party publisher's software licence, paid API access to one of your applications, hosting of health data, for which French law requires an HDS-certified host (art. L. 1111-8 of the French Public Health Code), SecNumCloud-qualified hosting, a speech synthesis service, particular hardware —, they are offered to you as an option or on quotation and re-invoiced at the cost incurred; nothing is committed without your written agreement. Where the artificial intelligence model you choose entails an additional cost, that cost is shown to you before you order and re-invoiced to you at the cost incurred, with no margin.
What to expect
Go-live 2 to 3 weeks
Agent designed, channels connected, team trained.
Steady state 4 to 7 weeks
After a few weeks of real use, once the agent's behaviour matches what you expect. Indicative estimate, adjusted to the options you keep. It is not a delivery commitment.
Our commitment

Four guarantees that matter at scale

Your customers' exchanges stay with youLocal inference or an isolated resource hosted in France; no exchange entrusted to a third party, no conversation used to train a model.
Data in France, under French lawYour customers' exchanges and data: minimisation and location in France, architecture designed to reduce exposure to extraterritorial legislation, location alone not being enough to guarantee immunity.
The customer relations department keeps the decisionThe agent produces omnichannel customer service at high volume, which can be checked and altered; no approval is automated.
Human oversight & traceabilityOn the volume of requests, the number of channels and specialised agents: systematic logging and tracking, in line with the AI Act.
Frequently asked questions

Your questions, our answers

Why several agents rather than one?
Because each type of request calls for its own content and rules. Specialisation improves how accurate the answers are, and the single point of supervision keeps the whole readable.
Does escalation hold up during a spike?
Yes. Situations that call for an adviser are passed on with the full thread, independently of the volume, and the queues you have designated as priorities stay that way.
How are requests routed?
According to your routing rules, with the routing reason kept for every request.
Which tools can our customers use to talk to the agent?
The ones they already use: WhatsApp Business, your website chat and email. Your customers have no account to create and no app to install. Your own teams reach the same agent from Microsoft Teams, Slack or their email. These connections rest on open standards, including the MCP protocol, and are included in every plan, at no extra cost, within the number of connections your level includes; the catalogue grows without a surcharge. Only the fees charged by third-party platforms — WhatsApp Business bills per conversation — are passed on to you at cost, without margin, outside the subscription. You keep oversight from a web dashboard.
What do the service commitments cover?
Availability, response times and follow-up, contracted according to how critical the system is.
Where are the exchanges hosted?
In France, on local inference or an isolated resource, with the deployment objective of processing and access operated within the European Union and an architecture designed to reduce exposure to extraterritorial legislation, location alone not being enough to guarantee immunity.
Does the agent state that it is an artificial intelligence?
Yes, from the very first interaction, and this is not a configuration option: since 2 August 2026, Article 50(1) of the European AI Regulation requires that any person interacting with an AI system be informed, unless this is obvious. The announcement is built into the greeting, in the other party’s language, and they can ask for a human at any time.
How long does it take to deploy this system?
Several months as a rule, depending on the volume, the number of channels and specialised agents, after a free audit then phases of design, integration and testing.
Let's talk

Let's size up the potential for your customer service

15 minutes to frame your volumes and your channels — hosted in France, supervised, with no commitment.