One assistant per number
The number someone messages already says which desk they reached: a market, a language, an audience. So each Cloud API number gets its own assistant, with its own name and its own instructions. That is what lets a Turkish line and a Malaysian line answer completely differently, and it is what lets you switch one on while the other stays off. Find them under Settings → Agents. Every Cloud API number is listed whether or not it has an assistant yet.The assistant runs on Cloud API numbers only, and the number has to be set to two-way. Personal numbers paired by QR are deliberately left alone: those are a rep’s own line for warm one-to-one conversations, which is the last place an automatic reply belongs. A number that isn’t two-way doesn’t keep incoming messages, so there would be nothing for the assistant to read — the row tells you so and won’t let you switch it on.
Off, Suggest, Live
Every assistant is in one of three states, set on its row or on its page.- Off — nothing happens. This is where every number starts.
- Suggest — when a contact writes, a draft reply appears in the conversation, visible only to your team, with Send and Discard. Nothing goes out on its own.
- Live — it replies by itself, about fifteen seconds after the contact stops writing.
Writing its instructions
Click the assistant’s name to open its page. Instructions is where you say how this desk answers: who it’s talking to, what it should offer, what it must never say, and when to get a person involved. Instructions do not save as you type. A half-written sentence would otherwise be live on the very next message a customer sends. The page says when you have unsaved work and offers Save or Discard changes. A short brief that is specific beats a long one that is vague. The most useful lines are usually the refusals — the figures it may not quote, the promises it may not make.What it knows
You don’t have to tell it who it’s talking to. Before it writes anything it already has:- The contact’s record — their name, and the details on their card, in plain wording rather than raw values.
- What your calls collected — the answers your voice agents captured, plus a short summary of recent calls. A chat can pick up where a phone call left off.
- The last ten messages in the conversation.
- What the business knows — see below.
Ten messages is short on purpose. The conversation is short-term memory; the contact’s record is the long-term memory. A person picking up a chat scrolls back a screen, not a year.
What the business knows
Under the instructions is a shared list of facts: opening hours, parking, refund policy — the answers you give twenty times a week and that don’t live in a field or a template. It is shared by every assistant, so write it once rather than into each one’s instructions. Don’t retype things Nudge already holds, like a contact’s own details: prose that repeats a record can contradict it, and the contradiction is invisible until someone is told the wrong thing. Write it in whatever language you think in. The assistant replies in the customer’s.Saved messages it can send
The assistant can send your saved messages — a price list, a booking link, a brochure — but only the ones you allow. On a WhatsApp template, tick The assistant can send this. You can add one plain line saying when it applies, like “send when they ask about opening times”. That is guidance, not a keyword rule: “what time do you open” and “are you open Sundays” both land on it. Nothing is offered by default, and that is the point. Most of a library is outreach copy, and one of those going out in the middle of a live conversation is exactly what this prevents.When it hands over
Sooner or later someone asks for something the assistant shouldn’t answer — an exact figure, a commitment, or simply a person. When it hands over decides what happens next.- Keep answering — it flags the conversation for your team, says so once, and carries on helping with everything else. Not being able to quote a price is no reason to stop answering the opening hours.
- Stop replying — it says a colleague will follow up and then goes quiet until someone hands it back. The right choice when half an answer is worse than none.
Changing it for one conversation
The setting on the number is a default. Any single conversation can disagree with it, from the strip above the message box: the same Off / Suggest / Live, for this chat only. That covers the ordinary cases — a delicate conversation you want to write yourself on an otherwise live number, or one straightforward chat you’re happy to let run while the rest are drafts. To put a conversation back to normal, pick whatever the number itself is set to. There’s no separate “back to default” button because there doesn’t need to be: a conversation you never touched follows the number, including when you change it later, and one you set by hand keeps your choice.Introducing itself
Introduce itself is off by default. Switch it on and the first reply a contact ever gets opens with a line naming the assistant as a digital one. Whatever you choose, it will never claim to be a person, and if a contact asks outright whether they’re talking to a bot it says so plainly and carries on helping. That part isn’t a setting.Some markets and some industries expect the disclosure. If you’re messaging people in the EU, or in a regulated field, turning it on is the safer read.
Trying it before anyone sees it
Try it on the assistant’s page talks to the real assistant with your saved instructions. Nothing is sent and no contact is involved. It answers “does this sound the way I want”, which is the question you ask twenty times while writing a brief. It does not answer “will this handle my customers” — there’s no contact record behind it. Suggest mode on real conversations answers that one.What it costs
Replies are charged by the AI provider you connected, per message, and it is small — fractions of a penny for a typical reply. Recent replies on the assistant’s page shows every turn with its cost, and the month’s total against your budget. The budget is workspace-wide, because the bill is. It’s checked before each reply, so the ceiling can be passed by at most one reply’s worth rather than found out about afterwards.What it won’t do
Worth knowing, because these are deliberate:- It never changes the contact’s record. If someone tells it something worth keeping, it acknowledges and carries on — your team reads the conversation and records it.
- It answers at most ten times an hour per contact. A loop can’t run away with your budget or your customer’s patience.
- It waits about fifteen seconds after the contact’s last message, so three messages typed in a row get one answer to all of them rather than three replies to a question nobody finished asking.
- It can’t see attachments. Sent a photo or a voice note, it says so and gets a person rather than guessing.
- If the AI provider fails, the contact gets nothing — never an error message from a system they didn’t know was there. The conversation just stays unread and your team answers it.
- Meta’s 24-hour rule still applies. The assistant answers people who have messaged you, so it is almost always inside the window — but a reply that lands outside it needs an approved template like any other. See WhatsApp.
Getting started
1
Connect an AI provider
Settings → Connections, then add an OpenAI or Anthropic key. Without one the assistant has nothing to think with.
2
Set the number to two-way
Settings → WhatsApp. Incoming messages have to be kept for the assistant to read them.
3
Name it and write the brief
Settings → Agents → Set up. Give it a name your customers would find natural, write its instructions, and save.
4
Put it on Suggest
Read the drafts for a week. Watch the share your team actually sends.
5
Go live on one number
When the drafts read the way you’d write them, switch that number to Live — one number, not all of them.