Skip to content
Mirai Minds

WhatsApp AI agents for sales and support

In short

Mirai Minds builds WhatsApp AI agents that hold long conversations with customers on the WhatsApp Business Platform. They run on Gemini, remember people across weeks with Mem0 and rolling summaries, answer from your documents and serve several business numbers from one system. We also build team inboxes so staff can read any conversation and reply.

Problems we take on, and the systems we ship for them

Long conversations that stay coherent

ProblemCustomers come back days later and the bot has forgotten them, or the prompt has grown too long to afford.

SystemEvery 20 messages are folded into a rolling summary of at most 150 words. The agent sees that summary, the last 20 messages and relevant long-term memories from Mem0.

ResultContext stays small and relevant in conversations that run for weeks.

Answers from your documents

ProblemAnswers must match current price lists, policies or product sheets.

SystemDocuments are chunked and indexed in a knowledge service shared with our voice agents, and the agent retrieves passages before it answers.

ResultOne knowledge base serves both phone and WhatsApp agents.

A whole signup inside one chat

ProblemFounders drop off long signup forms.

SystemFor iKoMatch, the founder sends a pitch deck in chat, the agent reads it, a voice call fills the gaps and the profile publishes itself. A 45-day campaign then delivers a daily slate of matched investors in the same thread.

ResultSignup, onboarding and daily delivery happen in one WhatsApp conversation, with no forms.

People step in when it matters

ProblemSome conversations need a person, and WhatsApp limits what a business may send after 24 hours of silence.

SystemA team inbox shows every conversation with media and read receipts. Staff reply in the thread; once the 24-hour window closes, the composer locks and only approved templates can go out.

ResultStaff can take over a conversation without breaking WhatsApp's messaging rules.

Replies that read like texting

ProblemA customer sends three short messages in a row and gets three separate bot replies.

SystemThe agent waits 3 seconds after the last message, answers once, and splits long replies into short bubbles with natural pauses.

ResultConversations read like a person typing, not a form.

How the pieces fit together

Reference architecture for WhatsApp AI agentsINPUTCustomer messagevia Meta webhookSYSTEMAccount routing and3 s debounceHUMAN REVIEWTeam inbox forstaff repliesSYSTEMSummary, Mem0memories, last 20messagesMODELGemini with toolsand documentsOUTPUTReply on WhatsApp

How it flows

  1. 01 Customer message via Meta webhook → Account routing and 3 s debounce
  2. 02 Customer message via Meta webhook → Team inbox for staff replies
  3. 03 Account routing and 3 s debounce → Summary, Mem0 memories, last 20 messages
  4. 04 Summary, Mem0 memories, last 20 messages → Gemini with tools and documents
  5. 05 Gemini with tools and documents → Reply on WhatsApp
  6. 06 Team inbox for staff replies → Reply on WhatsApp

Published

How is a WhatsApp agent different from a website chatbot?

A website chat lasts minutes. A WhatsApp conversation can run for weeks, with photos, voice notes and documents in between, and the customer expects you to remember. WhatsApp also has rules a website doesn't: replies are free-form only within 24 hours of the customer's last message, and after that only templates Meta has approved can go out. An agent that ignores those rules has its messages rejected.

Mirai Minds builds WhatsApp agents that are designed around this. They run on our WhatsApp Agents service, which already handles long-running conversations for iKoMatch.

How does it remember without a huge prompt?

Sending the whole history to the model on every message gets slow and expensive, and it gets worse as the conversation grows. Instead, every 20 messages are folded into a rolling summary of at most 150 words. The agent sees that summary, the last 20 messages and the long-term memories Mem0 finds relevant to the current message. A founder who mentioned their funding stage a month ago doesn't have to repeat it.

The agent also waits 3 seconds after a customer's last message before it answers, so three quick messages get one reply.

How do people stay in control?

Three ways. First, staff can read every conversation in a team inbox and reply in the thread; once the 24-hour window closes, the inbox only lets them send approved templates. Second, the prompt says which topics the agent must hand to a person, and we test those cases before launch. Third, templates and broadcasts go through your approval and Meta's before any customer sees them.

What does a build involve?

Setting up the WhatsApp Business account and webhook, writing and approving templates, loading your documents into the knowledge service, and writing the prompt with the cases where a person takes over. Then we test on real past messages, launch with staff watching the inbox, and tune from what customers actually send. The same knowledge base can serve a voice agent, so phone and chat give the same answers.

Where we've built this

What we usually build it with

Chosen per project. We'll tell you when something simpler will do.

Model
Gemini (google-genai)
Memory
Mem0Rolling summaries in PostgreSQL
Messaging
WhatsApp Business Platform (Cloud API)Approved message templates
Knowledge
Shared retrieval serviceChunked document collections
Backend
FastAPIPostgreSQLSQLAlchemy (async)AlembicAmazon S3 for media
Documents
unoserver for deck and document conversion

How the engagement runs

  1. Week 01

    Discovery call

    Thirty minutes with an engineer. You describe the job; we say whether AI is the right tool and what it would take.

  2. Week 12

    Scope and evals

    We agree what done looks like and build a test set from your real data before writing the system.

  3. Weeks 2–63

    Build

    Working software every week, measured against the test set, with your team trying it early.

  4. Launch4

    With human review

    The system goes live with a person checking the risky steps and a clear way to reach a human.

  5. Ongoing5

    Run and improve

    We watch real traffic, fix what breaks and tune on real cases, or hand over with documentation.

Asked on the first call

Do we need the WhatsApp Business Platform?

Yes. The agent sends and receives through Meta's WhatsApp Business Platform with a verified business number. We help with the setup, message templates and webhook; Meta approves templates and sets the rules for when a business may message a customer.

What happens after 24 hours without a reply from the customer?

WhatsApp only allows approved template messages once 24 hours have passed since the customer's last message. Our agents and team inboxes respect that: free-form replies stop, and follow-ups go out as templates you have approved.

How does the agent remember a customer?

Three layers: the last 20 messages, a rolling summary of everything older (rewritten every 20 messages, 150 words at most), and long-term facts in Mem0, such as a preference or a company's stage. The agent only pulls in memories relevant to the current message.

Can a person take over a conversation?

Yes. Staff read every conversation in a team inbox, with media and read receipts, and reply in the same thread. We agree the rules with you before launch: which topics go to a person and how the agent should behave once someone has replied.

Can one system run several WhatsApp numbers?

Yes. Each number is its own account with its own credentials and prompt, picked by the phone number ID on each incoming message. Account owners can also bring their own Gemini key and model.

Which languages does it handle?

Gemini reads and writes most major languages, including Hindi and mixed Hindi and English. We test prompts on real messages from your customers before launch, because tone matters as much as grammar on WhatsApp.

Other services

Have a system in mind? Let's scope it.

A 30-minute call with an engineer who has shipped this before. You leave with a plan, a rough timeline and what it would take — whether or not we build it.