
Agentic AI · 2026
Chatbot vs AI agent — pick the product, not the label
The distinction that actually matters
A chatbot is a conversational surface. It interprets language and returns language. It may retrieve a document. It does not, by itself, have authority to change a record, send money, or write to a live system.
An agent pursues a goal across steps. It plans, calls tools, keeps enough state to continue, and stops when a rule says stop. The risk is not a wrong sentence. The risk is a wrong action.
We keep judgement in code wherever we can, and we let a language model phrase, classify, or draft. That boundary — deterministic where it must be trusted, generative where it adds value — is the same rule we used on our on-device daily brief.
Side by side
Planning table — not a vendor scorecard.
| Question | Chatbot | Agent |
|---|---|---|
| What does success look like? | The user got a correct answer. | A bounded workflow finished, or stopped cleanly for a human. |
| What can it touch? | A knowledge base, maybe a search index. | Named tools with scoped credentials — CRM, inbox, ledger, calendar. |
| What fails first? | Hallucinated copy. Embarrassing, usually reversible. | A write to production. That is an incident, not a typo. |
| What must exist on day one? | A source of truth and a fallback to a human. | Ownership, an audit trail, and a gate on irreversible steps. |
| When is it the right product? | FAQs, policy lookup, guided intake. | Multi-step work you can name, bound, and review. |
When we refuse the agent brief
If the brief cannot name the tools, the owner, and the actions that must never run unsupervised, it is not an agent project. It is a demo. We will say so in the first call.
If the only requirement is “ChatGPT on our website,” a retrieval chatbot with a human handoff is usually the honest product. Dressing it as an agent adds cost and risk without finishing more work.
If the workflow commits money, sends external communications, or updates a regulated record, those steps stay behind a human gate until the owner writes otherwise. Autonomy is a setting, not a personality.
A hybrid is often the actual product
Intake can be a chatbot. Execution can be an agent. Review can be a person. That split is how you get a usable interface without giving a model the keys to the estate.
The playbook on this site is the production version of that argument: why pilots die after a beautiful demo, and a 90-day path to one owned system.
FAQ
Can a better language model turn a chatbot into an agent?
No. A stronger model writes better sentences. An agent still needs tools, permissions, state, and a stop condition. Those are product and infrastructure decisions, not prompt upgrades.
Do we need an agent if we only want Arabic and English support?
Usually not. Bilingual retrieval is a chatbot or search problem. Build the language surface first. Add an agent only when there is a workflow to complete.
Is this legal advice?
No. If an agent will process personal data or sit in a regulated workflow, your counsel and DPO own the mapping. We design the architecture so that mapping is possible — named tools, logs, and human gates.
