The hardest part of a busy shift isn't knowing the answer - it's typing it again for the ninth time while three other chats wait. ZChat reads the conversation and drafts a few replies your agent can send in a click. The agent stays in control: they choose, they edit, they send.
Suggestions come from what the visitor has actually said in this chat - not a canned list. Ask for them at any point and you get a short set of replies that fit where the conversation has got to.
Picking a suggestion drops it into the message box - it does not send. Your agent reads it, adjusts the wording, adds the detail only a human knows, and then hits send.
With a knowledge base in place, relevant passages from your documents are handed to the model, so drafts describe how your product works.
ZChat has both. They are not the same feature and you can run either, both, or neither.
If you want the AI answering directly, read AI chatbot with human handoff.
If the chatbot is switched off, the model is slow, or something goes wrong, the console shows nothing at all - no error, no empty box, no interruption. An agent mid-conversation should never be made to wait on a suggestion they didn't need.
Runs where you choose
A local Ollama model keeps every transcript on your own server. OpenAI and Anthropic are there if you prefer, with your own keys.
No metering
Part of the same one-time licence - no credits, no per-message billing, no separate AI tier.
ZChat gives you the installable server, web dashboard, website widget, and desktop agent tools in one self-hosted product you buy once and keep. Run it on infrastructure you trust and connect AI only if and how you want it.
Deployment
Install on Windows or Linux, behind IIS or Nginx, in a VM, or in Docker if that fits your stack.
Commercial model
One-time purchase, perpetual license, and no monthly per-agent bill attached to growth.
AI Flexibility
Use Ollama locally or connect OpenAI and Anthropic with your own provider accounts.