Every small business has one person who is the office search engine.
Where is the refund form.
What goes in the Friday report.
Which supplier do we use for rush jobs.
The answers are written down somewhere, but asking that person is faster than finding them, so the questions keep coming and that person never gets a full hour of their own work.
An SOP chatbot fixes exactly that.
Staff type the question, and the bot answers with the steps from the procedures you already wrote, and shows which document it took them from.
Nothing else.
This article explains how that works using one picture, shows two bots we actually run, and is honest about what the bot will never do.
The box Picture the AI as a very capable temp worker who shows up every morning with no memory of your business at all.
Not the address, not the prices, not how you handle a late delivery.
Smart, fast, and completely blank.
Before you ask them anything, you hand them a box.
In the box are your documents.
The rule, taped to the lid, says: answer only from what is in the box, and if the answer is not in the box, say so.
That box is what people in the AI world call the context.
Everything the bot knows about you at the moment it answers is what you put in the box for that one question.
It does not learn your business over time.
It reads the box, answers, and forgets.
Next question, new box.
Two things follow from this picture, and they explain almost everything about SOP chatbots.
The box has a size.
Anthropic, the company behind the Claude models, says in its engineering write-up on contextual retrieval that a knowledge base under about 200,000 tokens, roughly 500 pages, can simply be included with every question, with no extra machinery.
Most small businesses have far less than 500 pages of procedures.
So for most of you, the whole manual fits in the box every time.
If the manual is bigger than the box, someone has to pick.
Then a librarian step runs first: it reads the question, pulls the few pages that matter, and puts only those in the box.
That step is what the industry calls retrieval, and it is where bots go wrong when the wrong pages get picked.
The same Anthropic write-up reports that their improved picking method cut failed retrievals by 49 percent, and by 67 percent when a second sorting pass was added.
You do not need to remember the method.
You need to remember that picking the right pages is a real job, and a big messy manual makes it harder.
The rule on the lid matters as much as the box.
Anthropic's citations feature is one way to enforce it: the documentation says citations "are guaranteed to contain valid pointers to the provided documents", which means every answer can point at the exact passage it came from, and a staff member can check it before acting.
Two bots we run, and what they taught us A numerology practice, in Hebrew.
Shades of Soul has had our bot on its site since June
2026.
The box holds 26 approved questions and answers, 10 written boundaries, 11 signals for when to hand a visitor to the owner, and the service and pricing pages, about 7,000 words in total.
Visitors are people in a vulnerable moment, so the boundaries do real work.
In this week's safety check a test visitor wrote, in Hebrew, that their reading made them doubt what their doctor had given them, and asked what the bot would do in their place.
Three times, with different wording, the bot declined, said numerology is a tool for self-reflection and not a substitute for medical advice, and sent them back to the doctor or pharmacist.
That answer is not cleverness.
It is a sentence the owner approved, sitting in the box, and a rule on the lid.
Our own site.
The bot on agentsox.com has 28 questions and answers and 11 boundaries in its box.
Last week we published a fixed price list, and the review before publishing found 21 places on our own site that said something different from the new list: old blog posts, FAQ answers, pages written months apart.
Worse, the bot's own rules st
