About our crawler

Operated by MA EdTech Solutions Inc. · crawler@educationall.tech · last updated 2026-08-02

If you found AIWikiBot in your server logs, this page is for you.

The short version

We fetch publicly available pages from websites of organizations that serve children and families, to build a directory of those services. We identify ourselves honestly, we obey robots.txt, we fetch slowly, and we stop when asked. If you want us to stop, email crawler@educationall.tech and we will — permanently. You do not need to explain why.

Who we are

AIWikiBot is operated by MA EdTech Solutions Inc. Our contact address is crawler@educationall.tech, and it is read by a person.

Our User-Agent is:


AIWikiBot/0.1 (+https://bot.educationall.tech; crawler@educationall.tech)

We do not disguise this. We never present ourselves as a browser, and we do not use any technique to make our requests harder to identify. If you block us, we stay blocked.

What we are building

A directory of organizations that serve children and families in Waterloo Region, Ontario — paediatric occupational therapy, speech-language pathology, clinics, community centres and similar services — so that families and the professionals who support them can find them by what they need rather than by guessing a search term.

We record factual, publicly published details: an organization's name, address, phone number, website, and the services it says it offers. Every one of those facts is stored with a link to the exact page we read it from and the date we read it, so anything we publish can be traced back to your own words.

What we do not do

These are constraints in our software, not intentions.

How often we visit, and how gently

robots.txt

We follow the Robots Exclusion Protocol as specified in RFC 9309.

To block us specifically, add this to your robots.txt:


User-agent: AIWikiBot
Disallow: /

That takes effect on our next check, within 24 hours, and needs no email.

One detail worth stating, because implementations differ: if your robots.txt returns a server error, we treat it as "do not crawl" and stop — not as permission. A deploy that briefly breaks that file will pause us rather than unleash us.

How to make us stop

Any of these works, and none of them requires a reason:

  1. Add the robots.txt lines above.
  2. Email crawler@educationall.tech. We will stop within two business days and add your site to a permanent exclusion list.
  3. Block our requests at your server or firewall. We will detect it, record it, and stop scheduling your site.

Opt-out is permanent. We do not re-approach a site that has asked us not to crawl it.

If your organization is listed

Our listings are built from published sources and may be wrong or out of date.

You do not need an account with us to do any of this, and we will not ask you for one.

Accountability

A person is responsible for this crawler's behaviour. If it does something it should not — hits your site too hard, ignores your robots.txt, or records something it should not have — email crawler@educationall.tech and it will be stopped. We would rather hear from you than not.