nectr

FAQ

Questions

Scans

Why would AI crawlers miss part of my page?
Many of them fetch the HTML and don't run JavaScript. Text that a page adds with scripts, such as prices, menus or product details, isn't in what they get.
What does a scan do?
It fetches the page twice: as plain HTML, and in a headless browser that runs its JavaScript. It compares the words in the two and lists the parts that only appear with JavaScript. It also reads the site's robots.txt for the main AI crawlers, and checks whether up to three PDFs the page links to on the same site contain text.
Do I need an account?
No. Anyone can scan a public page, one at a time.
Can I scan a site that isn't mine?
Yes, one public page at a time, as long as the site's robots.txt lets nectr's crawler in.
How do I keep nectr off my site?
Disallow Nectrbot in your robots.txt. A scan reads it before fetching anything and stops if the page is disallowed. For anything else, get in touch.

Your data

Does nectr keep the pages I scan?
No. A scan's result goes back to you and isn't stored.
Does nectr train AI on content?
No. nectr doesn't train models or sell content.
Where does nectr run?
In Switzerland, including its backups.

Coming

When will Integrate, Publish and Monitor be ready?
There are no dates yet. Each one will be announced on the updates page when it opens.
Will they need a verified domain?
Yes. nectr will only serve copies or publish data for a site whose owner has verified the domain. Analytics will also need Integrate, because nectr only sees requests for the copies it serves.
What about llms.txt?
Integrate will generate one from your pages and keep it current. Few AI crawlers request llms.txt today, so the page copies matter more.

Another question? Get in touch.