Skip to content

Journal

What Oernoe Search crawls, and what it does not

Search engines are not magic. They fetch pages, store what they can parse, and rank what they stored. This is how that works here.

Updated 24 August 2026 · Oernoe Editorial Team

Most explainers of “how search engines work” stop at crawl, index, rank. That outline is true and also too thin to be useful. The interesting part is what a given engine chooses to keep, and what it refuses to turn into a product.

Oernoe Search lives at search.oernoe.com. This journal lives at www.oernoe.com. They share an operator, Anoepal, and they do not share an advertising file of your queries. If you only remember one split from this page, remember that one.

Crawling is just fetching

A crawler starts from URLs it already knows: sitemaps, links from pages it fetched earlier, and addresses people submit. It requests the page, reads the HTML, and follows links that look public. If robots.txt tells it to skip a path, it should skip it. If a page is behind a login, it should not pretend it saw the contents.

Oernoe Search is built to index publicly visible pages. It is not a way to open someone’s Health notes, Chat thread, Drive folder, or account settings. Those surfaces are not the public web, and they are not part of the search index.

We also do not treat crawl as a harvesting job for email lists or advertising audiences. The point of the fetch is to know what a page says, not to build a dossier on the person who might click it later.

The index is a working copy, not a profile

After a fetch, the engine stores a working copy of the page: text, titles, links, and enough structure to retrieve it later. That store is the index. If the page changes, the copy goes stale until the crawler comes back. That lag is normal. Anyone who tells you an index is always current is selling something.

What we do not store as “who you are” is the query you typed. Search history is not an advertising profile at Oernoe. We have said that on the About page, the Privacy Policy, and the Search guide because reviewers and readers both need it stated in the same way every time.

There is a separate fact that often gets flattened into the first one: selected publisher pages on this website may show Google ads. Google may then process page context, cookies, IP address, and device data for those ads. That is not Search selling your queries. It is this publisher site using AdSense on pages that already have original writing. The Privacy Policy names the pages and the opt-outs.

Ranking without a surveillance shortcut

Ranking is the ordered list you see. Engines that fund themselves with personalized ads have a shortcut: they already know a lot about the person asking. Oernoe Search is not built on that shortcut. Ranking here is supposed to follow the query and the page, not a file of previous clicks sold to advertisers.

Signals we do use are ordinary: whether the page is about the words you typed, whether the page looks like a finished document rather than a doorway, how it links, and whether it is still reachable. The longer write-up is How Oernoe Search works.

Signals we do not want in the ranking loop: a dossier of your searches, a purchased audience segment, or a guess about your health or chat contents. If a ranking idea needs those, it does not ship. That is a product rule, not a slogan on a landing page.

Listings people submit

Oernoe Search is not only a crawler of other people’s HTML. People can add public people and company listings, and those go through an approval process you can see in the product. That is a separate pile of documents: they are public because someone asked them to be, not because we scraped a private inbox.

Approval is slow on purpose. A search index that accepts every submitted name without a check becomes a spam folder. If a listing is rejected, that is the product working. If you think a listing about you should come down, write support@oernoe.com or privacy@oernoe.com with enough detail to find it.

What this means if you publish

If you want Oernoe Search to find a page, make a real page. Give it a title that matches the document. Write the answer on the page, not only in an image. Link to it from somewhere the crawler can already see. Keep a sitemap. Do not block the crawler in robots.txt and then wonder why it never arrives.

If you want a person to trust the page, say who wrote it, when it was updated, and how to reach you. That is also how this site is supposed to work. The journal, How-To, and guides exist so a reader can check our claims against public text instead of a press quote.

What this page is not

This is not a generic tour of Google, Bing, and “the algorithm.” Those write-ups are everywhere, and they were part of why a reviewer could look at this site and see low value content. The useful version is narrower: what Oernoe Search fetches, what it stores, what it ranks on, and where Google ads on this publisher site still apply.

If a later product change affects crawling or ads, this page should be updated or taken down. The date above is the last time we checked it against the live services.