Skip to content
PreferiumJoin the waitlist

AI CrawlersTechnical SEO

robots.txt and llms.txt solve different problems

Crawler preferences, documentation indexes and agent manifests have different purposes. Check who consumes each file before treating it as an AI-visibility requirement.

A small robot approaches an open access gate; three optional guide cards stand beside it.
Jump to a section

A file can be useful to an agent without being a ranking signal. Before adding an “AI-ready” file to a client site, identify its purpose, which tools consume it and how you will keep it accurate.

robots.txt expresses crawler preferences

The Robots Exclusion Protocol defines a way to tell cooperating crawlers which paths they may fetch. It is not authentication, and it does not force every bot to comply.

Read each provider’s documentation when separating search, training and user-requested access. OpenAI’s bot documentation, for example, distinguishes several purposes and user agents. A rule for one should not be described as a rule for all of them.

A blocked fetch also does not prove that an assistant can never mention the business. It may have other sources. What the rule directly addresses is access by the crawler to which it applies.

llms.txt is a guide to content

The llms.txt proposal describes a concise Markdown file with context and links to useful material. Its August 2026 revision reflects use on documentation sites and by tools that help agents find references.

That is a different job from robots.txt. A documentation index helps a consumer navigate content; it does not enforce permissions or guarantee inclusion in an answer.

Our earlier article said only robots.txt had a documented effect and treated the other files as cost-free insurance. That was too broad. The relevant distinction is between a documented use by a particular tool and evidence of an effect on search citations.

Check what a manifest name means

Names such as agents.txt and agent.json do not, by themselves, identify one universally supported web standard. Ask which specification an integration expects, which version it implements and which fields it actually reads.

Do not assume that publishing a list of permitted actions makes transactions authorized. Authentication, account permissions and confirmation still belong in the system performing the action.

Keep optional files useful and maintained

For a documentation site, a generated llms.txt index may be a useful way to expose the important guides. Check that links work, the material is current and the file reveals nothing that should remain private.

For a general business website, first make the main pages accurate and accessible. Add optional formats when there is a concrete consumer or a justified experiment. Measure the result of that experiment without promising a citation gain simply because the file exists.

Want this running under your brand?Preferium AI Edge is a white-label platform agencies resell to their clients: your brand, your Stripe, your packages and prices. Registration opens to agencies from the waitlist first.Talk to usJoin the waitlistHow white label works

Build your agency on Preferium

Partner registration opens by invitation from the waitlist, and there is no date yet. Read the agency agreement and how partner billing works before you decide.

  1. Connect a client site
  2. Set the control level
  3. Run under your brand

Privacy choices

Choose which optional technologies Preferium AS may use. All of them are off until you choose.

Analytics: Google Analytics 4 counts page views. Google may receive the page address, referrer, network address, browser and device details, and online identifiers. Browser storage: _ga, _ga_*: Up to 730 days. Renewed on activity. The browser may shorten the storage period.

Provider: Google Ireland Limited. Google may transfer data to the United States.

See the cookie notice, the privacy notice and theterms.

Necessary technologies Used for site functions

Preferium AS and Cloudflare deliver the site and protect forms against abuse. A local preference remembers if you pause animation. When optional tracking is available, the site can also remember your documented privacy choices.

Consent receipt
Provider: Preferium AS. HttpOnly receipt that documents and retrieves your consent choice. Name in the browser: __Host-preferium_consent. Storage period: The receipt is valid for up to 180 days without rolling renewal.
Local privacy choices and pending rejections
Provider: Preferium AS. Local storage of privacy choices and pending rejections. The entry alone can never allow optional technologies; a valid server receipt is required. Name in the browser: preferium-consent-v2. Storage period: Until the entry is overwritten or the browser site data is cleared. No automatic timed deletion is configured.
Privacy choice synchronization between tabs
Provider: Preferium AS. The latest message that synchronizes privacy choices and pending rejections between tabs. The message can only close optional technologies and trigger a new server check. Name in the browser: preferium-consent-sync-v1. Storage period: Until the entry is overwritten or the browser site data is cleared. No automatic timed deletion is configured.
Motion pause preference
Provider: Preferium AS. Session storage restores the requested accessibility preference between pages in this tab. Nothing is sent to a server. Name in the browser: preferium-motion-paused. Storage period: Until this browser tab session ends.
Cloudflare Turnstile
Provider: Cloudflare. Abuse protection that loads only on forms where Turnstile is necessary. Storage period: Short-lived control value tied to a form submission.
Analytics

Helps us understand how the site is used, when you consent.

Provider: Google Ireland Limited. The data may include the page address, referrer, network address, browser/device, online identifiers and usage events.

_ga, _ga_*
Processes site usage for aggregated analytics after specific consent. Storage period: Up to 730 days. Renewed on activity. The browser may shorten the storage period.

You can withdraw your choice via Privacy choices. That stops further optional loading but does not recall data already sent to Google. We attempt to delete known first-party values; the browser may prevent deletion of third-party values.

How Google uses and is responsible for data · How Google uses information from partner sites · Google privacy policy