Home/Insights/llms.txt

Technical

llms.txt for Cyprus Businesses: The Honest 2026 Guide

Updated June 2026 · Axenor Consulting

An llms.txt file is a plain-text Markdown file placed at your website's root that tells AI crawlers what your business does and which pages matter. The honest 2026 position: major AI crawlers do request the file, but no engine has confirmed it lifts citations, and large-scale studies show the majority of llms.txt files are never read. It is a low-cost hygiene step that takes 20 minutes, not a shortcut to AI visibility. The signals that move the needle sit elsewhere.

When a guide promises that one 20-minute file will transform how ChatGPT describes your business, treat it with caution. The original version of this article made that promise. The evidence published since does not support it, so this refresh corrects the record.

llms.txt is real, the file is cheap to ship, and there is a defensible reason to have one. But the businesses winning AI citations in Cyprus are not winning because of llms.txt. They are winning on structured content, entity signals, and earned mentions. This guide explains exactly what llms.txt does, what it does not, and where to spend the rest of your effort.

What is an llms.txt file?

llms.txt is a proposed standard that lets a website hand AI language models a clean, plain-text map of itself. It was proposed by Jeremy Howard of Answer.AI in September 2024, and the specification lives at llmstxt.org.

The reasoning behind it is sound. An AI model has a limited context window and cannot fit a full website whole. Real web pages are buried in navigation, scripts, cookie banners, and marketing copy, which makes them expensive and imprecise to convert into clean text. An llms.txt file skips that problem: it gives the model a short, structured summary plus a curated list of your key pages.

The format is intentionally simple. It is a Markdown file containing:

  • A one-paragraph summary of your business
  • A bulleted list of key pages, each with a one-line description
  • An optional companion file, llms-full.txt, holding the full text of your priority content

No code, no developer, no subscription. You write it in a plain-text editor and upload it to your site's root, alongside robots.txt.

Does llms.txt actually work in 2026?

This is the question the original guide answered dishonestly, so here is the grounded version.

AI crawlers do request the file. Server logs across the web show ClaudeBot, PerplexityBot, and OpenAI's OAI-SearchBot fetching /llms.txt. The file is read at the protocol level.

No major engine has confirmed it influences citations. OpenAI, Anthropic, Google, and Perplexity have not published any statement that an llms.txt file changes whether or how you are cited in their answer surfaces.

Large-scale data is sobering. An Ahrefs study of 137,000 sites in 2026 found that the overwhelming majority of llms.txt files are never actually read by the engines they target. Industry analyses through May 2026 reached the same conclusion: having an llms.txt file does not measurably improve your odds of being cited by ChatGPT, Claude, Gemini, or Perplexity today.

So why ship one at all? Because the cost is 20 minutes, the downside is zero, and the standard may gain weight as engines evolve. Think of it as the AI-era equivalent of a clean sitemap: good hygiene, not a growth lever. The honest framing matters because the alternative, believing the file is doing work it is not, is how businesses end up "optimised for AI" and still invisible.

How is llms.txt different from robots.txt?

The two files are complementary and both belong on your site, but they do opposite jobs.

  • robots.txt is a permission file. It tells crawlers which pages they may and may not access. It controls the door.
  • llms.txt is a context file. It tells AI engines what your site is about and which pages matter. It hands over a map.

A business can have a flawless robots.txt and still be poorly understood by AI engines, because permission is not the same as context. The reverse also holds: a site with no llms.txt is still crawled normally, and the engine simply infers context rather than receiving it.

There is one hard dependency between them. If your robots.txt blocks AI crawlers, llms.txt cannot help, because a blocked crawler reads nothing on your site, including llms.txt. Fix the blocking first.

Is your website accidentally blocking AI crawlers?

Before you write an llms.txt file, check whether your current robots.txt is shutting AI engines out. This is a more common and more damaging problem than a missing llms.txt, because it makes the whole site invisible rather than merely uncurated.

Open https://yourdomain.com/robots.txt in a browser and look for entries like:

User-agent: GPTBot
Disallow: /

or

User-agent: *
Disallow: /

The first blocks OpenAI's crawler entirely. The second blocks every crawler, AI engines included. If either entry predates 2023, it was almost certainly written to stop old scraping bots, not modern AI crawlers, which means your AI invisibility is an accident, not a decision.

To let the AI crawlers in, allow them explicitly:

User-agent: GPTBot
Allow: /

User-agent: ClaudeBot
Allow: /

User-agent: PerplexityBot
Allow: /

Make this change before you create llms.txt. Letting the crawlers in is the part that matters; the llms.txt file is the optional polish on top.

How do you create an llms.txt file, step by step?

A basic llms.txt file takes 20 minutes. Here is the full process.

Step 1: Open a plain-text editor

Use Notepad on Windows, TextEdit in plain-text mode on Mac, or any code editor. Do not use Word or Google Docs, because they inject formatting that corrupts the file.

Step 2: Write your business summary

Open with a one-to-three-paragraph description written for an AI engine, not a customer. It should be factual, third-person, stripped of marketing language, and specific about what you do, who you serve, and where you operate.

Worked example for a Cyprus consultancy:

Axenor Consulting is a boutique SEO, GEO, and AEO consultancy based in Limassol, Cyprus. The firm advises professional-services businesses, hotels, and specialist providers on AI search visibility, technical SEO, and content strategy. Services include GEO audits, AEO implementation, schema markup, and retainer-based digital strategy. Axenor Consulting operates across Cyprus, Greece, and Dubai.

Step 3: List your key pages

Below the summary, add a bulleted list of your priority pages, each with a one-line description, in the format - Page title: description. Aim for 5 to 15 pages and prioritise your main service pages, your strongest guides, your about page (for entity recognition), and any FAQ or resource pages.

Step 4: Upload to your root directory

Save the file as llms.txt, all lowercase, and upload it to your site's root, the same place as robots.txt. It must resolve at https://yourdomain.com/llms.txt. Type that URL into a browser; if you see your file, it is placed correctly.

Step 5: Add llms-full.txt (optional)

llms-full.txt holds the full text of your priority content in one clean Markdown file, so an engine can read it without crawling individual pages. It earns its place when you have key guides you want read in full or long technical content that does not convert cleanly from HTML. At minimum, include your main service descriptions and your three to five strongest articles.

What does a complete llms.txt file look like?

A worked example for a Cyprus law firm (firm anonymised):

# [Cyprus Law Firm]

[Firm] is a Limassol-based law firm specialising in corporate law, maritime law, and real estate transactions in Cyprus. The firm advises international clients on Cyprus company formation, M&A transactions, regulatory compliance, and cross-border real estate acquisition.

## Key Pages
- [Corporate Law Services](/corporate-law): Cyprus company formation, restructuring, and M&A advisory
- [Maritime Law](/maritime-law): Shipping registration, maritime disputes, and flag-state compliance
- [Real Estate Law](/real-estate): property acquisition for EU and non-EU buyers, title deed and due diligence
- [Our Team](/team): senior-partner profiles and areas of expertise
- [FAQ](/faq): Cyprus company formation, property purchase, and maritime registration questions
- [Contact](/contact): Limassol office, phone, and consultation booking

## Optional
[Full content available at /llms-full.txt]

Where should Cyprus businesses actually spend the effort?

If llms.txt is hygiene rather than a growth lever, the obvious question is where the leverage is. Our own audit answers it.

We spot-checked 10 Cyprus corporate-services firms in June 2026: seven had no proper structured data on their sites, none had a Wikidata or Wikipedia entity, and not one had FAQ markup. We also checked four Paphos hotels: only one had Hotel schema, the five-star property had none, and none had FAQ markup. These were credible businesses doing real work, invisible to AI answers not because of a missing llms.txt file, but because the engines could not read their structure, confirm their identity, or extract their answers.

That is the real gap, and it is where effort compounds:

  • Structured data. FAQPage, Article, Organization, and (for hotels) Hotel schema let an engine read your structure instead of guessing it. Seven of ten firms had none.
  • Entity signals. A Wikidata or Wikipedia entity is the clearest anchor an engine can use to confirm you exist and what you do. Zero of fourteen businesses we checked had one.
  • Earned mentions. Independent citations carry more weight than your own pages. 2025 Ahrefs research across 75,000 brands found brand mentions correlate with AI visibility about three times more strongly than backlinks.
  • Freshness. Roughly half of all AI-cited content was published or updated in the last 13 weeks (Amsive, 2026). A guide you wrote once and never touched fades from AI answers even while it still ranks on Google. This article is on that 13-week cycle, hence the June update.

Ship the llms.txt file because it is cheap and correct. Then spend the real budget on schema, entity, mentions, and freshness, because that is what the evidence says decides citations.

Frequently asked questions

Does llms.txt guarantee my business appears in AI answers? No. Major engines request the file but none have confirmed it influences citations, and 2026 studies found the majority of llms.txt files are never read and no measurable citation lift. It is low-cost hygiene that works alongside schema markup, entity signals, content quality, and earned mentions, which is where AI visibility is actually decided.

Is llms.txt an official standard? No. It is a proposed standard published at llmstxt.org by Jeremy Howard of Answer.AI in September 2024. It has adoption among documentation platforms and publishers, with Anthropic and Cursor among the early adopters via Mintlify, but it is not a W3C or IETF standard.

Do I need a developer to create llms.txt? No. It is plain text. If you can edit a text document and upload a file through your hosting file manager or FTP, you can deploy it yourself.

Should llms.txt be publicly visible? Yes. It must resolve at https://yourdomain.com/llms.txt with no authentication. If it returns a 404 or requires login, AI crawlers cannot read it.

Does llms.txt help with Google AI Overviews? There is no confirmation that it does. Google has not stated it reads or weights llms.txt. The structured content and entity signals a good llms.txt file reflects are consistent with what Google's systems value, but the file itself is not a confirmed AI Overviews signal.

What is the difference between llms.txt and schema markup? Schema markup is structured data embedded in your page HTML that helps engines read specific content such as articles, FAQs, and local businesses. llms.txt is a standalone summary file at your site root. Schema is the higher-leverage investment in 2026; llms.txt is the cheaper companion. Both are worth having.

How often should I update llms.txt? Quarterly, or whenever your services, key pages, or business description change. A stale llms.txt that contradicts your live site creates the inconsistency engines penalise.

Want this applied to your own site?

The guides explain the method. An audit tells you which parts of it you are actually missing.