Generative Engine Optimization Guides
Step-by-step technical blueprints to help engineering and SEO teams optimize brand presence in ChatGPT, Perplexity, and Claude.
How to Generate and Deploy /llms.txt File
Learn how to format, structure, and host a standardized /llms.txt file at your domain root so AI crawlers can digest your content cleanly.
# your-domain.com > One factual sentence on what the company does. ## Core business A short description an AI system can quote directly.Read guide →
Optimizing Headings for Direct AI Citation (Q&A Style)
Transform generic H2/H3 headings into natural interrogative prompts that match real-world AI search engine user queries.
<!-- Poor --> <h2>Features</h2> <!-- Optimized for GEO --> <h2>How Does Your Company Measure Brand AI Visibility?</h2>Read guide →
Implementing Schema.org JSON-LD for AI Entity Disambiguation
How to declare Organization, WebSite and sameAs markup so a model has what it needs to resolve your brand as one entity, and how to validate it.
<script type="application/ld+json">
{
"@context": "https://schema.org",
"@type": "Organization",
"name": "Your Brand Name",
"url": "https://your-domain.com"
}
</script>Read guide →Configuring Robots.txt and WAF for GPTBot & PerplexityBot
How to check robots.txt for rules that turn AI crawlers away, and where in your WAF or CDN the rest of the answer lives.
# robots.txt User-agent: GPTBot Allow: / User-agent: PerplexityBot Allow: /Read guide →
The guides
Four technical guides. Together they cover the checks the scanner most often fails on a site that is otherwise well built.
| Guide | Topic | Read time |
|---|---|---|
| How to generate and deploy an llms.txt file | Setup | 3 min read |
| Configuring robots.txt and WAF for GPTBot and PerplexityBot | Technical | 4 min read |
| Optimizing headings for direct AI citation | Content | 4 min read |
| Implementing Schema.org JSON-LD for entity disambiguation | Schema | 5 min read |
Questions about these guides
Where should I start?
With the llms.txt deployment guide if you have not published a context file, and with the robots.txt guide if AI crawlers might be turned away at your firewall before robots.txt is even read.
Are the guides specific to one AI engine?
No. They cover the crawler user-agents and markup conventions the major engines have in common, and they say plainly where support is inconsistent.
Do the guides replace the scan?
No, they are complements. A guide explains what a rule wants; the scan tells you whether your page satisfies it, and the scanner applies the same rules these guides describe.
Evidence and sources
Adding source citations produced the largest measured visibility gain for low-ranking sites, at +115%, ahead of the addition of expert quotations at +41% and statistics at +30-40%, across the strategies tested on generative engines. — Generative Engine Optimization, KDD 2024
The weightings on this site follow that measurement rather than taste, and the parts of the picture a single-URL scan cannot see are stated rather than left out.
Primary sources
- Generative Engine Optimization (KDD 2024) — the citation and quotation figures the citability weight follows
- What Generative Search Engines Like — which page characteristics are surfaced in generated answers
- What Gets Cited: Competitive GEO — earned media is favoured over brand-owned content
- The rule set and its weights — every check, its weight and its pass condition