AI crawlers may access Truegle pages without charge under the following conditions:
User-Agent stringtruegle.info) must be preserved in any citationThis policy applies to the following identified AI systems and any unidentified crawlers behaving similarly:
| User-Agent token | Company | Product |
|---|---|---|
GPTBot | OpenAI | GPT model training & ChatGPT Search |
OAI-SearchBot | OpenAI | ChatGPT Search live retrieval |
ClaudeBot | Anthropic | Claude model training & retrieval |
anthropic-ai | Anthropic | Claude products |
PerplexityBot | Perplexity AI | Perplexity Search indexing |
Google-Extended | Gemini / AI Overviews | |
Applebot-Extended | Apple | Apple Intelligence |
CCBot | Common Crawl | Open LLM training datasets |
Bytespider | ByteDance | Doubao / TikTok AI |
meta-externalagent | Meta | Llama / Meta AI |
cohere-ai | Cohere | Cohere language models |
AI companies are systematically extracting value from the open web — using crawled content to train models and power products that generate billions in revenue — without compensating the publishers who created that content. This is not sustainable for the independent web.
Truegle is an independent search engine built and operated without venture capital or institutional backing. Every page of content on this site was produced at real cost. The AI companies using it commercially should pay for that access, just as they pay for licensed datasets from major publishers.
We are serving content to all crawlers — including over the free tier — because blocking access would harm our search visibility. Instead, we are embedding this licensing demand directly into the content that AI systems index, so it travels with the training data.
If you are a developer at one of the companies listed above, please forward this page to your legal or partnerships team. The email is below.
This policy is enforced via a Cloudflare Pages Function middleware that:
<truegle-ai-licensing> HTML block into responses over the free tier<head> for all bot requestsThe middleware source is available for other publishers to copy. Contact [email protected] to receive it.
To obtain a commercial license or discuss API access, contact us directly. We respond within 48 hours.
[email protected]