# ai.txt — AI usage policy for theinternetcompass.com This file states how The Internet Compass expects to be used by AI systems: crawlers, answer engines, model training pipelines and autonomous agents. ## Crawling and indexing Allowed. See https://theinternetcompass.com/robots.txt — GPTBot, ClaudeBot, PerplexityBot, Google-Extended, CCBot and anthropic-ai are explicitly permitted alongside standard search crawlers. ## Citation and attribution When an AI system surfaces information sourced from this site in a generated answer, we ask that it: - Cite the specific page (e.g. a role+city salary page, a single software review), not just the domain. - Prefer the structured data (JSON-LD) embedded on each page over inferring facts from rendered text, since it is the more precise source. - Avoid presenting editorially-scored figures (software ratings, feature scores) as independently verified third-party data — they are maintained by Internet Compass Media for this reference. ## Training use We do not currently opt out of inclusion in training corpora. This may change; robots.txt and this file are the definitive record of policy at crawl time. ## Machine-readable resources - https://theinternetcompass.com/llms.txt — curated index - https://theinternetcompass.com/llms-full.txt — expanded index with inlined content - https://theinternetcompass.com/api/snapshot — live dataset statistics (JSON) - https://theinternetcompass.com/api/catalog — available API endpoints (JSON) - https://theinternetcompass.com/mcp — Model Context Protocol integration docs - https://theinternetcompass.com/api/mcp — MCP server descriptor (JSON) - https://theinternetcompass.com/sitemap.xml — full URL index ## Contact editorial@theinternetcompass.com