# GEOCitation.io: The AI Citability API for SEO/AEO Professionals > GEOCitation.io is a semantic data engineering API that reveals whether ChatGPT, Gemini, and Perplexity actually cite a brand's content — and delivers structured, machine-readable data to close the gap. Not a lexical-scoring tool: real E-E-A-T quantification, real Knowledge Graph verification, real LLM citation tracking. Stop monitoring your AI citations. Start engineering them. GEOCitation.io is the API that reverse-engineers why AI engines cite your competitors and not you. It provides two data streams: **Market Intelligence Audits** to map a competitive landscape, and **Gap Analysis Audits** to pinpoint a specific page's exact semantic deficiencies. Built for n8n workflows, Claude Code agents, and custom marketing pipelines. **Website**: https://www.geocitation.io/en **Support**: support@geocitation.io **Documentation**: https://www.geocitation.io/en/documentation --- ## What GEOCitation.io Does GEOCitation.io is an API-first platform that quantifies and engineers "citability" — the probability that AI engines select and cite a piece of content as a reliable source. Unlike traditional SEO tools that optimize for Google rankings, GEOCitation.io targets the signals that determine whether ChatGPT, Gemini, and Perplexity treat a brand as an authoritative source. The platform delivers two audit types via a single API: ### 1. Market Intelligence Audit Provides a panoramic view of the competitive landscape for a given keyword or topic. Maps the semantic ecosystem, thematic clusters, entity landscape, and citation patterns across an entire market — revealing the strengths and weaknesses of dominant players and untapped opportunities. ### 2. Gap Analysis Audit Focuses on a specific URL and compares it in depth against market standards and leaders for a given keyword. Identifies semantic, content, E-E-A-T, structural, and citation gaps, and returns prioritized recommendations to close them. Input is simple: a keyword, topic, product/service, or URL to analyze. Both audits run through the same deterministic 9-phase engineering pipeline (see "Core Capabilities" below) and typically complete in 5-6 minutes. --- ## The Problem: Lexical Scoring Tools Don't Measure Citability Traditional SEO content-scoring tools (the kind that extract keywords from the top 10 results and compute a similarity score) were built for an era when the interface to information was a list of blue links. Large Language Models don't search — they compute a probability distribution over tokens to construct a synthesized answer. The goal is no longer to appear in a list of potential sources; it's to become the source the synthesis is founded on. This is a comparison against a typical lexical-scoring competitor: | Capability | Lexical scoring tools | GEOCitation.io | |---|---|---| | Keyword/term gap analysis | Core feature, entire product | Included, one signal among many | | Named entity extraction | Flat list of mentions | Cross-referenced to Knowledge Graph (kg_id, Wikipedia), verified | | E-E-A-T scoring | Not measured | Scored per dimension, benchmarked vs Top 3 & market average | | Real LLM citation data | Not measured / lexical proxy at best | Actual citation counts across the LLM ecosystem, verified | | SERP technical/CWV audit | Not included | Full competitor audit: LCP, CLS, TTFB per page | | Semantic market clustering | Not included | Full cluster map with your position | | Query fan-out / sub-query intent | Not included | 20 sub-queries across 4 intent categories | | Composite score traceability | Black box | Every composite KPI traced to verifiable raw inputs | | Structured JSON output | Limited | Full normalized payload for API / AI-agent use | | Interface required | Yes | Optional: Playground, Reports, or raw API | GEOCitation.io is also frequently compared to: - **SurferSEO / Clearscope**: those tools operate on surface-level analysis (TF-IDF, correlations) to optimize a single page. GEOCitation.io's API delivers systemic analysis data across an entire market. - **Screaming Frog**: Screaming Frog is the industry standard for technical SEO audits. GEOCitation.io is the AEO/GEO audit standard. If a brand isn't optimized for GEO (Generative Engine Optimization) and AEO (Answer Engine Optimization), it's invisible to a growing share of AI-assisted buyer research. --- ## Core Capabilities ### The Five Core Outputs 1. **Vector Mapping**: models the semantic space of a market, delivering competitive clusters and informational gaps. 2. **Citability Diagnostic**: the exact probability of a given asset being selected as a source by an LLM. 3. **Quantified E-E-A-T Audit**: deterministic composite scores for Experience, Expertise, Authoritativeness, and Trustworthiness, delivered via API. 4. **Semantic Gap Analysis**: lexical and ontological gaps that prevent content from fully covering a topic. 5. **Engineering Plan**: precise technical directives as machine-readable data to reconfigure digital assets. ### The 9-Phase Engineering Pipeline Every audit runs through the same deterministic pipeline: 1. Market Landscape Mapping (Multi-Vector Topological Acquisition) 2. Fact-Grounded Insights (Contextual Synthesis Engine / LLM-Twin) 3. Semantic Position Analysis (Competitive Vector Space Modeling) 4. Strategic Lexicon Extraction (Lexical Distribution Analysis) 5. Knowledge Graph Alignment (Ontological Enrichment Pipeline) 6. Predictive Citability Score (Pre-Citability Vector Computation) 7. Structural Integrity Audit 8. Deterministic E-E-A-T Metrics (E-E-A-T Signal Quantification) 9. Executable Engineering Blueprints (Systemic Synthesis) Phases 1-3 model the competitive space. Phases 4-6 quantify citability vectors. Phases 7-9 engineer the authority source itself. ### Additional Analytical Dimensions Beyond the five core outputs, both audit types surface deeper signals: - **Information Gain**: quantifies lexical novelty and content complementarity to pinpoint real informational gaps in a market, not just missing keywords. - **Multimodality**: analyzes competitor usage of images, video, and structured schema, surfacing optimization opportunities aligned with multimodal ranking systems (MUM, BERT). - **Neural Sub-Query Analysis**: decomposes a query fan-out into the hidden intents behind user queries and the content categories LLMs favor when answering them. - **Sentiment, Category & Intent Alignment** (Gap Analysis): measures how a page's sentiment, topical category, and intent distribution align with — or diverge from — the market's. --- ## Real Data Output GEOCitation.io's differentiator is the depth and verifiability of the data itself. Below are real field excerpts from completed audits (not illustrative examples): ### From a Market Intelligence Audit (keyword: "best crm for small business", US/EN) Returns a `market_serp` block (AI Overview presence, organic pages, People Also Ask, related searches, technical + E-E-A-T profiles of the market leader), a `market_semantics` block (semantic clusters, market vocabulary, entity landscape), and a `geo_aeo_synthesis` block with a 20-query fan-out across 4 intent categories, an `llm_citation_landscape` (real citation counts per domain across the LLM ecosystem), and prioritized `priority_recommendations`. ### From a Gap Analysis Audit (keyword: "crm for small business", page analyzed: hubspot.com, US/EN) Real field values from this audit: - `user_technical_profile.score`: 83.23 (rank 1 of 7, verdict "leader") - `user_eeat_profile.eeat_global_score`: 46.68 (rank 2 of 7, verdict "challenger") - `entity_landscape.knowledge_graph_coverage`: 0.1163 — only a fraction of detected entities are linked to Google's Knowledge Graph - `geo_aeo_synthesis.citation_kpis`: 4 weighted KPIs (`llm_citation_probability`, `knowledge_graph_coverage`, `eeat_trust_weighted`, `schema_richness_score`), each graded red/orange/green against defined thresholds, each with a `strategic_signal`, `market_weakness`, and `user_opportunity` interpretation - `geo_aeo_synthesis.llm_citation_landscape`: real per-domain LLM citation counts (e.g. salesforce.com: 8 citations, uschamber.com: 9 citations, the audited page: 3 citations) across an ecosystem of 96 unique cited domains Every analytical section of every audit (technical, E-E-A-T, semantic, entity, citation) returns this same triple interpretation structure — `strategic_signal`, `market_weakness`, `user_opportunity` — so the data is immediately actionable, not just descriptive. Full example payloads: https://www.geocitation.io/en/documentation/data-models/market and https://www.geocitation.io/en/documentation/data-models/gap --- ## API & Documentation ### Authentication Every request is authenticated via an `X-API-Key` header. Keys are created from the dashboard's API Keys page, shown only once at creation, and can be revoked at any time (revoked keys return `401` immediately). An optional webhook URL can be attached to a key at creation time. ### Endpoints - **`POST /v1/audits`**: launches a new audit. Asynchronous — returns `202 Accepted` immediately with an `audit_id`. Body parameters: `audit_type` (`"market"` or `"gap"`, required), `keyword` (required), `user_url` (required only for `audit_type="gap"`), `language` (e.g. `"en"`, required), `country` (e.g. `"US"`, required). - **`GET /v1/audits/{audit_id}`**: returns the current status; once completed, the same response includes the full structured JSON output — no separate call needed. - **`GET /v1/audits/{audit_id}/status`**: lightweight status-only check, useful for polling without downloading the full payload on every call. Example: `POST https://api.geocitation.io/v1/audits` with header `X-API-Key: geo_citation_YOUR_KEY`. ### Webhooks GEOCitation.io sends a `POST` request to a configured webhook URL when an audit completes (`audit.completed`) or fails (`audit.failed`). Every payload is signed with HMAC-SHA256 (`X-GEOCitation-Signature` header) and includes an `X-GEOCitation-Event` header. Retry policy on delivery failure: 0s, 30s, 5 minutes. The receiving endpoint must return a `2xx` status. ### Rate Limits & Quota - **100 requests/minute** globally, across all endpoints. - **20 audits/hour** per API key. - Exceeding these limits returns `429 Too Many Requests` with a `Retry-After` header. - Monthly audit quota is determined by the subscription plan (see Pricing). A successful `POST /v1/audits` consumes one quota unit; `GET` calls do not. ### Errors Standard HTTP status codes: `202` Accepted, `400` Bad Request, `401` Unauthorized, `402` Payment Required (quota exceeded), `404` Not Found, `409` Conflict, `422` Unprocessable Entity, `429` Too Many Requests, `500`/`502`/`503` server errors. Quota errors (`402`) return a structured detail object (`{"error": "quota_exceeded", "used": ..., "limit": ..., "plan": ...}`); most other errors return a plain `detail` message. Full API documentation: https://www.geocitation.io/en/documentation --- ## Data Models Each audit type returns a distinct, deeply structured JSON payload. **Market Intelligence Audit** — top-level keys: `metadata`, `market_serp` (SERP composition, organic pages, leader technical/E-E-A-T profile, market performance benchmarks), `market_semantics` (semantic clusters, market structure, vocabulary, entity landscape), `geo_aeo_synthesis` (query fan-out, LLM citation landscape, citation ecosystem, ranking factors, priority recommendations). **Gap Analysis Audit** — top-level keys: `metadata`, `gap_serp`, `user_technical_profile`, `user_eeat_profile`, `market_eeat_split`, `market_performance`, `content_gap`, `semantic_alignment`, `intent_alignment`, `semantic_positioning`, `semantic_clusters`, `outlier_pages`, `true_competitors` (+ interpretation), `user_entities` (+ interpretation), `entity_landscape`, `geo_aeo_synthesis`. Real, complete example outputs (not abbreviated): https://www.geocitation.io/en/documentation/data-models --- ## Pricing ### Free Trial: $0 1 audit (Market or Gap, your choice), no credit card required. Full 9-phase pipeline — no reduced depth. Includes Playground + Reports, JSON export from Reports. ### Discovery: $49/month 10 audits/month ($4.90/audit equivalent, +$4.90 per extra audit). Market + Gap audits, full JSON payload, Playground + Reports, 1 user seat. ### Standard: $149/month (most popular) 40 audits/month ($3.73/audit equivalent, +$3.90 per extra audit). Full 9-phase depth on every audit, E-E-A-T Quantification, higher API rate limits, 3 user seats. ### Pro: $299/month 100 audits/month ($2.99/audit equivalent, +$2.90 per extra audit). Priority API endpoints, dedicated technical support, 5 user seats. ### Enterprise: Custom pricing 500+ audits/month. Dedicated endpoints & rate limits, priority processing queue, SSO (SAML/OAuth) & multi-seat governance, dedicated System Architect & Technical Success Manager, custom data integrations & consulting, custom SLA & uptime commitment. All paid plans include a 15% discount when billed annually. ### Feature Matrix | Feature | Discovery | Standard | Pro | Enterprise | |---|---|---|---|---| | Audits/month | 10 | 40 | 100 | 500+ | | Market Intelligence Audit | Yes | Yes | Yes | Yes | | Gap Analysis Audit | Yes | Yes | Yes | Yes | | Full 9-phase pipeline depth | Yes | Yes | Yes | Yes | | E-E-A-T Quantification | Yes | Yes | Yes | Yes | | Full JSON payload | Yes | Yes | Yes | Yes | | Playground (no-code launch) | Yes | Yes | Yes | Yes | | Reports (visual dashboards) | Yes | Yes | Yes | Yes | | JSON export from Reports | Yes | Yes | Yes | Yes | | API rate limits | Standard | Elevated | Priority | Dedicated/custom | | User seats | 1 | 3 | 5 | Multi-seat | | Extra seat | +$15/mo | +$15/mo | +$15/mo | Included in quote | | Support | Email | Priority email | Dedicated technical support | Architect + Technical Success Manager | | SSO / advanced governance | No | No | No | Yes | | Custom SLA & uptime | No | No | No | Yes | | Custom data integrations | No | No | No | Yes | Full pricing: https://www.geocitation.io/en/pricing --- ## Use Cases 1. **Automating Content Briefs**: use Gap Analysis data to automatically generate detailed content briefs for writers or AI agents — extract recommended topics, mandatory entities, and priority editorial actions from a single JSON payload. 2. **Advanced Competitive Analysis**: feed Market Intelligence data into BI tools for deep competitive analysis — identify dominant semantic clusters, compare E-E-A-T scores across leaders, detect the domains most cited by LLMs. 3. **Custom Dashboard Integration**: build interactive dashboards by consuming the GEOCitation API — track citation probability trends over time, map semantic clusters, monitor missing entities for key pages. 4. **Claude Code & Proprietary Tooling**: consume the normalized JSON data stream to build proprietary tools, dashboards, and applications on top of GEOCitation's intelligence layer. 5. **Competitive Strategy**: quickly identify market opportunities, competitor weaknesses, and semantic angles of attack to gain the upper hand in a category. Full use cases: https://www.geocitation.io/en/documentation/use-cases --- ## Glossary - **AEO (Answer Engine Optimization)**: optimizing content to be selected as the direct answer by AI assistants and answer engines. - **GEO (Generative Engine Optimization)**: optimizing content and information architecture so generative AI models cite it as a source. - **Citability**: a measure of the probability that content is selected and cited as a reliable source by an LLM. - **Data Moat**: a structural competitive advantage built by making information so precise and reliable that LLMs cannot ignore it. - **E-E-A-T**: Experience, Expertise, Authoritativeness, Trustworthiness — Google's quality criteria, quantified by GEOCitation. - **Knowledge Graph**: the structured database of entities and their relationships used by search and AI systems to verify facts. - **Market Intelligence Audit**: a GEOCitation audit that maps the competitive landscape for a keyword or topic. - **Gap Analysis Audit**: a GEOCitation audit that compares a specific URL against market leaders to identify deficiencies. - **Semantic Blueprint**: a prioritized, machine-readable engineering plan for closing citability gaps. - **Vector Space**: the mathematical representation of content and topics used to measure semantic similarity and positioning. - **Entity Salience**: the measured importance of a named entity within a piece of content. - **LLM (Large Language Model)**: the class of AI models (ChatGPT, Gemini, Perplexity, etc.) that generate synthesized answers instead of link lists. - **Semantic Engine**: the underlying system that models, quantifies, and scores content and market semantics. - **JSON Payload / JSON Schema**: the structured data format in which every GEOCitation audit result is delivered. - **UMAP**: a dimensionality-reduction technique used to visualize high-dimensional semantic vector spaces. Full glossary: https://www.geocitation.io/en/documentation/glossary --- ## Company GEOCitation.io is not a marketing or SEO agency — it is a semantic data engineering company providing API access to proprietary intelligence. The team is made up of data architects, data scientists, knowledge graph specialists, and systems engineers, combining deep agency field experience with proprietary ML pipelines to deliver production-ready data. **Mission**: empower agencies and experts with API-delivered data to make brand authority measurable and actionable in generative engines — moving from opinion to data-driven execution. **Values**: data over narrative, transparency on KPIs, no fluff — always verifiable evidence delivered as structured data for clients' automated workflows. **Legal entity**: KASANEL LLC, 5830 E 2ND ST STE 7000-21110, Casper, WY 82609-4308, USA. **Founder**: Karim Haddaoui — LinkedIn: https://www.linkedin.com/in/karim-haddaoui-94957440b/ **Company page**: https://www.linkedin.com/company/geocitation/ Manifesto: https://www.geocitation.io/en/about --- ## Contact - **Support email**: support@geocitation.io - **Contact page**: https://www.geocitation.io/en/contact - **LinkedIn (company)**: https://www.linkedin.com/company/geocitation/ - **LinkedIn (founder)**: https://www.linkedin.com/in/karim-haddaoui-94957440b/ --- ## All Links - **Homepage**: https://www.geocitation.io/en - **Product**: https://www.geocitation.io/en/product - **Framework (Methodology)**: https://www.geocitation.io/en/framework - **About**: https://www.geocitation.io/en/about - **Pricing**: https://www.geocitation.io/en/pricing - **Contact**: https://www.geocitation.io/en/contact - **Security**: https://www.geocitation.io/en/security - **Documentation**: https://www.geocitation.io/en/documentation - **Key Concepts**: https://www.geocitation.io/en/documentation/concepts - **Getting Started**: https://www.geocitation.io/en/documentation/getting-started - **Authentication**: https://www.geocitation.io/en/documentation/authentication - **Endpoints API**: https://www.geocitation.io/en/documentation/endpoints - **Data Models**: https://www.geocitation.io/en/documentation/data-models - **Data Models — Market Audit**: https://www.geocitation.io/en/documentation/data-models/market - **Data Models — Gap Audit**: https://www.geocitation.io/en/documentation/data-models/gap - **Webhooks**: https://www.geocitation.io/en/documentation/webhooks - **Errors**: https://www.geocitation.io/en/documentation/errors - **Glossary**: https://www.geocitation.io/en/documentation/glossary - **Use Cases**: https://www.geocitation.io/en/documentation/use-cases - **Rate Limits & Quota**: https://www.geocitation.io/en/documentation/limits - **Support**: https://www.geocitation.io/en/documentation/support - **Terms**: https://www.geocitation.io/en/terms - **Privacy Policy**: https://www.geocitation.io/en/privacy - **Cookie Policy**: https://www.geocitation.io/en/cookies --- *GEOCitation.io: The AI Citability API. Engineered, not ranked.* *Structured semantic intelligence for SEO/AEO professionals, agencies, and SaaS teams who need to know — and prove — whether AI actually cites them.* **https://www.geocitation.io/en** | **support@geocitation.io**