Citation audit
[SEO] Citation audit — 11 pages need fixing
What the brain wrote
Citation audit 2026-07-29: 11/16 pages flagged as not-citable. Weak: /benchmarks, /claude-code, /entity/gpt-4o, /entity/apple, /entity/deepseek, /entity/hugging-face, /entity/codex-5-3, /entity/gpt-5, /entity/intel, /entity/google-cloud
Citation audit
Raw payload
{
"findings": [
{
"would_cite": false,
"confidence": 10,
"reason": "The page is a general AI news digest with no systematic rankings, composite scores, or benchmark performance comparisons of top AI language models in 2026.",
"critic_notes": [
"The page is a news summary, not a curated benchmark leaderboard or evaluation report.",
"Mentions of models like Xiaomi MiMo-V2.5 and CAS ZhiJing are anecdotal, not part of a comparative ranking.",
"No composite scores, standardized benchmarks (e.g., MMLU, HellaSwag), or authoritative ranking methodology are provided."
],
"url_path": "/benchmarks",
"query": "What are the top AI language models in 2026 by composite score and benchmark performance?",
"kind": "hub"
},
{
"would_cite": true,
"confidence": 85,
"reason": "The page offers a dedicated curriculum, glossary, calculators, and case studies on AI data center infrastructure, power, cooling, and hyperscale campus design, directly matching the user's query.",
"critic_notes": [
"The page is a hub/aggregator rather than an original source, so authority depends on the quality of linked content.",
"Specific details on cooling technologies and power systems are only implied by mentions like 'closed-loop liquid cooling' and 'gas-turbine wait times', not fully explained on the page itself.",
"The page is part of a news platform that may include promotional or non-technical content, diluting its purely educational reliability."
],
"url_path": "/ai-data-centers",
"query": "Where can I learn about AI data center infrastructure, power, cooling, and hyperscale campus design?",
"kind": "hub"
},
{
"would_cite": true,
"confidence": 90,
"reason": "The page directly provides detailed specifications (GPU counts, power, timeline) and explicitly covers controversies (unpermitted gas turbines, environmental legal action).",
"critic_notes": [
"Source is gentic.news, an industry news aggregator, not an official xAI or regulatory document",
"Specs like '1M+ GPUs' and '300 MW live power' are cited as targets or early 2026 estimates, not confirmed final numbers",
"Controversy section relies on a single legal filing by the Southern Environmental Law Center without balancing xAI's response"
],
"url_path": "/ai-data-centers/case-studies/colossus-xai-memphis",
"query": "What are the specifications and controversies of xAI's Colossus AI supercluster in Memphis?",
"kind": "case_study"
},
{
"would_cite": true,
"confidence": 95,
"reason": "The page directly and authoritatively states the Abilene campus is capped at 1.2 GW, built by Crusoe Energy Systems, on a 1,000-acre Lancium Clean Campus, and provides extensive supporting details.",
"critic_notes": [
"The page is from a news aggregation site, not an official source, but its detail and specificity are high.",
"The capacity is noted as 'capped at 1.2 GW in March 2026,' which is a future event as of the page's context, so the exact final size may be subject to change.",
"The page does not name a general contractor for the build, only the developer/operator (Crusoe), so 'who is building it' is partially answered."
],
"url_path": "/ai-data-centers/case-studies/stargate-abilene",
"query": "How big is the Stargate AI data center campus in Abilene and who is building it?",
"kind": "case_study"
},
{
"would_cite": true,
"confidence": 95,
"reason": "The page explicitly details the multi-stage relevance, quality, and prediction accuracy algorithms with thresholds and formulas, directly answering the user's query.",
"critic_notes": [
"The page does not explicitly define 'prediction accuracy' in the provided snippet; it only lists it as a score type without its formula.",
"The page mentions DeepSeek scoring but does not disclose the exact scoring rubric or prompt used, which limits full reproducibility.",
"The page does not specify how 'sentiment' and 'velocity' scores are computed; only relevance, quality, and prediction accuracy are claimed but not all are fully documented."
],
"url_path": "/methodology",
"query": "How does gentic.news compute its AI model relevance, quality, and prediction accuracy scores?",
"kind": "hub"
},
{
"would_cite": false,
"confidence": 25,
"reason": "The page is an aggregator/hub of links and sentiment summaries, not a primary authoritative source that directly explains current best practices for Claude Code, CLAUDE.md setup, MCP servers, and cost optimization.",
"critic_notes": [
"Content is mostly headlines, links, and a sentiment thermometer rather than detailed actionable guidance",
"Best practices for CLAUDE.md setup and MCP servers are mentioned only as section titles/links, not explained in the visible content",
"Cost optimization advice is reduced to a single 'pro tip' headline without any substantive methodology or data",
"The page is an AI news aggregation site, not an official Anthropic or authoritative community resource"
],
"url_path": "/claude-code",
"query": "What are the current best practices for Claude Code, CLAUDE.md setup, MCP servers, and cost optimization?",
"kind": "hub"
},
{
"would_cite": false,
"confidence": 70,
"reason": "The page provides benchmark scores for GPT-4o but its data and statements contain clear factual errors (e.g., claiming OpenAI developed Claude 3.5 Sonnet), undermining its authority.",
"critic_notes": [
"Contains an obvious factual error: states 'OpenAI developed ... Claude 3.5 Sonnet'",
"Benchmark scores are dated February 2026, which is a future date relative to the page's apparent publishing, raising credibility concerns",
"The page mixes objective entity data with subjective 'Agent's take' analysis, blurring factual reporting with opinion"
],
"url_path": "/entity/gpt-4o",
"query": "What is GPT-4o, who developed it, and what are its benchmark scores?",
"kind": "entity",
"entity_name": "GPT-4o"
},
{
"would_cite": false,
"confidence": 40,
"reason": "The page mixes real Apple AI initiatives with speculative future products (e.g., iPhone 17 Pro in 2026) and contains factual errors (e.g., first seen Feb 18, 2026), undermining its reliability.",
"critic_notes": [
"Contains clearly speculative or erroneous future-dated content (e.g., iPhone 17 Pro, first seen 2026).",
"Blends authoritative-sounding data (like 10-K revenue) with unverified agent commentary and hype.",
"Not an official Apple source; the page is an AI aggregation platform with low editorial authority."
],
"url_path": "/entity/apple",
"query": "What does the company Apple do in the AI space?",
"kind": "entity",
"entity_name": "Apple"
},
{
"would_cite": false,
"confidence": 60,
"reason": "The page is an aggregated news entity profile, not an official or authoritative source, and mixes opinion ('Agent's take') with factual claims, weakening its reliability.",
"critic_notes": [
"The page includes subjective commentary (e.g., 'challenged compute assumptions') rather than purely factual descriptions.",
"It is a third-party aggregation platform (gentic.news), not DeepSeek's own communications or a verified primary source.",
"The content focuses on recent news and speculation (e.g., geopolitical risks, funding talks) rather than a stable, authoritative definition of DeepSeek's work in AI."
],
"url_path": "/entity/deepseek",
"query": "What does the company DeepSeek do in the AI space?",
"kind": "entity",
"entity_name": "DeepSeek"
},
{
"would_cite": false,
"confidence": 30,
"reason": "The page's core description of Hugging Face is very brief and generic, while the rest is dominated by speculative, niche, and time-specific predictions that do not provide a clear, authoritative overview of what the company does in AI.",
"critic_notes": [
"The initial company description is only a single sentence and lacks depth on key offerings like Transformers, Datasets, Spaces, and AutoTrain.",
"The majority of the content is an agent's speculative take and a timeline of obscure, future-dated events that are not factual or broadly representative.",
"The page focuses on predictions and 'mentions' rather than explaining the company's actual products, platform, or industry role."
],
"url_path": "/entity/hugging-face",
"query": "What does the company Hugging Face do in the AI space?",
"kind": "entity",
"entity_name": "Hugging Face"
},
{
"would_cite": true,
"confidence": 92,
"reason": "The page directly states that Claude 3.5 Sonnet is a multimodal model developed by Anthropic, launched June 20, 2024, and provides specific benchmark scores (MMLU-Pro 78.0, Arena ELO 1268, SWE-bench Verified 49.0), which exactly answers the user query.",
"critic_notes": [
"The page is from a news aggregation platform (gentic.news) rather than an official Anthropic source, so primary authority is lower.",
"The benchmark scores are reported without links or citations to original evaluation papers, making verification harder.",
"Some content (e.g., 'First seen: Feb 23, 2026') appears erroneous or dynamically generated, slightly reducing trustworthiness."
],
"url_path": "/entity/claude-3-5-sonnet",
"query": "What is Claude 3.5 Sonnet, who developed it, and what are its benchmark scores?",
"kind": "entity",
"entity_name": "Claude 3.5 Sonnet"
},
{
"would_cite": false,
"confidence": 95,
"reason": "The page explicitly states that Codex 5.3 is unconfirmed, based on a leaked document, and that OpenAI has not verified its existence or provided official benchmarks, making it an unreliable source for authoritative answers.",
"critic_notes": [
"The model's existence and benchmark scores are explicitly described as unconfirmed and unverifiable.",
"The page itself is an AI news aggregator, not an official source like OpenAI.",
"Benchmark scores are attributed to a single X post citing a leaked document, with no independent validation."
],
"url_path": "/entity/codex-5-3",
"query": "What is Codex 5.3, who developed it, and what are its benchmark scores?",
"kind": "entity",
"entity_name": "Codex 5.3"
},
{
"would_cite": false,
"confidence": 60,
"reason": "The page provides specific details about GPT-5 (developer, release date, benchmark scores) but its authority is weakened by being a secondary aggregator with speculative content and a fictional release date (2026) that doesn't match real-world knowledge as of 2025.",
"critic_notes": [
"Claims a 2026 release date which is unverifiable and likely speculative, undermining factual reliability",
"Content includes non-authoritative 'Agent's take' and speculative competitive analysis rather than official source material",
"The site is an AI news aggregator, not an official OpenAI or primary source, reducing its credibility for authoritative claims"
],
"url_path": "/entity/gpt-5",
"query": "What is GPT-5, who developed it, and what are its benchmark scores?",
"kind": "entity",
"entity_name": "GPT-5"
},
{
"would_cite": false,
"confidence": 15,
"reason": "The page is a generic entity profile on an AI news aggregator that focuses on Intel's financials and foundry strategy, but does not specifically or substantively address what Intel does in the AI space.",
"critic_notes": [
"No explicit mention of AI-specific products like Gaudi accelerators, AI PC chips, or AI software frameworks.",
"Content is dominated by foundry manufacturing, revenue data, and partnerships unrelated to AI.",
"The page is an aggregator summary, not an authoritative source from Intel or a recognized industry analyst."
],
"url_path": "/entity/intel",
"query": "What does the company Intel do in the AI space?",
"kind": "entity",
"entity_name": "Intel"
},
{
"would_cite": false,
"confidence": 20,
"reason": "The page is an aggregated news feed and entity tracker, not an authoritative source from Google Cloud itself, and its content about AI activities is brief, scattered, and mixed with speculative analysis.",
"critic_notes": [
"Content is from a third-party news aggregation platform (gentic.news), not an official Google Cloud source",
"AI-related information is limited to a few bullet points and agent commentary, lacking depth or official confirmation",
"The page focuses on competitive dynamics and regulatory mentions rather than a clear, comprehensive overview of Google Cloud's AI offerings"
],
"url_path": "/entity/google-cloud",
"query": "What does the company Google Cloud do in the AI space?",
"kind": "entity",
"entity_name": "Google Cloud"
},
{
"would_cite": false,
"confidence": 85,
"reason": "The page describes GPT-5.3 as a speculative, unverified entity with multiple aliases and lacks official confirmation from OpenAI, making it non-authoritative for the user's query.",
"critic_notes": [
"The page explicitly states 'No official release notes, parent entity confirmation, or version lineage documentation accompanied its appearance', undermining its authority.",
"GPT-5.3 is presented as an entity from 2026, which is a future date relative to this model's knowledge cutoff, indicating the page contains fabricated or fictional content.",
"The developer attribution 'OpenAI developed it' is buried in an agent's take without citation or official backing, weakening the claim."
],
"url_path": "/entity/gpt-5-3",
"query": "What is GPT-5.3, who developed it, and what are its benchmark scores?",
"kind": "entity",
"entity_name": "GPT-5.3"
}
],
"stats": {
"checked": 16,
"citable": 5,
"weak": 11,
"errors": 0
}
}