DETERMINISTIC MATHEMATICAL SPECIFICATION5 Pillars • 18 Modules

How AI Search Scores Are Calculated.

The exact mathematical formula, weight distributions, and deterministic probe logic powering the Composite Generative Index (0–100). Zero LLM hallucinations, zero probabilistic drift.

The Master Composite Generative Index Equation
Range: [0.0, 100.0] • Deterministic SHA-256
Formal Linear Combination
S = 0.25·P_crawlers + 0.20·P_entity + 0.20·P_schema + 0.20·P_tech + 0.15·P_llms

Where each pillar parameter \(P_i\) is evaluated independently on a scale of 0 to 100 via deterministic AST parsing, RFC 9309 rule compilation, network socket latency benchmarking, and Schema.org graph traversal.

CRAWLERS25%
ENTITY20%
SCHEMA20%
EDGE & WAF20%
LLMS.TXT15%

The 5 Diagnostic Pillars: Scoring Criteria & Rules

Each pillar evaluates a mission-critical vector of synthetic retrieval. If any single pillar fails completely, generative visibility drops precipitously.

Frontier AI Crawler Accessibility

25%

Deterministic RFC 9309 path evaluation for all 11 frontier search bots. Evaluates live citation crawlers vs offline training scrapers.

ALGORITHM:P_crawlers = 100 * (AllowedCitationBots / TotalCitationBots) - Penalties
Audit Checklist:
OAI-SearchBot (ChatGPT Search) explicitly allowed (+20 pts)
PerplexityBot (Sonar Search) explicitly allowed (+20 pts)
Claude-SearchBot (Claude Web) explicitly allowed (+20 pts)
Applebot-Extended (Apple Intelligence) explicitly allowed (+20 pts)
Google-Extended (Gemini Grounding) explicitly allowed (+20 pts)

Entity Grounding & Knowledge Graph

20%

Unambiguous brand anchoring in global knowledge graphs to eliminate hallucination during multi-source synthetic retrieval.

ALGORITHM:P_entity = WikidataLink (40) + WikipediaLink (30) + SameAsArray (15) + KnowsAbout (15)
Audit Checklist:
Canonical Wikidata QID resolvable in sameAs (+40 pts)
Canonical Wikipedia entity URL linked (+30 pts)
Multi-platform entity federation (GitHub, LinkedIn, Crunchbase) (+15 pts)
knowsAbout subject domain ontology linked (+15 pts)

Structured Data & Schema Graph

20%

Syntactic validity and interconnected topology of JSON-LD metadata, enabling LLM parsers to map domain capabilities.

ALGORITHM:P_schema = ValidJsonLd (30) + ConnectedGraph (25) + CoreEntities (25) + ZeroErrors (20)
Audit Checklist:
W3C-compliant JSON-LD with zero syntax errors (+30 pts)
Unified Schema.org @graph interconnected nodes (+25 pts)
Primary Organization and WebSite entity definitions (+25 pts)
Article / Product / FAQPage domain-specific schemas (+20 pts)

Technical Edge Performance & WAF

20%

Edge response speed within AI retrieval timeout budgets (sub-120ms) and zero false-positive WAF challenges.

ALGORITHM:P_tech = TtfbScore (40) + Http3Quic (20) + ZeroWafChallenge (25) + ModernTls (15)
Audit Checklist:
Edge Time-To-First-Byte under 120 milliseconds (+40 pts)
Zero WAF JavaScript challenges or 403 blocks on AI user agents (+25 pts)
HTTP/3 (QUIC) protocol negotiation active (+20 pts)
TLS 1.3 handshake with HSTS preloaded (+15 pts)

Semantic Context & /llms.txt

15%

Zero-noise markdown context files designed specifically for ingestion by frontier generative model reasoning loops.

ALGORITHM:P_llms = EndpointExists (50) + HighSnrMarkdown (25) + CuratedLinks (25)
Audit Checklist:
Live, reachable /llms.txt endpoint returning HTTP 200 (+50 pts)
High-density, boilerplate-free markdown context (+25 pts)
Clean architectural links and API specifications (+25 pts)
RFC-compliant LLMs-Txt header or robots.txt directive (+20 bonus pts, capped at 100)

The 18 Deterministic Diagnostic Modules

Every audit executes 18 automated probes synchronously. The outputs are verified against published RFC protocols and cryptographic signatures.

IDDiagnostic ModuleCategoryDeterministic Probe MethodologyScoring Impact
M01Frontier Citation Bot ParserCrawlerRFC 9309 line-by-line deterministic evaluation+25% weight
M02Training Scraper IsolationCrawlerSeparation of commercial harvesters from citation botsPrevents false blocks
M03WAF Challenge DetectorSecuritySimulation of AI crawler User-Agent & ASN fingerprintingUp to -35 penalty
M04Edge TTFB Latency BenchmarkPerformanceSub-120ms millisecond socket handshake timing+40 pts in P_tech
M05HTTP/3 (QUIC) NegotiatorPerformanceALPN negotiation check for h3 protocol+20 pts in P_tech
M06TLS 1.3 & HSTS Preload CheckSecurityCipher suite audit and Strict-Transport-Security header+15 pts in P_tech
M07JSON-LD Syntax ValidatorSchemaW3C AST parser verifying bracket parity and keys+30 pts in P_schema
M08Schema.org @graph TopologySchemaEntity node connectivity via @id cross-referencing+25 pts in P_schema
M09Primary Organization GroundingSchemaPresence of canonical Organization metadata+25 pts in P_schema
M10Wikidata QID ResolutionEntityRegex scan for valid Q-identifier in sameAs URIs+40 pts in P_entity
M11Wikipedia Knowledge LinkageEntityCross-link detection to verified Wikipedia entry+30 pts in P_entity
M12Federated Entity MeshEntityPresence of GitHub, LinkedIn, and registry profiles+15 pts in P_entity
M13/llms.txt Endpoint ProbeContextDirect GET request to domain root for llms.txt standard+50 pts in P_llms
M14RAG Markdown SNR AnalysisContextSignal-to-noise token ratio calculation on markdown body+25 pts in P_llms
M15512-Token Semantic Chunk SizerContentRolling token-count segmentation testing window boundsGEO readiness
M16Princeton Quotation DensityGEOHeuristic detection of attributed expert quotes+41.5% citation lift
M17Verifiable Statistics RatioGEOQuantitative numerical density per 1,000 words+37.8% citation lift
M18SHA-256 Deterministic SealIntegrityCryptographic hash generated from canonical audit payloadAudit reproducibility
Anti-Fabrication & Deterministic Reproducibility Seal

Why Your Score Is Guaranteed to Be 100% Deterministic

Many AI audit tools call an LLM directly to “guess” whether a website is optimized, producing divergent scores on subsequent audits. AI Search Fixer operates differently:

1. Code-Only Ingestion

Raw HTTP headers, robots.txt tokens, and JSON-LD AST nodes are parsed strictly by deterministic TypeScript algorithms.

2. Zero Prompt Guessing

The LLM report writer is never permitted to calculate scores, infer crawler status, or modify numerical ratings.

3. SHA-256 Signature

Every audit payload is hashed into an immutable cryptographic seal, ensuring identical site configurations yield identical scores.

Interactive AI Search Ranking Probability Calculator

Adjust live parameters to see how your Composite Generative Index and citation probabilities shift in real-time.

AI Search Ranking Probability Engine

Unlike Google 10 blue links that rely on PageRank backlink graphs, generative answer engines synthesize citations using vector embedding similarity, RFC 9309 crawler permissions, and knowledge graph grounding.

Empirical 2026 RAG Scoring Model
Generative Ranking FactorPredicted Impact
RFC 9309 AI Crawler AllowlistProtocol Access

Authorizing OAI-SearchBot, PerplexityBot, and Claude-SearchBot in robots.txt allows direct real-time context ingestion.

+28% Retrieval Weight
Wikidata QID & Authority TriplesEntity Disambiguation

Canonical Schema.org @id and sameAs links disambiguate your brand from generic word homonyms in the global knowledge graph.

+22% Grounding Weight
ASF GEO Quotations & StatisticsEmpirical Density

Empirical research proves quotation clusters and statistical claims increase citation inclusion by up to 41.5%.

+18% Citation Lift
Direct Answer Capsules (40-60 Words)Chunk Extraction

Placing concise TL;DR definitions immediately below H2 headers enables instantaneous RAG vector chunk extraction.

+14% Zero-Shot Retrieval
Machine-Readable /llms.txt ManifestoAgent Discovery

Standardized Markdown documentation deployed at /llms.txt feeds clean, token-efficient summaries to reasoning models.

+10% Agent Discovery
Server-Side Rendering (Zero JS Blankness)Crawl Parity

Delivering full HTML without client-side JavaScript execution ensures non-headless AI crawlers never see empty pages.

+8% Ingestion Parity
Predicted Citation RateMulti-Engine AI Grounding
6 Heuristics
98%
Primary Cited Source #1
Projected Inclusion by Frontier Engine
ChatGPT Search (GPT-4o)99% Grounded
Perplexity Sonar 3.099% Cited
Google AI Overviews99% Included
Claude Web Grounding99% Verified

Grounding Telemetry & Citation Cockpit.

Synthesized vector retrieval instrumentation. Inspect live neural branches, sweep the holographic scanner beam, and toggle pre/post remediation states.

AI Citation Monolith v2.4 · Spatial Telemetry
Live Sync Active
Cinematic 3D Floating Telemetry Glass Monolith
Node TB4
Multi-Hop Citation Root0.984 Cosine Sim

Direct factual grounding linked across 4 frontier LLMs with multi-hop verification.

Model Waveform
Vector Hist
Coherence Radar
Citation Velocity
98.4%
+14.2% post-remediation
Allowed Crawlers
11 / 11
Frontier engines authorized
Wikidata Verification
Q183921
Canonical entity disambiguated
ASF GEO Boost
+41.0%
Quotation density verified

Ready to Audit Your Live Domain?

Receive your deterministic Composite Generative Index and 18-module technical diagnostic in 120ms.

Start Free Audit