- GEO — Generative Engine Optimization
- The practice of making a brand more likely to be cited, linked, recommended, or named in AI-generated answers — ChatGPT, Google AI Overviews, AI Mode, Perplexity, Gemini, Grok, Claude. GEO is to answer engines what SEO is to the ten blue links.
- Presence probability
- A calibrated estimate of the fraction of engine runs in which a given URL, brand, or product is surfaced for a given prompt. Expressed as a percentage and always reported with a 95% confidence interval. Never derived from a single run.
- Confidence interval (95% CI)
- The range [lo%, hi%] within which the true underlying probability is expected to fall in 95 out of 100 repetitions of the sampling procedure. A wide CI means more uncertainty; a narrow CI means more evidence. Always displayed alongside the point estimate — never hidden.
- Surface class
- The five levels of how a brand appears in an AI answer: cited (URL in a citation card), linked (inline hyperlink), recommended (recommendation language), named (prose mention only), and semantic (descriptor match without an explicit name). Each class has its own rate and CI.
- Share of voice (SoV)
- Lokrix AI computes share of voice as the presence-ratio SoV: p̃_target / (p̃_target + Σp̃_competitors), where p̃ is the Beta-shrunk presence probability per competitor. A position-weighted variant uses PAWC (1/log2(rank+1)) so a #1 citation is not treated the same as a #5 citation.
- Sequential Wilson stopping
- The adaptive sampling rule that controls how many draws Lokrix AI takes per (prompt × engine × cell). It samples while the Wilson 95% CI width exceeds 2ε (default ε=0.10), with a minimum of 8 and maximum of 60 draws. Decisive cells (presence probability > 0.8 or < 0.2) exit early to save cost. This concentrates draws exactly where uncertainty is high.
- Mode separation
- The rule that only ORGANIC draws (recommendation-elicitation, intent fan-out) enter the headline composite. DIAGNOSTIC-ELICITED draws (citation-forced, brand-named, comparison) are reported separately. Blending a diagnostic mode — where the brand is named in the prompt — into the organic presence probability is the cardinal statistical-honesty violation.
- Provenance
- A badge on every data point indicating the measurement method: "live" (a real-time engine run), "cached" (a stored result from a recent run), or "predicted" (the calibrated ML model). Live probabilities carry an "uncalibrated" flag until the Platt/Isotonic calibration map has been fit from historical predicted→realized outcomes. Provenance is always visible.