GEO GEOHub 0.3.1 · Experimental

GEO-first · SEO-active · Protocol-first

GEOHub

GEOHub organizes GEO research, diagnosis, content, measurement, and one-line SEO into a routable, reproducible, auditable super skill, with one entry point across six active capabilities

Registry-driven Traceable evidence Artifact Bus 1.0.0 Bilingual routing
Unified routing
Artifact Bus
Discover
Opportunities
Diagnose
Readiness
Content
Generation
Measure
Measurement
SEO
Planning
Research
Evidence

Six active skills share one Registry, Evidence Ledger, and artifact protocol

6
active Skill
3
planned Skill
56
Research sources
7
Content modes
10
Deterministic packages
1.0.0
Artifact Bus protocol
01 · POSITIONING

An operating system forGEO workflows!

It turns natural-language requests into the smallest executable skill or a stable workflow, then records every run with evidence status, a quality report, and a manifest

GEOHub turns natural-language requests into five capability entry points:discover, diagnose, content, measure, and one-line SEO. The unified entry resolves intent, the Research Kernel governs scientific boundaries, and the Artifact Bus supports delivery and replay

GEO and SEO share query, evidence, entity, content, and site-structure capabilities. The dedicated seo provider handles technical audits, keyword-to-page mapping, Search Console incidents, migrations, and experiment plans
6 / 9Currently executable
Active:geo、discover、diagnose、content、measure、seo6
Planned:strategy、knowledge、publish3
Protocol compatibility1.0.0
01

Registry-driven

A single Registry declares skill status, intents, entries, and outputs. The Router compiles its Action and Intent indexes from that source

02

Deterministic execution

Inputs pass strict schema, length, and file-boundary validation. Equivalent valid inputs produce stable structures without implicit web research

03

Traceable evidence

Factual claims link to evidence IDs. Missing sources remain unverified, source_gap, or blocked-by-evidence

04

Research boundaries

Fifty-four papers and two datasets inform 11 runtime principles with explicit correlation, benchmark, causal, and governance status

05

Atomic delivery

The complete file set is generated and validated in a hidden staging directory, then published in one operation so failed runs remain invisible

06

Explicit degradation

If optional DOCX or PDF rendering fails, Markdown, JSON, and HTML remain available while the manifest records missing dependencies and renderer errors

02 · ARCHITECTURE

Five layers turn each request into anauditable run record

The architecture separates natural-language selection, deterministic execution, research constraints, and delivery protocols while keeping shared data structures compatible

FIGURE 01 · SYSTEM ARCHITECTURE
Five-layer super skill architecture A user request passes through the unified entry, Registry resolver, six executable skills, Research Evidence Kernel, and Artifact Bus to produce multiple output formats INPUT User GEO request Chinese / English RESOLVE Unified Router Registry · Intent · DAG EXECUTE Discover Diagnose Content Measure SEO DELIVER Artifact Bus Atomic run directory JSON · Markdown · HTML RESEARCH EVIDENCE KERNEL 56-source registry 11 research principles Evidence Ledger Research Context Causal · Scope · Limitations OUTPUT SURFACES Structured JSONSchema validated Markdown reportHuman readable Standalone HTMLEscaped local render Quality reportWarnings and checks Run manifestFiles · Status · Degradation
Figure 1. The Research Evidence Kernel governs both execution and delivery. Each run records its research context, while repository documentation preserves the full audit. On mobile, scroll horizontally to view the architecture.
Keep selection lightweightThe Router returns only the skill, status, entry, and optional DAG. Execution logic stays inside each provider
Encapsulate executionDiscover, Diagnose, Content, Measure, and SEO each own an input contract and output specification
Preserve shared compatibilityThe Evidence Ledger, Research Context, Quality Report, and Run Manifest provide cross-skill interfaces
03 · ROUTING

The Router checksavailability boundariesbefore choosing the smallest execution unit

Inputs first pass character and byte budgets, Unicode normalization, quote-span detection, and negation-scope parsing. Planned capabilities always remain non-runnable

  1. Preflight validationLimits are 8,000 characters and 16,384 UTF-8 bytes. Oversized input returns a deterministic validation error
  2. Single semantic precomputationQuote spans, connectors, negation ranges, Action Index, and Intent Index are each scanned once per request
  3. Planned precedenceStrategy, knowledge, and publish matches return required_inputs, closest artifact, and nearest active capability
  4. Exact workflow matchingOnly Registry-declared stage combinations and orderings generate a DAG. Intent words inside explanatory text do not trigger workflows
  5. Smallest active skillA single goal routes directly to discover, diagnose, content, measure, or seo. Broad requests return to geo

brand-baseline-lite

DiscoverDiagnose
Use it to understand audience questions before auditing brand or website readiness

content-campaign

DiscoverContent
Use it to expand queries before producing titles, explainers, blueprints, or articles

composed workflow

DiscoverDiagnoseContent
When discovery, diagnosis, and content are explicitly requested, Discover becomes the shared upstream stage

planned boundary

StrategyKnowledgePublish
Returns status=planned, runnable=false, and entry=null without simulating execution
04 · CAPABILITIES

Six active skills cover GEO workflows andone-line SEO

Each skill can be installed and executed independently or selected through the geo Router. The table summarizes inputs, core processing, and primary outputs

geo

ACTIVE · ROUTER
Unified entry and capability status surfaceHandles broad GEO requests, single-capability selection, planned states, and two stable DAGs
INPUT
Natural-language requests in English or Chinese
OUTPUT
skill_id、status、runnable、entry、reason、suggestion、workflow

geo-discover

ACTIVE · OFFLINE
Expands seed terms and business topics into question clusters and opportunity mapsGenerates learn, compare, evaluate, and act task forms, removes duplicates, and preserves parent-query lineage
INPUT
subject、seed_queries、audience、scenario、evidence
OUTPUT
query-map、opportunity-map、evidence-ledger、research-context

geo-diagnose

ACTIVE · BOUNDED NET
Diagnoses GEO readiness at brand, site, or page scopeAnalyzes title, description, canonical, headings, FAQ, JSON-LD, extractable structure, and evidence signals
INPUT
Explicit URL, HTML snapshot, or evidence with brand, site, or page scope
OUTPUT
diagnosis、funnel、report、remediation query-map、opportunity-map

geo-content

ACTIVE · 7 MODES
Generates traceable content from a structured brief and evidenceBuilds the mode object first, then derives content-spec, typed Markdown, escaped HTML, and evidence units
INPUT
Topic, mode, brand facts, evidence, and optional source content
OUTPUT
content-spec, content, evidence units, Markdown, HTML, and optional DOCX or PDF

geo-measure

ACTIVE · OFFLINE
Summarizes user-provided AI platform observationsPreserves eligible trials, missing answers, exclusion reasons, platform strata, and Wilson intervals under a descriptive label
INPUT
Query, engine, interface, locale, timestamp, answer, and citation status
OUTPUT
answer rate、citation rate、conditional rate、interval、platform strata

seo

ACTIVE · OFFLINE PLAN
Turns a one-line SEO request into an evidence-constrained execution planRecognizes technical audits, keyword-to-page mapping, traffic incidents, migrations, experiments, international SEO, ecommerce, and implementation requests while preserving evidence and authorization gaps
INPUT
request、target URLs、market、language、authorization、rollback、evidence
OUTPUT
seo-plan、report、evidence-ledger、quality-report、run-manifest
01 / titleTitlesDiverse candidates with intent and compliance scoring, removing unsupported absolutes and generated years
02 / explainerExplainerDefinitions, rationale, steps, selection criteria, misconceptions, FAQ, and sources
03 / comparisonNeutral comparisonRequires a complete entity × dimension evidence matrix, with gaps blocking conclusions
04 / rankingMethod-based rankingRequires a complete score matrix and transparent method. Incomplete data produces no ranking
05 / page-blueprintPage blueprintInformation architecture, summary, FAQ, tables, schema candidates, and acceptance checks
06 / refineContent refinementPreserves source facts and records before and after scores plus change notes without adding data
07 / article-friendlyArticle-friendly adaptationReuses the refine contract with publication-ready Markdown, evidence markers, and risk notices
05 · SCENARIOS

The current release coversfourteen GEO and SEO tasks

Scenarios are derived from the Registry, two stable DAGs, seven Content modes, and one-line SEO work modes. Every scenario traces to an active entry or an explicit planned state

geo-discover

Seed expansion and question clusters

Generate learn, compare, evaluate, and act questions from a product, service, or category seed

Delivers: query-map + opportunity-map
geo-discover

Audience intent and content opportunities

Organize FAQ, explainer, comparison, and landing-page opportunities by audience, scenario, and evidence gap

Delivers: priority, asset type, and evidence tasks
geo-diagnose

Brand panorama diagnosis

Check fact coverage for identity, offering, audience, differentiation, proof, and contact

Delivers: brand findings + remediation
geo-diagnose

Site and page audit

Check discoverability, structure, extractability, evidence, authority, and freshness proxy signals

Delivers: diagnosis funnel + HTML snapshot
geo-content

Titles and explainers

Generate compliant titles, definitions, steps, misconceptions, FAQ, and source sections

Delivers: content.md + content.html
geo-content

Product comparisons and rankings

Generate neutral comparisons from aligned evidence or transparent rankings from complete score matrices

Delivers: matrix gaps or evidence-backed conclusions
geo-content

Page information architecture

Design summaries, FAQ, tables, semantic HTML, schema candidates, CMS fields, and acceptance checklists

Delivers: page blueprint + acceptance checks
geo-content

Existing content refinement

Preserve source fact boundaries while improving structure and expression. Unlinked facts remain unverified

Delivers: before / after + change notes
geo-measure

AI answer and citation measurement

Summarize answer rate, citation rate, conditional citation rate, missing answers, exclusion reasons, and uncertainty intervals

Delivers: measurement-report.json
seo

One-line SEO planning

Recognize technical audits, keyword-to-page mapping, international SEO, ecommerce, and experiment modes from natural language

Delivers: seo-plan + evidence gaps
seo

Traffic incidents and migration control

Stratify by page, query, market, device, and time while recording migration maps, validation, and rollback boundaries

Delivers: segmented plan + rollback gate
brand-baseline-lite

Brand baseline workflow

Discover audience questions, then audit how well the brand, site, or page addresses them

Flow: Discover → Diagnose
content-campaign

Content campaign workflow

Generate a query and opportunity map, then turn selected opportunities into titles, explainers, blueprints, or articles

Flow: Discover → Content
geo · planned

Strategy, knowledge, and publishing requests

Return required inputs, the nearest active capability, and the closest deliverable artifact

Status: planned · runnable=false
06 · PROMPTS

A prompt only needs a cleargoal, input, and evidence boundaryto trigger the right capability

These examples come from Registry intents and the fixed routing contract. Add your topic, audience, URL, or evidence file after copying

Query expansion
Expand the seed “GEO optimization company” into AI-search intents for marketing leaders evaluating service providers.
geo-discover · active
Brand diagnosis
Run a brand diagnosis using the official website URL and brand facts I provide, then list every evidence gap.
geo-diagnose · scope=brand
Page audit
Audit this page for GEO readiness, including its title, description, FAQ, structured data, and extractable evidence.
geo-diagnose · scope=page
Explainer
Using only the provided evidence, write an explainer on how to choose a GEO optimization company and preserve every factual source.
geo-content · mode=explainer
Neutral comparison
Compare Company A and Company B across service scope, method, deliverables, and evidence. List missing evidence for every incomplete dimension.
geo-content · mode=comparison
Method-based ranking
Use my evaluation dimensions and complete candidate scores to generate a GEO tool ranking with a transparent calculation method.
geo-content · mode=ranking
Page blueprint
Design a page blueprint for enterprise GEO services with a summary, FAQ, comparison table, schema candidates, and acceptance checklist.
geo-content · mode=page-blueprint
Content refinement
Improve the GEO readability of this article, preserve its facts, add no data or citations, and explain every change.
geo-content · mode=refine
Measure
Measure GEO results by summarizing answer rate, citation rate, missing answers, and platform strata from these offline observations.
geo-measure · active
one-line SEO
One-line SEO: create a technical audit plan for this website with evidence gaps and a validation sequence.
seo · technical-audit
SEO incident
Analyze the organic traffic decline in this Search Console export and list competing hypotheses by page, query, market, and device.
seo · incident
Baseline workflow
Discover audience intents, then run a brand diagnosis.
brand-baseline-lite
Content workflow
Expand the query set, then generate an article.
content-campaign
Composed workflow
Discover questions, audit our site, then write an explainer.
baseline + campaign
Roadmap
Create a GEO strategy roadmap to improve the brand’s evidence coverage in AI-search research.
geo-strategy · planned
07 · ARTIFACTS

Every execution produces aself-contained run directory

The Artifact Bus places input snapshots, structured results, human-readable reports, the Evidence Ledger, research boundaries, and quality status in one delivery unit

runs/run-<id>/
├── input/
│   ├── geo-brief.json
│   └── sources/...
├── primary-result.json
├── report.md
├── content.html
├── evidence-ledger.json
├── research-context.json
├── quality-report.json
└── run-manifest.json

# Added by each skill
query-map.json
opportunity-map.json
diagnosis-funnel.json
content-evidence-units.json
measurement-report.json
seo-plan.json
01
Normalized inputStrict JSON, Unicode, and path validation places replayable input under input/
02
Structured coreGenerate schema-valid objects first, then derive Markdown, HTML, and optional formats
03
Evidence and research contextEach factual unit links to evidence IDs, while Research Context records source scope and limitations
04
Quality and degradationQuality gates record warnings, failed checks, missing dependencies, and renderer errors
05
Atomic publicationPublish the final run directory after validating its staged file set. The Manifest becomes the delivery index
SkillCore JSONHuman-readable resultEvidence and qualitySpecial artifacts
geo-discoverquery-map、opportunity-mapStructured opportunity summaryledger、research-context、quality、manifestFour task forms with parent-query lineage
geo-diagnosediagnosis、diagnosis-funnelreport.mdledger、research-context、quality、manifestHTML snapshots、remediation maps
geo-contentcontent-spec、content、evidence-unitscontent.md、content.htmlledger、research-context、quality、manifestOptional DOCX / PDF
geo-measuremeasurement-reportreport.mdledger、research-context、quality、manifestDenominators, intervals, and platform strata
seoseo-planreport.mdledger、quality、manifestWork mode, evidence gaps, implementation authority, and rollback boundaries
08 · RESEARCH

Fifty-six sources inform11 runtime principles

Fifty-four papers and two datasets from a pinned research commit were audited individually. Runtime uses a compact Registry and emits research-context.json for every execution

56Pinned research sources
Papers54
Datasets2
paper-reported43
not-reproducible11
reproduced1
source-reported1
Outcome claims require direct measurementReadiness, visibility, and influence scores are proxy variables
Query diversity stays conditionalIt supports coverage planning without inferring demand or conversion
Diagnosis scores express readiness onlyScores do not become platform citation probabilities
Semantic relevance leads format tacticsContent fidelity and evidence links have stronger support
Comparisons require aligned evidenceIncomplete matrices retain gaps and produce no winner
Rankings need manipulation safeguardsTransparent methods, complete scores, and threat models work together
Structural edits must be reversibleStructure can improve or degrade, so before-and-after records are required
Measurement must preserve denominatorsRepeated trials, missing answers, intervals, and platform scope are reported together
Knowledge needs provenance and conflict controlRights, freshness, conflicts, and poisoning risk enter governance boundaries
SEO and GEO share upstream retrieval boundariesQueries and evidence can be reused while SERP and traffic metrics remain separate
Publishing needs authorization and reversibilityPublic changes require permission, preview, records, and rollback
Research boundary: current materials support planning hypotheses, descriptive measurement, and governance controls. They provide no universal guarantee for cross-platform citation, ranking, traffic, conversion, or revenue. Read the full audit in geo-citation-lab-alignment-audit.md, with the machine-readable matrix in research-evidence-matrix.json
09 · QUALITY

Engineering gates verify behavioral consistency andkeep outcome experiments explicit

The local 0.3.1 snapshot passes the full test suite, fixed evaluations, cross-Python checks, deterministic packaging, isolated installation, and yao-meta Library gates

Router precision100%
Router recall100%
Trigger compliance100%
Output contract100%
Fabricated citations0
The fixed evaluation set contains 374 routing cases, 40 trigger cases, and 30 deterministic output cases
540Full tests passed, with one optional renderer environment case skipped
3.11-3.14Four Python minor versions verified in fresh virtual environments
10 / 9Ten deterministic ZIPs and nine isolated non-source package installs
79 / 15 / 0yao-meta commands / explicit waivers / release blockers

Available now

  • Offline question expansion and evidence-constrained opportunity discovery
  • Brand, site, and page diagnosis from explicit URLs or HTML
  • Seven content modes with optional rich-format rendering
  • Descriptive statistics and Wilson intervals for offline observations
  • One-line SEO mode detection, evidence gaps, and replayable action plans
  • English and Chinese routing, stable DAGs, and planned-state guidance

Current boundaries

  • No AI-platform account connections or continuous sampling
  • No inference of search volume, keyword difficulty, or SERP rank
  • No guarantee of citations, ranking, traffic, conversion, or revenue
  • No CMS publishing, knowledge-base writes, or external changes
  • Diagnose fetches only explicit public URLs without crawl expansion
geo-strategyRequires subject, goals, and constraints; nearest current artifact: opportunity-map
geo-knowledgeRequires subject and approved_sources; nearest current artifact: brand-fact-card
geo-publishRequires content_artifact and target_channel; nearest current artifact: run-manifest