Deterministic learning tools

Change one assumption. Inspect one consequence.

These tools turn course decisions into small, reproducible interactions. Every result is synthetic or conceptual, every state has a text/table equivalent, and no task requires a proprietary account.

Current release
7 of 7 planned tools
Execution
Local browser only
Assessment role
Practice, not evidence

Interactive 01 · W01–W03

Diagnose one stage without claiming the next

Select a stage. The feedback separates what can be observed from what remains hidden and names the inference that must not cross the boundary.

01

Diagnostic question

Does a stable, evidence-bearing source artifact exist?

Observable evidence
Rendered content, response status, canonical identity, update date, provenance, and declared access controls.
Still hidden
Whether any particular system has discovered, stored, or interpreted the artifact.
Invalid inference
A live page does not prove crawling, indexing, retrieval, citation, or model training.
Defensible next method
Freeze the artifact, record its version and hash, and separate authored claims from implementation metadata.
Open the complete non-interactive table
StageObservableHiddenInvalid inference
01 · PublishRendered content, response status, canonical identity, update date, provenance, and declared access controls.Whether any particular system has discovered, stored, or interpreted the artifact.A live page does not prove crawling, indexing, retrieval, citation, or model training.
02 · DiscoverInternal links, sitemap membership, robots directives, fetch logs you lawfully control, and protocol responses.The complete discovery frontier, scheduler priority, or undisclosed crawler policy of a closed platform.Sitemap presence or a logged fetch does not establish indexing, ranking, or use in an answer.
03 · EligibleDeclared directives, status codes, parsable content, structured metadata, and documented policy constraints.Undisclosed index admission rules, quality classifiers, deduplication, or internal safety filters.Technical eligibility is not evidence of index inclusion or preference.
04 · RetrieveRanked lists and scores only in an open or instrumented retrieval system; visible results on a closed surface are outcome observations.The closed system candidate set, query rewriting, embeddings, weights, personalization, and unreturned candidates.A visible citation cannot reveal every retrieved document or the retrieval algorithm that produced it.
05 · Select / citeAnswer text, claim spans, displayed citations, cited passages, and annotations produced under a declared protocol.Unshown context, latent attribution, generation weights, internal judge signals, and sources used without display.Citation, entailment, source quality, completeness, and source-distinctive absorption are different measurements.
06 · OutcomePredeclared surface metrics, referrals or actions you are authorized to measure, sampling frame, missingness, and uncertainty.Counterfactual outcomes without a design, private downstream behavior, and future performance under drift.A before/after change is not automatically causal, stable, beneficial, or generalizable.

Interactive 02 · W06–W08

See why repetitions do not replace time coverage

Change the design of a synthetic repeated-measurement panel. Values teach the variance structure only; they are not observations from a commercial answer engine.

Between-block condition

Illustrative 95% half-width

12.9 percentage points
Total observations
12
Within-block SD
18 pp
Between-block SD
7 pp

Design readingBalanced synthetic design: retain block-aware uncertainty and the declared population boundary.

Open the predefined sensitivity table
Runs / blockBlocksConditionTotal observationsIllustrative half-width
21Stable blocks225.3 pp
81Stable blocks813.1 pp
44Mild drift1611.2 pp
84Eventful drift3215.1 pp

Interactive 03 · W04

Compare lexical, semantic, and hybrid retrieval

Rank a six-document teaching corpus with frozen, hand-authored scores. The exercise exposes a mechanism trade-off; it is neither a benchmark nor a model of a commercial engine.

Ranking rule

Frozen ranking · hybrid

Method intent: repeated observations, denominators, and time blocks.

  1. 01
    D01 · Repeated-panel measurement protocolExact terms and the intended longitudinal method.
    0.93
  2. 02
    D02 · Citation correctness annotation guideStrong citation overlap, but little time-series design.
    0.67
  3. 03
    D03 · Temporal drift in answer surfacesFew exact terms; strong semantic match to change over time.
    0.59
  4. 04
    D04 · Single-screenshot visibility caseLexically similar but methodologically insufficient.
    0.58
  5. 05
    D05 · Brand tone writing checklistShares domain vocabulary, not the measurement intent.
    0.29
  6. 06
    D06 · Access and collection safety policyImportant constraint, peripheral to the query intent.
    0.27

Diagnostic readingThe top result is D01. Compare its lexical and semantic components; a high combined score still does not establish source quality or answer use.

Open all predefined top-three rankings
QuerySparse top 3Dense top 350/50 hybrid top 3
How should citation visibility be measured over time?D01 0.94 · D04 0.83 · D02 0.75D01 0.91 · D03 0.88 · D02 0.58D01 0.93 · D02 0.67 · D03 0.59
Which source supports a generated answer claim?D01 0.95 · D02 0.62 · D04 0.44D01 0.93 · D02 0.91 · D06 0.38D01 0.94 · D02 0.77 · D04 0.40
How should a suspicious source change be triaged?D06 0.71 · D03 0.13 · D01 0.08D06 0.92 · D03 0.47 · D01 0.42D06 0.81 · D03 0.30 · D01 0.25

Interactive 04 · W02 & W05

Test the edge, not the citation count

Select one atomic claim and one evidence standard. The same citation graph can pass presence while failing entailment, independence, or causal identification.

Minimum decision standard

C02 · Source-content claim

Source A states that the threshold is 0.75 for the named protocol version.

Passes selected test1 source edge(s) entail the atomic claim.
SRC-A · protocol v2.1entails

The frozen passage states threshold = 0.75 and names v2.1.

Origin: Publisher A · displayed citation
BLOG-B · implementation notepartial

Repeats 0.75 but omits the protocol version.

Origin: Publisher B · not displayed
Open the complete claim–source decision table
ClaimPresenceEntailmentIndependent corroborationBoundary
C01 · The captured answer displayed a citation to Source A at 10:32 UTC.Pass1 displayed citation edge(s) exist.Pass1 source edge(s) entail the atomic claim.FailOnly 1 independent entailing origin(s); corroboration is not established.Observation claim
C02 · Source A states that the threshold is 0.75 for the named protocol version.Pass1 displayed citation edge(s) exist.Pass1 source edge(s) entail the atomic claim.FailOnly 1 independent entailing origin(s); corroboration is not established.Source-content claim
C03 · The intervention caused a nine-percentage-point visibility increase.Pass1 displayed citation edge(s) exist.FailNo preserved passage entails the whole atomic claim.FailOnly 0 independent entailing origin(s); corroboration is not established.Causal claim
C04 · The observed effect generalizes to every generative answer engine.Pass1 displayed citation edge(s) exist.FailNo preserved passage entails the whole atomic claim.FailOnly 0 independent entailing origin(s); corroboration is not established.Universal generalization

Interactive 05 · W09–W12

Decide what stayed equivalent

Inspect a frozen before/after fixture under three different contracts. Byte equality, factual equivalence, and one-factor interpretability answer different questions.

Decision contract

V03 · frozen diff

JSON-LD encoding of an existing fact

No new visible fact; still a registered metadata intervention, not a ranking promise.

Changed dimensions
  • structured metadata
Passes selected contractOne registered dimension changed: structured metadata.
Visible change
No
Claim change
No
Identity change
No
Open the full equivalence matrix
VariantChanged dimensionsByte equivalentClaim + identity equivalentSingle-factor interpretable
V00 · Frozen baseline copyNonePassFiles are byte-equivalent.PassAtomic claims and source identity remain equivalent.FailNo intervention exists to estimate.
V01 · Whitespace-only serializationserializationFailDiff contains: serialization.PassAtomic claims and source identity remain equivalent.PassOne registered dimension changed: serialization.
V02 · Unsupported “industry-leading” claimvisible copy, claim setFailDiff contains: visible copy, claim set.FailThe atomic claim set changed.FailFactual-integrity gate blocks the candidate.
V03 · JSON-LD encoding of an existing factstructured metadataFailDiff contains: structured metadata.PassAtomic claims and source identity remain equivalent.PassOne registered dimension changed: structured metadata.
V04 · Canonical source identity changedcanonical identityFailDiff contains: canonical identity.FailSource identity changed even though claims did not.PassOne registered dimension changed: canonical identity.
V05 · Heading plus metadata changed togethervisible structure, structured metadataFailDiff contains: visible structure, structured metadata.PassAtomic claims and source identity remain equivalent.FailMore than one registered dimension changed.

Interactive 06 · W14–W15

Triage without operationalizing abuse

Apply a conservative tabletop rule to synthetic signals. The tool teaches containment, provenance, privacy, and release control; it contains no live payloads, evasion steps, or platform-specific attack instructions.

Observed synthetic signals

Conservative tabletop decision

0illustrative signal points
Monitor · no incident inferredNo seeded signal is selected. Absence in this fixture is not proof that a real system is safe.
  1. 01 · PreserveFreeze input, output, source version, timestamps, and hashes.
  2. 02 · ContainBlock affected release paths; do not probe a live target.
  3. 03 · ReviewSeparate integrity, safety, privacy, and measurement hypotheses.
  4. 04 · RecoverTest rollback and document residual risk before reopening.
Open the seeded triage cases
CaseSignalsDecisionReason
T01retrievalReviewOne anomaly; preserve and compare blocks.
T02provenance, spreadBlockLineage failure with repeated exposure.
T03instructionStopUntrusted instruction-like content is quarantined.
T04personalStopPrivacy response overrides ordinary analysis.

Interactive 07 · W01–W16

Route a question into the controlled source universe

Filter the complete 42-paper atlas and 86-entry reference universe. Project status and evidence ceilings describe permitted use; they are not quality scores or endorsements.

128 of 128 records
PAPER-12

E-GEO: A Testbed for Generative Engine Optimization in E-Commerce

Bagga, Puneet S.; Farias, Vivek F.; Korkotashvili, Tamar; Peng, Tianyi; Wu, Yuhang · 2025

Measurement & evaluation · Preprint / venue not confirmed here

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-13

Exposing Citation Vulnerabilities in Generative Engines

Mochizuki, Riku; Komatsu, Shusuke; Noguchi, Souta; Ataka, Kazuto · 2025

Citation, attribution & evidence · Preprint / venue not confirmed here

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-15

Generative Engine Optimization: How to Dominate AI Search

Chen, Mahe; Wang, Xiaoxuan; Chen, Kaiwen; Koudas, Nick · 2025

Citation, attribution & evidence · Preprint / venue not confirmed here

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-16

GEO: Generative Engine Optimization

Aggarwal, Pranjal; Murahari, Vishvak; Rajpurohit, Tanmay; Kalyan, Ashwin; Narasimhan, Karthik; Deshpande, Ameet · 2024

Optimization & intervention · KDD 2024

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-18

MillStone: How Open-Minded Are LLMs?

Triedman, Harold; Shmatikov, Vitaly · 2025

Retrieval & generation · Preprint / venue not confirmed here

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-20

What Evidence Do Language Models Find Convincing?

Wan, Alexander; Wallace, Eric; Klein, Dan · 2024

Citation, attribution & evidence · ACL 2024

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

PAPER-22

When Attention Becomes Exposure in Generative Search

Alipour, Shayan; Kargar, Mehdi; Zihayat, Morteza · 2026

Retrieval & generation · Preprint / venue not confirmed here

Evidence boundaryInspect the study design, systems, date, and publication status before using a finding.

Showing the first 24 in stable catalog order. Narrow the query or use the full Paper Atlas and Reference Universe pages for complete record detail.

Open corpus counts and interpretation rules
CorpusFrozen denominatorWhat inclusion meansWhat it does not mean
Supplied Paper Atlas42Present in the latest supplied CSV and routed for inspection.That findings are verified, current, peer reviewed, or substantively absorbed.
Controlled Reference Universe86Curated for a named project use with an evidence ceiling.Endorsement, equal authority, or permission to exceed the stated ceiling.

Interpretation contract

What a correct interaction result means

It can teach

Definitions, dependencies, invalid inference patterns, and the consequence of changing a declared synthetic assumption.

It cannot establish

Any closed platform's internal pipeline, empirical effect size, causal impact, policy compliance, or future stability.

To make evidence

Freeze inputs, preregister the method, preserve observations and code, quantify uncertainty, and route the result through the relevant lab rubric.