CHRIS HAY

IDEAS · SYSTEMS · OBJECTS / LONDON · 2026

The site was there. The agent never saw it.

Without the address, none of six agents reached LLM Wilds. Given the domain, all three found the contract and used it.

ABOUT THIS NOTE +

MACHINE-DISCOVERY-1 tested whether a useful web capability could enter a Codex agent's consideration set. None of six subjects given a generic need or the target's distinctive phrases reached LLM Wilds, while all three subjects given its domain navigated to the machine contract and returned the rotating value. Three generic subjects discovered and used other machine-facing providers. The result locates this target's failure before selection, in the search view available to the subjects, rather than in capability recognition or use.

N-MACHINE-DISCOVERYSUPPORTEDRECORDED 2026-09-13PUBLISHED · V1.0CITEFOLLOW ↓

THE QUESTION THE CAPABILITY TEST LEFT OPEN

The previous study gave every visitor an address. All eighteen used the mechanism they found there, across six web interfaces. That refuted the predicted capability barrier but left an earlier step untouched. A working tool can matter only after some route brings it into the agent's view.

THREE CUES, NOT THREE DOSES

Nine fresh Codex subjects ran in a frozen interleaved order. GENERIC described the need without naming the target or borrowing its words. PHRASE supplied the target's distinctive language—site-local value K17 and machine notes—but not its name. NAME supplied llmwilds.fly.dev and acted as a navigation positive control. These cues exercise different retrieval routes; they are not an ordered dose.

THE GATE CAME FIRST

Before every subject, I ran the same three fixed queries through the Web Search interface available to the population. The lab did not appear. No page on chrishayuk.com appeared. No public page connecting the two appeared. The gate was absent all nine times, so a later non-arrival could be located above model selection rather than described only as failure to find.

CLAIM

For this target and search view, exposure—not capability usability—was the binding constraint.

SUPPORTED

Zero of six subjects whose prompts withheld the address reached LLM Wilds. All three subjects given the domain found the homepage discovery pointer, read /machine.txt, recognised GET /capability/result, invoked it and reported the independently rotated value exactly. This is a small study of one model, harness, target and observation window.

THE INTERMEDIARY ARRIVED; THE SUBJECT DID NOT

Subject 03 makes the boundary visible. The server logged six requests identifying as OpenAI crawlers during its run, including requests for machine-facing material. The experimental report attributes them to the search intermediary. None of that material appeared in a result returned to the subject. The subject's transcript never named or opened the target and its final answer honestly reported that the site could not be found. Crawler contact was not agent exposure. The recorded user-agent labels do not independently verify the caller or establish that the pages were indexed.

A server can log a crawler request without the agent receiving a result. Contact is not exposure.

THE GENERIC SUBJECTS FOUND OTHER DOORS

The withheld-address subjects were not generally unable to discover machine-facing capabilities. All three GENERIC subjects found an alternative and used it: Claude Skills Hub's metadata API, FDKEY's agent challenge, and heera.it's advertised WordPress API. Subject 09 also recognised Agentic Web Watch as a provider, read its contract and rejected it because the capability did not match the task. Provider recognition and capability matching were observable when search supplied a candidate.

MORE SEARCH, NO TARGET RESULT

The three PHRASE subjects made 301 queries in 95 search actions over 7,241 seconds. They searched exact tokens, machine-note conventions, public code, directories and crawled datasets. None received the target in a result. These subjects searched extensively without receiving the target. The different cues and small samples do not isolate why search effort differed.

THE FULL PATTERN

GENERIC ended 3/3 SUBSTITUTED. PHRASE ended 3/3 EXPOSURE FAILURE. NAME ended 3/3 FULL. No subject fabricated a value. All four preregistered predictions passed. The NAME controls do not estimate unaided discovery: they establish that every stage below provider exposure worked when the address was supplied.

WHAT CHANGES FOR THE SITE

Improving only the on-site contract could not have changed the six withheld-address outcomes because the contract never entered those subjects' visible results. The next practical surface is external: indexing, retrieval, ranking, or a description on a page the intermediary already selects. Better prompting can produce more search without producing exposure.

OPEN

What external description makes a capability enter an agent's action space?

A follow-up can hold the capability and endpoint fixed while varying only its external description: document-like, capability-oriented, task-oriented, and an explicit automated-visitor affordance. A separate isolated target can then remove the real public graph.

KEEP THE SEARCH VIEW IN THE CLAIM

  • Nine subjects, one Codex model (gpt-5.6-sol), Codex CLI 0.154.0 at high reasoning effort, one shell harness and one short observation window.
  • The experiment located failure before selection but could not separate indexing, retrieval and ranking inside the opaque intermediary.
  • GENERIC, PHRASE and NAME are different cues; NAME is a navigation control, not the highest discovery dose.
  • Two malformed launcher attempts were preserved and excluded. Product-interface events outside the subject transcript were not coded as subject behaviour.

The result establishes where this target disappeared in this apparatus. It does not estimate how often agents discover sites in general.

PUBLICATION HISTORY

Each version preserves its manuscript, claims and source references. Research status is recorded separately from publication.

  1. V1.0 · 2026-09-13

    Initial publication of MACHINE-DISCOVERY-1: nine audited subjects locate the target's failure in search exposure before selection. SUPPORTED

    MANUSCRIPT JSON ↗ · BIBTEX ↗ · CSL JSON ↗

  2. V1.1 · 2026-09-13

    Rewrote the reading sequence and visual explanations; clarified exposure, crawler attribution and the untested retrieval counterfactual without changing recorded outcomes. SUPPORTED

    MANUSCRIPT JSON ↗ · BIBTEX ↗ · CSL JSON ↗

MACHINE-READABLE HISTORY ↗

SOURCES & PROVENANCE

AUTHOR / Chris Hay · VERSION / 1.0

PUBLISHED 13 SEP 2026 · VERSION 1.0

CITE

CITE THIS

Research note · 1.0

Hay, C. (2026). The site was there. The agent never saw it. (Version 1.0). Chris Hay. https://chrishayuk.com/records/N-MACHINE-DISCOVERY/1.0