Reference Atlas

Follow the sources
Opening your graph…
Entire network
Web pagePDFUnresolved referenceArrows point to referenced documents

How the scraper works · coverage and limitations

The agent collects and publishes this network. Browsing, filtering, and opening evidence do not start scraping. The published snapshot loads on each visit; your selected document is remembered locally.

Incoming-source rankings count distinct source nodes, excluding navigation and self-links; they are local to the full published collection and do not change with display filters. OpenAlex counts cover its wider database. Larger nodes reflect the chosen ranking; citations measure attention, not research quality. How OpenAlex builds citations.

Starting from the SEI article, the scraper follows selected research links and parses HTML and PDF text. Each inspected document records its collection time, outgoing links, and extraction warnings. Uninspected neighbors retain only the label observed in their source.

Article hyperlinks, reference-section links, navigation links, embedded PDF links, DOI mentions, and bibliography candidates are separate edge types. Connection evidence preserves the observed label or bibliography text and PDF page where available. A hyperlink does not establish agreement or a semantic relationship.

PDF bibliography splitting is heuristic. Entries without an exact URL or DOI remain unresolved; they are not automatically merged with similarly titled papers. Scanned PDFs require a text layer; OCR is not included.

Fetches honor robots.txt, validate every redirect and public destination, and never execute source JavaScript. Each document is limited to 12 MB, 120 PDF pages, 100 connections, and 45 seconds of processing. Blocked documents and extraction failures are retained as explicit coverage gaps. This is a bounded research sample, not a complete citation index.