Reading a paper's page
Each paper with a page has one address, /paper/<its identifier>/, and one page in sections. A status is always said in words, with what was verified and when; a link always says where it was found.
The sections
A bar under the title leads to each section of the same page (no section has an address of its own, so a link to …/#data works as any other). A paper whose authors' code was read opens on the Code ↔ Paper reader; another opens on its overview.
- Code
- The Code ↔ Paper reader: the paper and its authors' code side by side, with their matches. For a paper whose code was not read, the list of its code repositories instead.
- Overview
- The record: authors in order (their pages when they have an ORCID iD), affiliations and institutions, journal, volume, issue, pages, dates, type, language, licence, identifiers, categories, keywords, MeSH terms, funding, citation count, references, RRIDs, OpenAlex's topic, and the abstract when its licence allows it.
- Repositories
- Each repository of the code: its state, licence, commit and date, languages, size, whether Software Heritage has archived it, where the paper cites it, what it holds (README, licence file, CITATION.cff, environment files, tests, continuous integration, notebooks), the tools found in it, and the history of its availability checks.
- Map
- The paper's tracing map: proposed by the harvester or validated by an author, what it holds, and its DOI once deposited.
- Data
- The datasets the paper cites (each with its page), its other data links and where each was found, and its data availability statement.
- Versions
- The record's history, newest first: each date and what changed in its public facts, its links included. A correction says by whom in role only: "a verified author", "a maintainer of its code".
- Cite
- The paper in an APA-like text, BibTeX, RIS and CSL-JSON, each with a Copy button; the map's citation too once it has a DOI.
- Similar
- Up to ten papers with a page that share the most with this one — tools, categories, datasets, cited references, authors — each shared thing weighed by its rarity, with the reasons in words.
- Contribute
- What a signed-in reader may do with this record, by their role: claim it, correct its links, validate its map, add the badge, or ask for a removal. It asks the site nothing while you are signed out.
- Discussion, Reproductions, Activity
- Muted in the bar: they say what they will hold; they are not built yet, and will open with sign-in.
The sidebar on the right (below, on a phone) gathers the essentials: the way into the reader and its number of matches, the paper at its publisher and on Europe PMC, the datasets, the map's state, the licences of the paper and of its code, and the removal request.
Statuses, in words
A paper's status sums up what was found for its code. It is never a coloured badge: words, in green when the code is verified, in brown when something is wrong or unconfirmed.
- code verified
- At least one repository the paper cites as its authors' code answered when it was last checked, and holds code.
- code found, not verified yet
- The paper cites its code, but no repository could be confirmed yet: not checked yet, unreachable at the last attempt, or a service that refuses robots.
- empty repository
- The repository answers, but holds no recognizable script yet (a README, a placeholder): the announced code is not there.
- dead link
- Every repository the paper cites as its code is gone: the link answers with an error, or the repository is missing or private.
- code on request
- The paper cites no code repository, but says its code is available on request. That is a promise, not code.
- data only
- The paper links to its data (a dataset, a data repository), but not to code of its own.
- no code
- No code and no data link found in the paper. Such a paper has no page; the DOI lookup knows it.
- full text unavailable
- Only the paper's metadata could be read, not its full text: nothing can be said about its code. No page.
A listing adds what can be counted: the files readable in the reader, the matches computed, the map validated by an author (with its DOI), the datasets cited.
A repository's state and evidence
Each repository says what its last check found:
- alive
- the link answers
- dead
- the link is dead
- unverified
- not verified yet
- unreachable
- unreachable at the last attempt
- unverifiable
- cannot be verified
And how far the evidence goes, level by level:
- found — the paper cites the link;
- the link answers — it was visited and answered;
- files inventoried — its files are listed, its scripts counted, its commit recorded;
- copy kept — the text of its scripts is kept, because its licence allows it.
"Where it was found" names the place in the paper — the availability statement, a section of the text, the references, a table of resources, the supplementary material — or the metadata (DataCite, Crossref), or an archive record that points back to the paper. The sentence itself is never shown.
Abstracts and availability statements
A paper's abstract and its code and data availability statements are shown in full only when the paper's licence is CC BY, CC0, CC BY-SA or CC BY-NC. Under any other licence the page says, from facts alone, what the statements point to (datasets, repositories) and whether they say "on request", and links to the paper (the licence policy).
Retractions and corrections
A retraction, an expression of concern, a correction or a reinstatement is said first, under the title, with a link to the notice (from Europe PMC's record and the Retraction Watch data). The reasons a database gives for a retraction are not repeated: the notice is.
Older papers' pages
The 6,000 most recent papers with a page have a page built ahead of time. An older one's page is made when it is asked for, from its record: it shows the record — title, notices, authors, journal, date, type, status, DOI, licence, institutions, categories, each code repository with its licence and state, tools, datasets, the links to the paper — and the Contribute section, and says at its top what it leaves out: the Code ↔ Paper reader, the tracing map and its validation, the versions, the citation formats, the similar papers and the badge. This keeps the site within the number of files its free host serves, whatever the catalogue's size.
Which papers have a page
Three kinds: those with their authors' code (whatever its state), those whose code is "on request", and those that share data only. Every other paper read is found by the DOI lookup. A paper outside the scope — not neuroscience — has no page and is in no list or figure (the scope).
