How to find the DOI of a PDF

Last reviewed on October 3, 2026

A DOI (Digital Object Identifier) is the most reliable way to cite an article, and reference managers can build a complete citation from it. Most journal articles published in the last two decades have one, but it is not always obvious where it is in the PDF. Work through these steps in order.

1. Look where publishers print it

A DOI always starts with 10., followed by a number, a slash, and a suffix. If you see something like 10.1038/s41586-020-2649-2, that is it.

2. Let a tool scan the PDF

The in-browser extractor searches the first pages of the PDF for a DOI pattern, and also reports arXiv IDs, PubMed IDs and ISBNs. The file is processed on your own device and is not uploaded. Reference managers do something similar: Zotero's "Retrieve Metadata for PDF" looks for a DOI or ISBN in the file and fetches the full record.

3. Check the PDF properties

Some publishers embed the DOI in the PDF's metadata rather than (or as well as) the visible text. In a PDF reader, open File → Properties (or Document Properties) and look at the subject, keywords or the advanced/XMP metadata. Be careful: these fields are sometimes empty or copied from a template.

4. Look it up from the title

Match on more than the title: check authors, year and journal, since different articles can share very similar titles, and corrections or errata have their own DOIs.

5. Verify it

Open https://doi.org/ followed by the DOI. It should land on the article you have. When copying a DOI from a PDF, watch for line breaks inside it, trailing full stops or brackets that are not part of it, and ligatures or odd characters introduced by text extraction.

When there is no DOI