Skip to main content

Extract DOI from PDF — DOI List from References

Extract DOI values from a PDF bibliography, normalize duplicates, and download one DOI per line for validation or import workflows.

No signup · No credit card · Works in your browser

How does it work?

1

Step 1

Upload your PDF bibliography

Drop a PDF that contains a references section. CiteMe extracts the bibliography server-side and discards the file.

2

Step 2

Verify extracted metadata

Each parsed reference is matched against OpenAlex, CrossRef, and Semantic Scholar so titles, authors, years, journals, and DOIs come from canonical records where possible.

3

Step 3

Download a DOI list

CiteMe returns doi-list.txt with one normalized DOI per line, deduplicated from the extracted references.

By CiteMe Editorial Team·

Which citation styles are supported?

Need it in a specific citation style?

Format the result in the style your paper requires — APA, MLA, Harvard, Vancouver, Chicago, or IEEE.

Related PDF tools

Also try

Frequently asked questions

What does the DOI list export include?
The DOI list includes one normalized DOI per line, deduplicated from the matched or parsed PDF references.
Are the extracted fields verified?
Yes. CiteMe matches parsed references against academic databases such as CrossRef, OpenAlex, and Semantic Scholar before export, so canonical metadata is used when available.
Does this work with scanned PDFs?
Scanned PDFs need OCR first. If no selectable text can be extracted, use the OCR retry flow or run OCR before uploading the bibliography.

Other tools

Generate Citations

Convert & Import

Each result shows its match evidence and metadata source.