What is a PDF file?
Last reviewed: October 1, 2026 · Markdown version
A PDF is a file in the Portable Document Format: a fixed-layout document that is meant to look the same on every device that opens it. The .pdf at the end of the name is the file extension for that format. Adobe created it in the early 1990s; today it is an international standard published by ISO. Whether a given PDF holds real text or only pictures of text is a separate matter, and that difference decides what you can do with the file.
What the letters and the extension mean
Adobe's own definition: "PDF is an abbreviation that stands for Portable Document Format. It's a versatile file format created by Adobe that gives people an easy, reliable way to present and exchange documents — regardless of the software, hardware, or operating systems being used by anyone who views them" (What is a PDF? Portable Document Format | Adobe Acrobat, read 1 October 2026).
The internet's registration for the format, RFC 8118, lists "File extension(s): .pdf" and gives the tell-tale first bytes: "All PDF files start with the characters "%PDF-" followed by the PDF version number, e.g., "%PDF-1.7" or "%PDF-2.0"." The same document gives the format's media type, application/pdf, which is how a web server tells your browser that a PDF is coming (RFC 8118: The application/pdf Media Type, read 1 October 2026). So renaming a file to end in .pdf does not make it one; the content has to start that way.
Who created it, and who maintains it now
Adobe dates the idea to 1991, when co-founder John Warnock started what he called The Camelot Project, and says "By 1992, The Camelot Project had developed into PDF." It also states the current position plainly: "The PDF format is now an open standard, maintained by the International Organization for Standardization (ISO)" (What is a PDF? Portable Document Format | Adobe Acrobat, read 1 October 2026).
RFC 8118 fills in the dates between: "The first version of PDF, 1.0, was published in 1993 by Adobe Systems Incorporated", and "In 2008, PDF 1.7 was adopted as an ISO standard (ISO 32000-1:2008)." The newer version, PDF 2.0, is ISO 32000-2. The RFC also lists specialised subsets that ISO standardised, among them PDF/A for archiving and PDF/UA for accessibility (RFC 8118, read 1 October 2026).
ISO's catalogue lists the current edition as ISO 32000-2:2020, "Published (Edition 2, 2020)", 986 pages, and adds that "This publication was last reviewed and confirmed in 2026." Its abstract describes the purpose in one sentence: a "digital form for representing electronic documents to enable users to exchange and view electronic documents independent of the environment in which they were created or the environment in which they are viewed or printed" (ISO 32000-2:2020 — Portable document format — Part 2: PDF 2.0, read 1 October 2026).
What can be inside one
More than most people expect. RFC 8118 says "PDF pages may include text, images, graphics, and multimedia content such as video and audio", along with "annotations, bookmarks, file attachments, hyperlinks, logical structures, and metadata", and that the format "supports encryption and digital signatures" (RFC 8118, read 1 October 2026). That last point is why some PDFs ask for a password and others simply refuse to let you print.
Text, or a picture of text
Here is the distinction that explains most PDF frustrations. Adobe says it in passing: "While many PDFs are simply pictures of pages, Adobe PDFs preserve all the data in the original file format" (What is a PDF?, read 1 October 2026). A PDF exported from a word processor carries the words themselves, so you can select, copy and search them. A PDF made by a scanner or a phone camera usually carries a photograph of each page, which you can read but your software cannot.
The standard does not settle which kind you get. ISO's abstract says the document does not specify "specific processes for converting paper or electronic documents to the PDF file format" (ISO 32000-2:2020, read 1 October 2026), so whether a scan gets a text layer is up to the scanning tool. Try to select a word to find out — the full test is here — and if it turns out to be pictures, OCR is the step that adds the words.
Where DocFind fits
Once you have more than a handful of PDFs, the question changes from "what is this file" to "which file was it in". DocFind, on iPhone and Android, answers that one: search inside many PDFs at once — results show the file, the page, and the surrounding text. It handles both kinds described above, because it can read scanned PDFs on your device before searching them. It works with PDFs only, not Word or Excel files.
Related: Does my PDF have a text layer? · What is OCR? · Which app is best for a PDF viewer? · Extract text from a PDF