Scroll to navigation

PDFTOHTML(1) General Commands Manual PDFTOHTML(1)

NAME

pdftohtml - program to convert pdf files into html, xml and png images

SYNOPSIS

pdftohtml [options] <PDF-file> [<html-file> <xml-file>]

DESCRIPTION

This manual page documents briefly the pdftohtml command. This manual page was written for the Debian GNU/Linux distribution because the original program does not have a manual page.

pdftohtml is a program that converts pdf documents into html. It generates its output in the current working directory.

OPTIONS

A summary of options are included below.

Show summary of options.
first page to print
last page to print
dont print any messages or errors
print copyright and version info
exchange .pdf links with .html
generate complex output
ignore images
generate no frames. Not supported in complex output mode.
use standard output
zoom the pdf document (default 1.5)
output for XML post-processing
output text encoding name
owner password (for encrypted files)
user password (for encrypted files)
force hidden text extraction
output device name for Ghostscript (png16m, jpeg etc)
do not merge paragraphs
override document DRM settings

AUTHOR

Pdftohtml was developed by Gueorgui Ovtcharov and Rainer Dorsch. It is based and benefits a lot from Derek Noonburg's xpdf package.

This manual page was written by Søren Boll Overgaard <boll@debian.org>, for the Debian GNU/Linux system (but may be used by others).