Image to PDF in Python

Overview

Convert images to PDF documents with the PDFCrowd Python client. The client handles communication with the API, while conversions run on PDFCrowd's servers.

Installation

Install the Python client with pip, or see other installation options.

python -m pip install pdfcrowd

Quick Start

Update the input URLs and filenames as needed.

Convert an Image URL

Convert a PNG image to PDF and save it locally as logo.pdf:

import pdfcrowd

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

client.convertUrlToFile('https://your-server.com/logo.png', 'logo.pdf')

The URL must return the image itself and be reachable from PDFCrowd's servers. The client raises pdfcrowd.Error on conversion or validation errors. See Handle Errors.

Convert an Image File

Convert logo.png to PDF and save it locally as logo.pdf:

import pdfcrowd

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

client.convertFileToFile('logo.png', 'logo.pdf')

Configure a Conversion

Authentication

Pass your PDFCrowd username and API key to ImageToPdfClient. Find your credentials on the API Keys page.

Choose an Input

The source image format is detected automatically. Each conversion accepts one image.

InputMethod
URLconvertUrlToFile()
Local fileconvertFileToFile()
Bytes in memoryconvertRawDataToFile()
Readable binary streamconvertStreamToFile()

For a local image or a URL on localhost, upload the file, bytes, or an open binary stream. Remote URLs must be reachable from PDFCrowd's servers.

Send Image Bytes

Read logo.png as bytes, convert it to PDF, and save logo.pdf:

import pdfcrowd

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

with open('logo.png', "rb") as source:
    image = source.read()
client.convertRawDataToFile(image, 'logo.pdf')

Add Conversion Settings

Set conversion options on the client before calling a conversion method. For example, client.setPageSize("A4") selects A4 paper.

Common settings are listed below. See the method reference for all available options or browse Python examples.

Image operations change the source image; page settings control the output area and image placement.

PurposeMethods
Resize or rotatesetResize(), setRotate()
Crop or remove solid-color borderssetCropArea(), setRemoveBorders()
Image resolution for layoutsetDpi()
PDF page dimensions and marginssetPageSize(), setPageDimensions(), setPageMargins(), setOrientation()
Image placementsetPrintPageMode(), setPosition(), setPageBackgroundColor()
PDF passwords and permissionssetUserPassword(), setOwnerPassword(), setNoPrint(), setNoCopy()

With an explicit page size, use setPrintPageMode() with "fit" to preserve the aspect ratio and fit the image inside the margins. The "stretch" mode fills the area and can distort the image. The default mode does not scale the image to fit and can crop it. Without an explicit size, margins add a border around the image.

Handle the Result

Choose an Output

The PDF contains the source image. Making text in a scan searchable requires a separate OCR step.

The methods below use URL input; local files, bytes, and input streams have corresponding methods.

OutputMethod
Local fileconvertUrlToFile()
Bytes in memoryconvertUrl()
Writable binary streamconvertUrlToStream()

Convert an image URL to PDF and return the result as bytes:

import pdfcrowd

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

data = client.convertUrl('https://your-server.com/logo.png')
print(f"Received {len(data)} bytes")

When serving the result from a web application, use Content-Type: application/pdf.

Handle Errors

The client raises a pdfcrowd.Error exception on conversion or validation errors. This example converts an image and logs any PDFCrowd error, including its HTTP status and reason code:

import logging
import pdfcrowd

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

try:
    client.convertUrlToFile(
        'https://your-server.com/logo.png', 'logo.pdf'
    )
except pdfcrowd.Error as error:
    logging.error("PDFCrowd: %s", error)
    logging.error("Status: %s; reason: %s",
                  error.getStatusCode(), error.getReasonCode())
    raise

Local Python errors, such as a failure to read an input file or write the output, may need separate handling.

pdfcrowd.Error provides these methods:

MethodReturns
getStatusCode()The HTTP status code, when available.
getReasonCode()The reason code identifying the specific error, or -1 if unavailable.
getMessage()The error message.
getDocumentationLink()A link to relevant documentation, when available.

str(error) returns the complete error, including available status and reason codes.

Common Status Codes

StatusWhat it meansWhat to do
400Invalid input, settings, or conversion failureRead the reason code and correct the input or settings.
401Missing credentials or an inactive licenseCheck your username, API key, and license status.
403Suspended service or no credits remainingCheck your account and available credits.
413Upload exceeds the 300 MB limitReduce the upload size.
429Request rate limit reachedWait and reduce the rate of new requests.
430Concurrent request limit reachedAllow active requests to finish before starting more.
503Temporary network issueCheck your retry policy before submitting another request.

See all status and reason codes for specific explanations and Limits and Retries for retry behavior.

Read Conversion Information

This information is available after a conversion and describes the client's last conversion.

MethodUse
getJobId()Identify the conversion in logs and support requests.
getDebugLogUrl()The URL of the conversion debug log when logging is enabled.
getOutputSize()Read the PDF size in bytes.
getConsumedCreditCount()Read the credits consumed by the conversion.
getRemainingCreditCount()Read the remaining credit count reported with the conversion.

Limits and Retries

Request rate and concurrency limits depend on your license. Control how quickly your application submits conversions and how many it runs at once. A 429 response concerns request rate; a 430 response concerns requests already in progress. The maximum upload size is 300 MB.

The Python client automatically retries a request once when it receives HTTP 502 or 503. Use setRetryCount() to change that count, or set it to 0 to disable automatic retries. Account for these retries when adding an application-level retry policy.

Troubleshooting

Inspect a Conversion

This example converts an image with debug logging enabled and records the debug log URL when available. Use the input and settings from the conversion you are investigating.

import logging
import pdfcrowd

logging.basicConfig(level=logging.INFO)

client = pdfcrowd.ImageToPdfClient('demo', 'demo')

try:
    client.setDebugLog(True)
    client.convertUrlToFile(
        'https://your-server.com/logo.png', 'logo.pdf'
    )
except pdfcrowd.Error as error:
    logging.error("PDFCrowd: %s", error)
    logging.error("Status: %s; reason: %s",
                  error.getStatusCode(), error.getReasonCode())
    raise
finally:
    if client.getDebugLogUrl():
        logging.info("Debug log: %s", client.getDebugLogUrl())

The debug log contains conversion settings and processing details. You can also find logs in your conversion history. A local or connection failure may occur before a conversion log is available.

Common Problems

ProblemCheck
The image cannot be loadedCheck that the URL returns an image, rather than HTML or a login page, and is reachable from PDFCrowd's servers. For local files, check the path and read permissions.
The image is distorted or croppedCheck dimensions, crop settings and margins. Use "fit" with setPrintPageMode() when setting an explicit output size.
The result looks blurryCheck the source image's pixel dimensions and how much it is enlarged. Increasing DPI cannot recover missing detail.
Text cannot be selected or searchedThe PDF retains the source image. Recognizing text from a scan requires a separate OCR step.

For help, contact support and include any available diagnostics, the client version, and enough detail to reproduce the problem.