Django / reports
Report PDFs in Django
Post your report HTML to a render API from a view that returns FileResponse, then hand the bytes back as a FileResponse with an explicit content type. No browser binary in your Django deployment.
Why reports become PDFs
A report is a snapshot of numbers at a moment in time. Its whole value is that it does not change when the underlying dashboard does. That is the difference between a document and a view of a database, and it is why analytics, operations and account teams keep asking for this.
What a report has to carry
Before any of the code below matters, the template has to produce a document that is actually a report. These are the fields that make it one, and the ones a reader or an auditor will look for first.
- the reporting period as an explicit start and end date
- the generation timestamp, distinct from the period
- the data source or filter set, so a reader can tell what was excluded
- each metric with its unit, since a bare number is not a figure
- comparison against the prior period where the report claims a trend
Turning your the Django template language template into a document
You already have the markup. Render_to_string("invoices/detail.html", context) gives you it as a string, which is the only input the render needs. Django resolves {% static %} to a path relative to STATIC_URL, which means the stylesheet URL in your rendered HTML is only meaningful to a browser that already has a session on your host. A renderer fetching it from outside gets a 404 and produces an unstyled document. Inline the CSS into the string you send, or make STATIC_URL absolute for this one code path.
The Django implementation
httpx is the client most projects here already have. The render happens in a view that returns FileResponse, and you hand the bytes back as a FileResponse with an explicit content type.
# views.py
import httpx
from django.conf import settings
from django.http import FileResponse
from django.shortcuts import get_object_or_404
from django.template.loader import render_to_string
from .models import Report
def report_pdf(request, pk):
report = get_object_or_404(Report, pk=pk)
# A template that extends nothing. The on-screen template carries base.html,
# the nav and the messages block, none of which belong in a report.
html = render_to_string("reports/pdf.html", {"report": report})
upstream = httpx.stream(
"POST",
"https://api.pdfpipe.xyz/v1/pdf",
headers={"Authorization": f"Bearer {settings.PDFPIPE_KEY}"},
json={"html": html, "options": {"format": "A4", "printBackground": True}},
timeout=60.0,
)
response = upstream.__enter__()
response.raise_for_status()
return FileResponse(
response.iter_bytes(),
content_type="application/pdf",
filename=f"report-{report.reference}.pdf",
)The CSS that makes a report page correctly
The layout problem specific to this document is that charts and tables must not split mid-element, and a report long enough to matter will page-break somewhere awkward unless break rules are set. These rules handle it.
/* Charts and tables must not split. Sections start on a fresh
page so a heading never ends up alone at the foot of one. */
.chart, figure, table { break-inside: avoid; }
h2 { break-after: avoid; }
section { break-before: page; }
section:first-of-type { break-before: auto; }
@page { size: A4; margin: 20mm 18mm; }What goes wrong in Django
The template that renders your on-screen view is almost always the wrong one to send. It carries the base layout, the nav, and the messages framework, all of which land in the PDF. Render a template that extends nothing, and keep it beside the view template so the two do not drift.
What people try first
Most Django projects reach for WeasyPrint or xhtml2pdf, both of which put layout in your Python process. WeasyPrint needs Cairo and Pango in the image; xhtml2pdf supports a subset of CSS that stopped growing years ago. That works until it is running on more than one machine, at which point the browser becomes the thing you operate rather than the thing you use.
Where the report lives afterwards
Rendering is the short part. Reports pile up faster than any other document here because they are produced on a schedule whether or not anyone reads them. Most teams keep twelve months and expire the rest, which makes storage lifecycle a design decision rather than an afterthought.
Getting the document right
- Check the period the figures cover, because a report without an unambiguous date range is worse than no report.
- Generation is triggered by a schedule, usually month end or week end, so size the timeout for that path rather than for a health check.
- This is a batch shape: render the run through the batch endpoint rather than firing thousands of individual requests.
- Backgrounds are painted by default here, so a design that uses colour needs nothing set. Only an explicit print_background of false turns them off.
Frequently asked
Do I need Chromium installed to generate reports from Django?
No. The render happens over HTTP, so your Django deployment stays the size it is now. That is the main reason to use an API rather than WeasyPrint or xhtml2pdf, both of which put layout in your Python process. WeasyPrint needs Cairo and Pango in the image; xhtml2pdf supports a subset of CSS that stopped growing years ago.
How do I stop a long report using all the memory?
Hand the bytes back as a FileResponse with an explicit content type. The template that renders your on-screen view is almost always the wrong one to send. It carries the base layout, the nav, and the messages framework, all of which land in the PDF. Render a template that extends nothing, and keep it beside the view template so the two do not drift.
What has to be on a report?
At minimum: the reporting period as an explicit start and end date; the generation timestamp, distinct from the period; the data source or filter set, so a reader can tell what was excluded. The one to get right before anything else is the period the figures cover, because a report without an unambiguous date range is worse than no report.
Do I need to store the generated reports?
Depends on the document, and this one has a clear answer: reports pile up faster than any other document here because they are produced on a schedule whether or not anyone reads them. Most teams keep twelve months and expire the rest, which makes storage lifecycle a design decision rather than an afterthought.
Can I keep my existing report template?
Yes, if it produces HTML. Whatever renders your report view today can render the same markup for the PDF, which is why the CSS above is the only new thing you write.
Related
Other Django documents, and the same report in other stacks.
Invoice PDFs in Django
an invoice is a legal record of a demand for payment
Receipt PDFs in Django
a receipt is proof a payment happened
Certificate PDFs in Django
a certificate is meant to be shown to a third party who has no relationship with the issuing system
Report PDFs in Node.js
Using fetch, in a route handler that returns the response body directly.
Report PDFs in Python
Using httpx, in an async endpoint that returns a streaming response.
Report PDFs in PHP
Using cURL, in a controller action that echoes the body with a PDF Content-Type.
The same thing in plain Python
Without the framework, using httpx directly.
When the output is wrong
Blank pages, missing backgrounds, breaks in the wrong place, by symptom.
All stacks and documents
The full grid of what this covers.
100 free documents a month, no card.