Scanning a book page by page produces a PDF where every page has a dark, curved shadow running down the center, the shadow cast by the book spine where the pages curve into the binding. The text near the spine is darkened and sometimes distorted by the page curvature. The shadow consumes ink or toner when printed, makes the page look unprofessional, and interferes with OCR accuracy on the affected text. Cropping the spine shadow out of each scanned page removes the visual artifact and focuses the reader attention on the content. The Crop PDF approach for spine shadow removal must be applied page by page, because the shadow position shifts slightly from page to page as the book thickness changes from front to back.

Why Spine Shadows Appear in Book Scans
Flatbed scanners press the book against a glass platen. The page near the spine cannot lie completely flat because the binding prevents it. This gap between the page and the glass creates a region where the scanner light reflects differently, producing a shadow that starts at the spine edge and fades as the distance from the binding increases. On the left page of an open book, the shadow is darkest along the left edge and extends toward the center. On the right page, the shadow is darkest along the right edge.
The curvature of the page near the binding also distorts the text. Characters near the spine appear compressed or stretched depending on the angle of the page relative to the scanner sensor. Cropping removes the shadow and the distorted text. The alternative, using image processing to lighten the shadow and deskew the distorted text, is more complex and rarely produces as clean a result as simply cropping the affected area away. The Scanned PDF from a book scan benefits more from cropping than from attempted correction of the shadow region.
Try Crop PDF
No installation needed. Works directly in your browser.
Determining the Crop Boundaries
Open the first scanned page in a PDF viewer that supports crop tools. Zoom in on the spine edge and identify where the shadow begins to noticeably darken the page background. The crop boundary should be placed just inside this point, removing the entire shadow zone. Include a small margin of normal page background next to the shadow to ensure no darkened pixels remain. A crop boundary that clips into the shadow leaves a dark strip along the edge of the cropped page.
Measure the distance from the page edge to the crop boundary. This distance becomes the crop value for all pages in the scan. However, the optimal crop distance may vary between the left and right pages of the book, and it may vary between the front and back of the book where the page curvature changes. Scan through a sample of pages from different sections of the book and check whether the spine shadow width is consistent. If it varies, group the pages into sections that share the same crop distance. Apply the appropriate crop distance to each group.
Applying the Crop to All Pages
After determining the crop boundary, apply it to every page in the PDF. Most PDF editors that support cropping can apply the same crop rectangle to a range of pages. Select all the pages that share the same crop distance, open the crop tool, enter the crop values, and apply. For a book scan where the left and right pages need different crop values, process the left and right pages as separate page groups. A Crop PDF batch operation that applies the same crop to all selected pages takes seconds.
If the PDF was scanned as two-page spreads, where each PDF page shows both the left and right pages of the open book, the cropping is more complex. The spine shadow runs down the center of the page. Crop the left half of the page to remove the left side of the spine shadow, and crop the right half to remove the right side. Some PDF tools can split each page down the center and crop each half independently, producing individual pages from the spread. WukongPDF and similar platforms provide crop tools that handle both single-page and spread-page book scan cropping.
Preserving Page Content Near the Spine
The spine shadow region may contain content that cannot simply be discarded. A table that spans the full page width may have columns that extend into the shadow zone. A diagram or photograph may be positioned close to the binding. A footnote or a marginal note may be partially obscured by the shadow. For these pages, cropping the entire shadow zone removes content the reader needs.
Handle these exceptional pages individually. If the content in the shadow zone is legible despite the darkening, crop less aggressively on that specific page, removing only the darkest portion of the shadow and leaving the text that is darkened but still readable. If the content is illegible, consider whether the physical book can be re-scanned with more careful handling to flatten the page further. Pressing the book more firmly against the scanner glass, using a book scanner with a V-shaped platen designed for bound volumes, or scanning the book with an overhead camera rather than a flatbed all reduce the spine shadow at the source.
Post-Crop Quality Verification
After cropping, scroll through every page of the PDF to verify that no shadow remnants remain and that no content was accidentally cropped. Pay particular attention to pages near the beginning and end of the book, where the page curvature and shadow width may differ from the middle pages. A crop that was perfect for page 100 may be too aggressive for page 5, where the thinner stack of pages on one side of the binding produces a different curvature.
Run OCR on the cropped PDF and compare the text recognition accuracy against the original uncropped scan in the shadow zone. If the cropped version recognizes text in the formerly shadowed area more accurately, the crop was successful. If the cropped version loses text that was previously recognized, the crop may have been too aggressive. The PDF Quality goal is a clean page with all essential content preserved and all distracting artifacts removed.
Alternatives to Cropping for Spine Shadow Removal
Cropping is the simplest approach to spine shadow removal, but it is not the only approach. Image editing software can apply localized brightness and contrast adjustments to the shadow zone, lightening the background and increasing the text contrast without removing the content. This approach preserves every word on the page but requires per-page image editing, which is practical only for short documents or individual pages of high importance.
Dedicated book scanning software such as ScanTailor includes a content selection feature that automatically detects the text area on each page and crops to that area, handling the spine shadow without manual measurement of crop boundaries. ScanTailor processes scanned images before they are assembled into a PDF. If you regularly scan books to PDF, investing in a scanning workflow that includes automatic content detection and cropping saves hours of manual Crop PDF adjustment.
For rare or fragile books that cannot be pressed flat against a scanner, consider photographing the pages with a camera positioned directly above the open book. The camera approach avoids the glass platen entirely and, with careful lighting from both sides, can minimize the spine shadow at the capture stage rather than relying on post-processing to correct it.
After cropping and finalizing the PDF, consider donating the cleaned digital version to an online archive or a digital library if the book is in the public domain. The hours spent cleaning up the scan can benefit not just your own project but anyone else who needs access to that text in a readable digital format.
For books with illustrations, photographs, or maps that span across the gutter between facing pages, cropping the spine shadow sacrifices the center portion of the image. In these cases, consider keeping the two-page spread intact for the illustration pages and cropping only the text-only pages. A note in the document metadata can explain that certain pages were left uncropped to preserve the cross-gutter content.
The Scanned PDF cleanup process for book digitization projects benefits from a structured workflow that separates the scanning, cropping, OCR, and quality-check steps. Attempting to crop and finalize each page as it is scanned leads to inconsistent results. Scan the entire book first, then process the complete set of page images through the cropping step as a batch.
The Crop PDF workflow for book scans transforms a rough scan into a clean digital copy suitable for reading, printing, and long-term preservation in digital library collections.
The cleaned Scanned PDF with spine shadows removed is ready for OCR processing, which benefits from the improved image quality
A well-cropped book scan preserves the author words while removing the mechanical artifacts of the scanning process
Book scanning is an act of preservation. Every page cleaned of scanner shadows is a page that future readers will be able to read without distraction. The care invested in the cropping step is part of the larger effort to carry the content of physical books into the digital future.
Try Crop PDF
No installation needed. Works directly in your browser.
