The Tool Desk
Outbyte PC Repair FREEClear out junk files and repair common Windows errorsFree Scan →Outbyte Driver Updater FREEFix the driver behind crashes, sound loss and screen glitchesFind Drivers →Playwright’s screenshot API saves a page as pixels, not searchable text. To make a searchable PDF that preserves those captured pixels, take a screenshot, convert the image to PDF, then add an OCR text layer with OCRmyPDF. If you want a PDF rendered from webpage content instead, use Playwright’s page.pdf(); it uses print CSS by default.
Choose screenshot-plus-OCR or Playwright’s PDF export
| Approach | What the PDF preserves | Searchable text | Page styling |
|---|---|---|---|
| Screenshot, image-to-PDF, then OCR | The captured image appearance | Added by OCRmyPDF; recognition quality depends on the source image and OCR | The screenshot’s rendered appearance |
page.pdf() |
Playwright’s PDF rendering of the webpage | Text is rendered as PDF content rather than added by OCR | Print CSS by default; use screen media when needed |
Use the first path when visual fidelity to the screenshot is the priority. Use page.pdf() when the page’s print layout is acceptable and you do not need a screenshot-faithful image. Playwright documents screenshots and PDF rendering as separate APIs; OCRmyPDF’s documented role is to add a text layer to image PDFs. Playwright Page API · OCRmyPDF 13.5.0 documentation
Install the tools
This workflow uses Playwright for browser automation, an image-to-PDF converter for the intermediate PDF, and OCRmyPDF for searchable text. Install Playwright in your project and install OCRmyPDF according to its documentation and your operating system. The image-to-PDF conversion command depends on the converter you choose; the example below uses ImageMagick’s magick command.
Install the browser binary for your selected Playwright browser if your environment does not already have it. The example uses Chromium via Playwright’s JavaScript package.
#1 Best Overall
- PORTABLE SCANNER FOR USE ON-THE-GO — The fastest and lightest mobile single-sheet-fed compact document scanner in its class¹
- QUICK DOCUMENT SCANNING ― This Epson ultra-fast scanner scans a single page as quickly as 5.5 seconds²; Windows and Mac compatible
- VERSATILE PAPER HANDLING ― Portable scanner scans documents up to 8.5 x 72 in; Also easily digitizes receipts and ID cards to make accounting, bookkeeping, and organizing simpler
- INTUITIVE, HIGH-SPEED SOFTWARE — Epson ScanSmart Software³ is a smart tool allowing you to easily scan, review, and save; Stay organized easily with the help of this Epson scanner
- EASY SETUP — USB-powered connect to your computer for quick and simple scanning; No batteries or external power supply required to operate portable document scanner; Standard Connectivity: USB 2.0
Capture a full-page image and add OCR
-
Navigate to the target page and wait for a meaningful readiness condition.
domcontentloadedis a practical initial wait, but pages that populate content asynchronously may require a site-specific selector or additional wait. -
Capture the screenshot. Set
fullPage: trueto capture the full scrollable page; omit it to capture only the current viewport. -
Convert the image to a PDF, then run OCRmyPDF against that image PDF to add a searchable text layer.
Rank #2
SaleBrother DS-640 Compact Mobile Document Scanner, (Model: DS640)- FAST SPEEDS - Scans color and black and white documents a blazing speed up to 16ppm (1). Color scanning won’t slow you down as the color scan speed is the same as the black and white scan speed.
- ULTRA COMPACT – At less than 1 foot in length and only about 1. 5lbs in weight you can fit this device virtually anywhere (a bag, a purse, even a pocket).
- READY WHENEVER YOU ARE – The DS-640 mobile scanner is powered via an included micro USB 3. 0 cable allowing you to use it even where there is no outlet available. Plug it into you PC or laptop and you are ready to scan.
- WORKS YOUR WAY – Use the Brother free iPrint&Scan desktop app for scanning to multiple “Scan-to” destinations like PC, Network, cloud services, Email and OCR. (2) Supports Windows, Mac and Linux and TWAIN/WIA for PC/ICA for Mac/SANE drivers. (3)
- OPTIMIZE IMAGES AND TEXT – Automatic color detection/adjustment, image rotation (PC only), bleed through prevention/background removal, text enhancement, color drop to enhance scans. Software suite includes document management and OCR software. (4)
-
Check that the output looks right and that you can find or copy text in a PDF viewer.
Recommended Free Tools
Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.
Runnable JavaScript example
Save as capture.cjs in a project with Playwright installed. It writes the screenshot, converts it with ImageMagick, and runs OCRmyPDF. The conversion and OCR commands must be available on the machine’s PATH.
const { chromium } = require('playwright');
const { execFileSync } = require('node:child_process');
(async () => {
const url = process.argv[2];
if (!url) throw new Error('Usage: node capture.cjs https://example.com');
const browser = await chromium.launch();
try {
const page = await browser.newPage();
await page.goto(url, { waitUntil: 'domcontentloaded', timeout: 60000 });
await page.screenshot({ path: 'page.png', fullPage: true });
} finally {
await browser.close();
}
execFileSync('magick', ['page.png', 'page-image.pdf'], { stdio: 'inherit' });
execFileSync('ocrmypdf', ['page-image.pdf', 'page-searchable.pdf'], { stdio: 'inherit' });
})();
Run it with node capture.cjs https://example.com. The output is page-searchable.pdf. For viewport-only capture, change the screenshot call to await page.screenshot({ path: 'page.png' }). Playwright documents saving screenshots by path and the fullPage option in its Screenshots guide.
Rank #3
- STAY ORGANIZED – Easily convert your paper documents into digital formats like searchable PDF files, JPEGs, and more.Power Consumption : 2.5W or less (Energy Saving Mode: 0.7W). Suggested Daily Volume : 500 scans..Does it contain liquid: no
- CONVENIENT AND PORTABLE –lightweight and small in size, you can take the scanner anywhere from home offices, classrooms, remote offices, and anywhere in between
- HANDLES VARIOUS MEDIA TYPES – Digitize receipts, business cards, plastic or embossed cards, reports, legal documents, and more
- FAST AND EFFICIENT – No technical hurdles or complicated setups here; easily scan both sides of a document at the same time, in color or black-and-white, at up to 12 pages-per-minute, and with a 20 sheet automatic feeder
- BROAD COMPATIBILITY – Works with both Windows and Mac devices, be it laptop or computer
PDF page sizing and long pages
A full-page screenshot is one tall image. When converted to PDF, it may become a single unusually tall page rather than a conventional multi-page document. If standard page breaks, selectable native text, or a print layout matter more than matching the screenshot, use page.pdf() instead. If a conventional page size is required but screenshot appearance is essential, split or resize the image as a separate image-processing step before or during PDF conversion; inspect the result because scaling can make text harder for OCR to recognize.
Use Playwright’s PDF rendering when screenshot fidelity is not required
For a browser-rendered PDF, call page.pdf(). Playwright uses print CSS by default. To render using screen styles, emulate the screen media type before exporting:
What’s actually slowing this PC down?
Pick the symptom - the matching free tool is one click away.
const { chromium } = require('playwright');
(async () => {
const browser = await chromium.launch();
try {
const page = await browser.newPage();
await page.goto('https://example.com', { waitUntil: 'domcontentloaded' });
await page.emulateMedia({ media: 'screen' });
await page.pdf({ path: 'page.pdf' });
} finally {
await browser.close();
}
})();
Remove the emulateMedia call if print styling is what you want. See the Playwright Page API for the PDF and screenshot methods.
Rank #4
- IRIScan Express, portable scanner : scans color and black and white documents a blazing speed up to 8ppm simplex. Color scanning won’t slow you down as the color scan speed is the same as the black and white scan speed.
- IRIScan Express mobile scanner is powered via an included micro USB 2. 0 cable allowing you to use it even where there is no outlet available. Plug it into you PC or laptop and you are ready to scan. USB cable provided. AC Adapter not provided and not needed.
- IRIScan flatbed scanner uses a simplex scanning mode allows for quick and straightforward scanning of single-sided documents. IRIScan with its full portable features is the ideal document scanners for computers.
- IRIScan document scanner : Versatile scanning capabilities, including scanning to Word, PDF, and Excel formats with companion software provided Readiris OCR
- Receipt scanner and card scanner with Additional features include scanning business cards directly to Outlook, photo scanning, and receipt scanning for efficient document management
Or skip the browser setup
For an API alternative, ScreenshotNeo returns a screenshot or PDF from one GET request. It is not the screenshot-plus-OCR Playwright workflow above; use it when an API-produced capture or PDF suits your task. Its API options include PDF settings, full-page capture, and a usage API. For a searchable image-based PDF, verify that the output meets your text-search needs.
Here is the one-call screenshot example; see the ScreenshotNeo documentation for API details:
curl -G "https://api.screenshotneo.com/v1/shot" -d access_key=YOUR_API_KEY --data-urlencode url=https://stripe.com -o shot.webp
- Cookie banners are accepted and removed before capture; supported consent platforms, newsletter popups, and chat widgets can also be removed, with each step configurable.
- Bot checks or CAPTCHAs, blank pages, timeouts, failed loads, and cache hits are not billed; response headers report the page verdict and billing status.
- An MCP server provides
take_screenshot,get_page_info, andcapture_pdftools for AI agents and MCP clients. - The free plan includes 1,000 screenshots per month with no card; paid plans start at $5 for 3,000 shots.
Sign up for ScreenshotNeo’s free plan to try 1,000 screenshots a month without a card.
Best Value
- Scanner type: Document
- Connectivity technology: USB
- With Auto Scan Mode, the scanner automatically detects what you're scanning
- Digitize documents and images
Troubleshooting and quality checks
The screenshot is blank or missing page content
The page may not have finished rendering when the screenshot was taken. Wait for a selector that identifies the content you need, or add a site-appropriate delay after navigation. A generic network-idle wait is not a guarantee that every application has finished rendering.
OCR text is missing or inaccurate
OCR works from pixels, so small type, low contrast, unusual fonts, scaling, or image compression can make recognition harder. Check the screenshot itself first; if the text is not legible in the image, OCR cannot reliably recover it. Increase the capture resolution or improve the visual conditions, then rerun the image-to-PDF and OCR steps.
The PDF opens but text cannot be searched
Confirm that you ran OCRmyPDF on the converted image PDF and are opening its output file, not the intermediate page-image.pdf. Try selecting text or searching for a clearly visible word in a PDF reader.
ImageMagick or OCRmyPDF command is not found
The corresponding program is not installed or its executable is not on PATH. Install it for the machine running the script, then check that magick and ocrmypdf run from the same shell used to launch Node.
Do these 3 things before closing this tab:
1Fix the driver behind crashes, sound loss and screen glitches2Clear out junk files and repair common Windows errors3Scan for outdated or missing drivers - takes under a minuteThe PDF is too tall or text is too small
That is a consequence of turning one full-page screenshot into a single image page or fitting it to a standard page. Use Playwright’s PDF export for conventional pagination, or process the captured image into page-sized sections before OCR. Check that any resizing preserves readable text.
The PDF differs from what the browser showed
page.pdf() uses print CSS by default, which can change or omit screen-oriented styling. Emulate screen media before PDF export when screen styles are required. A screenshot captures the rendered appearance instead, but does not supply searchable text until OCR is applied.
Quick Recap
Performance, reliability, and cost considerations
- Capture time: full-page screenshots require capturing more content than a viewport image, and OCR adds a separate processing step. Use a viewport capture when the full page is unnecessary.
- Reliability: choose a readiness signal based on the target site. A successful navigation event does not prove that delayed or lazy-loaded content has appeared; inspect the output when completeness matters.
- OCR review: searchable text is machine-recognized from the image, so verify critical names, figures, and small print against the visible screenshot.
- Cost: Playwright, image conversion, and OCRmyPDF run in your environment; any infrastructure or service costs depend on how you install and host them. No per-capture cost is specified by the cited tool documentation.
Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.

