Free tools Windows power users keep installed

One-click scans. No signup required.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

To export crawler results to Excel, save the crawler’s items as a CSV file, then open that file in Excel or import it through Data > From Text/CSV. CSV is a practical bridge between a crawler and a worksheet, but it is not an Excel workbook: it does not preserve formatting or multiple sheets. The exact export controls vary by crawler; the example below uses Scrapy.

Choose the right file format first

For ordinary scraped records—such as one row per product, article, or page, with a fixed set of fields—CSV is usually the simplest handoff. Scrapy’s Feed Exports documentation lists CSV, JSON, JSON Lines, and XML among its supported formats. CSV is convenient to inspect in rows and columns and easy to open in Excel. JSON, JSON Lines, or XML may suit a downstream program better when your items have nested structures or need to retain structure beyond a flat table. The best choice depends on the item shape and what will consume the data.

Format Useful when Excel consideration
CSV You want a simple row-and-column interchange file. Excel can open or import it, but it is still plain text rather than a formatted workbook.
JSON or JSON Lines A downstream program needs structured records; JSON Lines stores records line by line. These are not equivalent to a ready-made Excel worksheet; conversion or import steps may be needed.
XML Your next tool expects XML or the data workflow calls for that structure. It is a supported Scrapy feed format, but it is not an XLSX workbook.

Scrapy’s documented feed formats do not include XLSX as a built-in exporter. If the final deliverable must be a native Excel workbook—with multiple sheets, formulas, or formatting—export the crawler data and then save or create a workbook in Excel. Do not expect a CSV round trip to retain workbook features.

Export Scrapy items to CSV

Scrapy’s feed export settings can write a feed to a local path. Set the destination to a filename ending in .csv and the format to csv. The local filesystem feed storage uses a path and, according to Scrapy’s feed storage documentation, requires no extra library. Add FEED_EXPORT_FIELDS if you want a predictable set and order of columns; Scrapy also supports assigning output names to fields.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.
#1 Best Overall
Sale
The Microsoft Office 365 Bible: The Most Updated and Complete Guide to Excel, Word, PowerPoint, Outlook, OneNote, OneDrive, Teams, Access, and Publisher from Beginners to Advanced
  • The Microsoft Office 365 Bible: The Most Updated and Complete Guide to Excel, Word, PowerPoint, Outlook, OneNote, OneDrive, Teams, Access, and Publisher from Beginners to Advanced
  • ABIS BOOK

Basic project settings

Put the following in the Scrapy project’s settings file, such as settings.py. Adjust the field names to match the items your spider yields.

FEEDS = {
    "results.csv": {
        "format": "csv",
    },
}

Run the spider using your usual project command. When the crawl completes, look for results.csv in the configured working location. Scrapy’s feed export encoding defaults to UTF-8, which is suitable for many multilingual text exports.

Fix the columns and their order

For a stable spreadsheet layout, list the fields you want in the order they should appear. Replace the example names with your item fields. Scrapy also allows a field to be given a different output name; use that option if a concise or human-readable header is more useful than the internal item key.

FEED_EXPORT_FIELDS = [
    "url",
    "title",
    "price",
    "sku",
]

This is especially helpful when different items may have different keys: a CSV has a header row and a fixed column layout, so deciding on the fields before opening the file makes it easier to compare records. Check that the spider actually yields the selected fields. If a field is absent from an item, its cell may be empty rather than indicating a crawler failure.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Open the CSV in Excel

Quick option: open the file

  1. Finish the crawl and locate results.csv.
  2. Open the file in Excel. Microsoft says Excel displays the CSV data in a new workbook.
  3. Check the headers, a few records, and any columns that contain dates, identifiers, or values with leading zeroes.

This is the fastest route for a straightforward file. However, Microsoft cautions that when you open a CSV directly, Excel interprets it using its current default data-format settings. For example, a date can be interpreted in an unexpected order, or a product code or ZIP code such as 00127 can lose its leading zeroes if Excel treats it as a number.

More control: import through Text/CSV

  1. Open Excel and choose Data > From Text/CSV.
  2. Select the crawler’s .csv file.
  3. Review the import options and preview. Check that the values split into the expected columns and that the text is readable.
  4. For columns that must remain identifiers or preserve leading zeroes, use the import controls to treat those values as text rather than numbers. Review date columns before loading, too.
  5. Load the data into a new worksheet or, where appropriate, an existing workbook.
  6. Compare the imported row count and a few representative values against the crawler output before treating the worksheet as complete.

Import is preferable when you need to review delimiter, encoding, date interpretation, or data types rather than relying on automatic defaults. The exact controls shown can depend on your Excel version and platform; Microsoft’s general guidance is to review the import options before loading.

Keep CSV and an Excel workbook distinct

A CSV file contains delimited text, not the full contents of an Excel workbook. Microsoft’s guidance on saving a workbook to text format notes that text formats remove formatting and CSV saves only the active sheet. That means CSV is suitable for exchanging crawler rows, but it is not a safe archival format for a workbook that uses multiple worksheets, formulas, or presentation formatting.

If you need those workbook features, load the crawler data and save the finished file in an Excel workbook format such as .xlsx. Keep the original CSV as a separate data export when you need a plain-text interchange copy. Avoid overwriting a carefully formatted multi-sheet workbook with CSV: only the active sheet is represented in that text file.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Check whether the export fits Excel

Microsoft lists a text-file import/export capacity of 1,048,576 rows and 16,384 columns for Excel. These are product limits, not a guarantee that a large crawl will be convenient to work with in a worksheet. If the export is near or beyond the row limit, or is too large to inspect and analyze comfortably, use a storage or analysis workflow designed for the full dataset and send a smaller subset to Excel.

  • Estimate the record count before running a large crawl, especially if each page can produce multiple items.
  • After export, compare the number of crawler records with the rows you expect Excel to contain.
  • For large jobs, consider splitting the crawl into manageable exports or filtering the fields and records before creating a worksheet.
Independent reader supportYour contribution helps us test, update, and keep practical guides available for everyone.Support on Ko-Fi

Troubleshoot common export problems

The CSV is missing or empty

  • Confirm the crawl completed and that the configured feed path is where you are looking.
  • Check that the spider yielded items; a successful process that found no records can still leave you without useful rows.
  • Verify the feed format is set to csv and that the selected output filename ends in .csv.

Columns are missing, blank, or in the wrong order

  • Compare FEED_EXPORT_FIELDS with the actual keys in the items yielded by the spider.
  • Put the fields in the desired spreadsheet order in the setting, and use output names if the internal keys are unsuitable headers.
  • Inspect the raw CSV as text to distinguish an empty crawler value from a display or import issue in Excel.

Text looks garbled or values appear in one column

  • Use Excel’s Data > From Text/CSV import path and review the preview rather than accepting automatic interpretation blindly.
  • Check the import settings for the appropriate delimiter and encoding. Scrapy’s documented default feed encoding is UTF-8.
  • When the file opens as one column, review delimiter handling; when characters are garbled, review encoding before loading.

Excel changed dates or removed zeroes

  • Import through Text/CSV instead of opening the CSV directly.
  • Set identifier columns such as ZIP codes and product codes to text so Excel does not convert them to numbers.
  • Review date columns in the import preview and ensure the interpretation matches the source values.

The export exceeds worksheet capacity or loses workbook features

  • If the dataset exceeds Excel’s documented text-file limits, use a database or another large-data analysis workflow rather than expecting one worksheet to hold every record.
  • If formatting, formulas, or several sheets matter, keep the exported CSV as an input and save the finished result as an Excel workbook format such as XLSX.

Add page screenshots alongside crawler data

A screenshot is different from a crawler’s structured results: it records how a page looked, while the CSV holds fields such as URL, title, and price. If visual evidence is useful for a QA or review workflow, capture screenshots separately and associate them with the relevant page URLs or records. ScreenshotNeo is a website screenshot API and MCP server, not a crawler or CSV exporter.

Or skip the browser setup

For a screenshot of a crawled page, one GET request can return an image or PDF. This cURL example saves a WebP screenshot of a page; replace the target URL with a page from your crawl. See the ScreenshotNeo API documentation for request options.

curl -G "https://api.screenshotneo.com/v1/shot" -d access_key=YOUR_API_KEY --data-urlencode url=https://stripe.com -o shot.webp

ScreenshotNeo accepts cookie or consent banners as a visitor and removes 60+ known consent platforms, newsletter popups, and chat widgets before capture; each step can be turned off. Bot checks, blank pages, timeouts, failed loads, and cache hits are not billed, and response headers say which page verdict applied and whether the request was billed. An MCP server provides take_screenshot, get_page_info, and capture_pdf tools for AI agents and MCP clients. The free plan includes 1,000 screenshots per month with no card; paid plans start at $5 for 3,000 shots. Screenshot capture does not export your crawler’s item data to Excel.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Sign up free for 1,000 screenshots a month, with no card.

Frequently Asked Questions

Can I append new crawl results to the same CSV on every run?

The Scrapy example above configures a feed path for an export; it does not establish an append workflow. Check the behavior and settings for your Scrapy version and feed-storage setup before relying on repeated runs to preserve earlier rows.

Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.