Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

iTechGuides is reader-supported. When you buy through links on our site, we may earn an affiliate commission. As an Amazon Associate I earn from qualifying purchases. Learn more

Use a browser automation agent such as Playwright to open the page, save a screenshot, and read the title directly with page.title(). The screenshot records how the page looks; the title API returns the page’s title as text, without needing visual recognition.

Capture a screenshot and read the title with Playwright

This runnable JavaScript example opens a URL, waits for the page’s load event, saves a full-page PNG, and returns the title and final URL. Install Playwright and its browser first using the Playwright Page API documentation as a reference for the page methods used here.

const { chromium } = require('playwright');

(async () => {
  const url = 'https://example.com';
  const browser = await chromium.launch();

  try {
    const page = await browser.newPage();
    const response = await page.goto(url, { waitUntil: 'load', timeout: 30000 });

    if (!response) {
      throw new Error('Navigation did not return a response.');
    }

    await page.screenshot({ path: 'page.png', fullPage: true });
    const title = await page.title();

    console.log({
      requestedUrl: url,
      finalUrl: page.url(),
      status: response.status(),
      title,
      screenshot: 'page.png'
    });
  } catch (error) {
    console.error('Page capture failed:', error.message);
    process.exitCode = 1;
  } finally {
    await browser.close();
  }
})();

Replace https://example.com with the destination URL. A successful run writes page.png in the current working directory and prints the title, final URL, and response status. The final URL matters because redirects can take the browser somewhere other than the originally requested address.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Choose the capture scope

Select the screenshot mode that answers the task. Playwright’s screenshot guidance distinguishes full-page and target-element captures; these record different things. See Playwright’s MCP screenshots guide for the related scope and visual-versus-structure guidance.

#1 Best Overall
SunFounder PiDog AI Robot Dog Kit for Raspberry Pi 5/4/3B+/Zero 2W, Openclaw LLMs ChatGPT/Gemini/Grok, Voice&Video Recognition, Python, App, Gyroscope, Camera (RPI NOT Included)
  • AI-Powered Raspberry Pi Robot Dog — PiDog: Powered by Raspberry Pi (5/4B/3B+/3B/Zero 2W), OpenClaw, and multi-LLMs like ChatGPT, Gemini, Grok, DeepSeek, Qwen & Ollama. With 12 servos, camera, gyroscope, hearing & touch sensors, PiDog can see, listen, talk, move, and interact intelligently. Supports OpenCV, MediaPipe, TTS & STT, app control, FPV & Python. A great STEM robotics gift for students, makers & tech enthusiasts—perfect for birthdays and holidays. (Raspberry Pi not included)
  • Realistic Dog-like Movements: PiDog's 12 powerful servos enable 32 dog-like actions, including walking, sitting, standing, shaking its head, wagging its tail, and performing playful tricks, closely mimicking a real dog and providing an engaging experience. This is an AI development robot product designed for engineers, suitable for ages 15 and above
  • Rich Sensor Suite for Interactive Experiences: PiDog features ultrasonic, touch, gyroscope, sound, camera, speaker and microphone. These provide it with advanced hearing, vision, and touch, enabling it to see, detect obstacles, respond to touch, and recognize sounds, making interactions highly engaging
  • AI-Powered Interactions with OpenClaw & Multi-LLMs. PiDog combines voice, vision, and gesture recognition for immersive AI experiences. Powered by OpenClaw and multi-LLMs like ChatGPT, Gemini, Grok, DeepSeek, Qwen, Doubao, and Ollama (local LLMs), it can understand questions, respond naturally through TTS & STT, recognize math problems, interpret hand gestures, and hold smart conversations. OpenClaw also enables customizable AI behaviors and personalized robotics development, helping users create their own intelligent robotic companion
  • Comprehensive Learning Resources and Support: PiDog offers detailed online documentation, video tutorials, prompt technical support, and an active forum community, ensuring beginners can easily complete all projects and enjoy a great experience
Need Playwright method What it captures
What is visible in the browser window await page.screenshot({ path: 'viewport.png' }) The current viewport.
The whole scrollable page await page.screenshot({ path: 'full-page.png', fullPage: true }) A full-page image, including content beyond the viewport.
One component or region await page.locator('main').screenshot({ path: 'main.png' }) The selected element; replace main with a suitable CSS selector.

Full-page output may be much taller than a viewport capture. For repeatable jobs, use a descriptive output filename and keep the requested scope explicit in the result you return.

Wait for the page state you actually need

page.goto() navigates the browser, but the right point to capture depends on the site. A page may render meaningful content only after JavaScript runs, or an SPA may change its title after navigation. Playwright’s Page API documents navigation and title retrieval; it does not prescribe one universal wait condition for every site.

  • For pages whose relevant content appears after a known UI transition, wait for that specific selector or state before capturing and reading the title.
  • If the title is set asynchronously, wait for the expected title or another page-specific signal rather than adding an arbitrary fixed delay.
  • For a page that never reaches the expected state, report the timeout or missing state instead of presenting a stale title as verified.

Playwright’s browser-agent documentation describes accessibility snapshots as structured information for page text and interaction, while screenshots are visual evidence. Use the screenshot to inspect layout, rendering, or charts; use a snapshot or locator when the task is to understand or interact with page structure. Playwright’s screenshot guidance explicitly cautions that screenshots are for looking at, not acting on.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.
Rank #2
AI Robotic Arm Kit with Servo Motors – LeRobot SO-ARM101 Pro Low-Cost (Without 3D Printed Parts) | 6-DOF, Open-Source, Compatible with NVIDIA Jetson
  • Optimized AI Arm Kit for LeRobot & Hugging Face Projects – The SO-ARM101 is an upgraded low-cost robotic arm servo motor kit designed for AI robotics enthusiasts and developers. Fully compatible with LeRobot and Hugging Face frameworks, it supports imitation learning and reinforcement learning, making it ideal for real-world robotics applications. (3D-printed parts not included.)
  • Enhanced Wiring & Performance – Compared to the SO-ARM100, the SO-ARM101 features improved wiring to prevent disconnection at joint 3 and eliminates range-of-motion limitations. The leader arm uses optimized gear ratio motors for smoother performance—no external gearboxes required.
  • Real-Time Leader-Follower Functionality – New real-time tracking allows the leader arm to follow the follower arm, enabling human intervention and correction during reinforcement learning (RL) training. Perfect for hands-on AI robotics development and research.
  • Open-Source, DIY-Friendly & Nvidia-Compatible – Developed by TheRobotStudio, this open-source AI Arm kit integrates seamlessly with the LeRobot platform, offering PyTorch-based datasets, simulation, training, and deployment tools. Fully compatible with Nvidia Jetson edge devices, including reComputer Mini J4012 Orin NX 16 GB.
  • Comprehensive Learning Resources – Includes detailed open-source assembly and calibration guides, testing tutorials, and deployment instructions. From wiring to AI training, get everything you need to start building, teaching, and optimizing your robotic arm for grasping and placing tasks.

Use an agent tool for title extraction, not visual OCR

If an AI agent needs both a visual artifact and metadata, expose separate operations: one captures the image, and another returns the browser title as a string. The title comes directly from the page API, so asking a vision model to read it from screenshot pixels adds an unnecessary interpretation step.

For a Playwright agent CLI workflow, the page title may appear in page information. Snapshot references belong to the snapshot in which they were created, so take a fresh snapshot after navigation before using its references. See Playwright’s snapshots guidance.

Handle JavaScript-rendered pages and hosted browsers

A static HTTP response may not contain text produced after page scripts run. A real browser session can inspect the rendered page. For example, Cloudflare documents CDP-controlled browser sessions for screenshots, live DOM state, and information available after JavaScript execution. Its browser tools page was last updated June 24, 2026, and labels the tools Beta; this is one hosted option, not a requirement. Check its current availability and terms before depending on it: Cloudflare Agents Browser tools.

Rank #3
SunFounder AI Robot Kit with Raspberry Pi Zero 2 W+32G TF Card, ChatGPT-4o Enabled with Voice Command & Video Recognition, App Control, FPV, 12 Servos, Gyroscope, Camera, Mic
  • Raspberry Pi AI Robot: powered by Raspberry Pi (5/4B/3B+/3B/Zero 2W), features 12 servos and sensors for vision, hearing, and touch. Integrated with ChatGPT-4o, it responds to complex queries. With app control and FPV, users can manage and see its view in real-time. It supports Python programming
  • Realistic Movements: 12 powerful servos enable 32 actions, including walking, sitting, standing, shaking its head, wagging its tail, and performing playful tricks, closely mimicking a real and providing an engaging experience
  • Rich Sensor Suite for Interactive Experiences: features ultrasonic, touch, gyroscope, sound, camera, speaker and microphone. These provide it with advanced hearing, vision, and touch, enabling it to see, detect obstacles, respond to touch, and recognize sounds, making interactions highly engaging
  • Engaging Interactions with ChatGPT-4o: with ChatGPT-4o enables voice interactions and visual recognition, making it smarter and more responsive. Users can have natural conversations, solve math problems via the camera, and interpret gestures, creating diverse and fun interactions
  • Comprehensive Learning Resources and Support: offers detailed online documentation, video tutorials, prompt technical support, and an active forum community, ensuring beginners can easily complete all projects and enjoy a great experience

Report results and failures accurately

Return the title alongside enough context for someone to understand what was captured: the final page URL, screenshot path, scope (viewport, full page, or element), and any navigation or capture error. A browser response alone does not prove that a title reflects the intended content if the page failed to load or had not reached the relevant state.

Free tools Windows power users keep installed

One-click scans. No signup required.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Common problems and fixes

  • Navigation times out: record the timeout and check whether the destination is slow, blocked, or waiting on a long-running resource. Increase the timeout only when the workflow needs more time; do not treat a timeout as a successful capture.
  • Title is blank or stale: confirm the final URL, then wait for the page-specific navigation or render condition that sets the title before calling page.title().
  • Screenshot omits content below the fold: use fullPage: true if the whole document is required. Use a viewport capture when only the visible area is relevant.
  • Screenshot file is missing: check that the process can write to the selected path and that the screenshot call completed before the browser closed.
  • Agent uses an outdated snapshot reference: create a new snapshot after navigation; CLI references are invalidated when the page changes.
  • Sensitive information appears in the image: consider masking selected locators; the Page API documents screenshot masking options.
Independent reader supportYour contribution helps us test, update, and keep practical guides available for everyone.Support on Ko-Fi

Or skip the browser setup

ScreenshotNeo provides a screenshot API and MCP server. A single GET request can return an image or PDF, while its get_page_info MCP tool is designed to retrieve page information for an AI agent. The browser approach above is useful when you need direct control of Playwright; ScreenshotNeo is an alternative when you would rather call an API. Its clean-shot workflow accepts consent banners before capture and removes more than 60 known consent platforms, newsletter popups, and chat widgets; those steps can be turned off. Bot checks, blank pages, timeouts, failed loads, and cache hits are not billed, and response headers identify the page verdict and billing status.

For example, save a screenshot of the target page as WebP with cURL:

curl -G "https://api.screenshotneo.com/v1/shot" -d access_key=YOUR_API_KEY --data-urlencode url=https://example.com -o shot.webp

See the ScreenshotNeo API documentation for request options. ScreenshotNeo also has an MCP server for AI agents, with take_screenshot, get_page_info, and capture_pdf tools. For a page title, use its page-information tool rather than trying to read the title from the image. ScreenshotNeo’s site is screenshotneo.com.

Rank #4
AI Robotic Arm Kit Hiwonder SO-ARM101 Embodied Imitation Learning Open Source 6-Axis Robot Arm 12 High-Torque Bus Servo Motors AI Vision Recognition (Advanced Kit, Included 3D Printed Part, Assembled)
  • 【End-to-End Imitation Learning】Hiwonder SO-ARM101 robot arm is an embodied intelligent hardware platform compatible with the Lerobot open-source framework. It provides developers with streamlined access to shared code, templates, and pre-trained models to explore the latest advancements in AI research.
  • 【Dual-Camera Vision System】Equipped with both a gripper-mounted camera and an external camera, the system supports both precise manipulation and environmental awareness for accurate imitation learning.
  • 【Hiwonder High-Performance Bus Servos】Featuring 12 high-torque bus servo motors with magnetic feedback, the Hiwonder SO-Arm101 robotic arm delivers smooth, stable motion, eliminating issues like power deficiency and jitter.
  • 【Professional Control & Debugging】Integrated with the Hiwonder BusLinker V3.0 debugging board, the system supports servo scanning, real-time status monitoring, and trajectory control. The professional PC software simplifies device calibration and debugging, making it accessible for both researchers and hobbyists.
  • 【Open-Source Compatibility】The SO-ARM101 robotic arm is designed to be fully compatible with the LeRobot open-source project. We acknowledge the contributions of the open-source community; all trademarks and copyrights belong to their respective owners.

Sign up for 1,000 screenshots a month free, with no card required.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Cost and implementation trade-offs

With local Playwright, you operate the browser runtime and decide how to manage browser installation, output storage, and navigation behavior. A hosted browser service moves browser operation to that provider, so availability, terms, and deployment fit need checking. The cited documentation establishes no like-for-like performance or pricing comparison between those approaches.

ScreenshotNeo’s listed plans are Free (1,000 shots per month, no card), Starter ($5 for 3,000), Growth ($15 for 15,000), Pro ($39 for 60,000), Scale ($99 for 250,000), and Business ($249 for 1,000,000); yearly billing gives two months free. Every feature is available on every plan. These are ScreenshotNeo plan terms; confirm the current offering on its site before choosing a plan.

Frequently Asked Questions

Does a screenshot tell an AI agent the page title?

No. The screenshot is a visual file; retrieve the title separately from the browser page or an agent page-information tool.

Can I extract a page title from a JavaScript-rendered site?

Yes, when the browser session has reached the state that sets the title. Use a real browser and wait for the page-specific condition before reading it.

Special offer. See more information about Outbyte and uninstall instructions. Please review EULA and Privacy policy.

Product prices and availability are accurate as of the date/time indicated and are subject to change. Any price and availability information displayed on Amazon at the time of purchase will apply.