Requires
✓ Published0🌍 Public
NN0taN3rd
Last edited Sep 17, 2017
Created on Aug 20, 2017
This example demonstrates scraping public Twitter profiles with Puppeteer, a headless Chrome automation library, to extract tweet content, timestamps, embedded image URLs, and adaptive media containers. It also shows scraping a live webpage via the Wayback Machine to capture a full-page screenshot at 1920x1080 resolution. The script uses Puppeteer’s `page.evaluate` to run DOM queries, Bluebird’s `Promise.delay` for controlled waiting, and `fs-extra` to write the collected tweet data as JSON. It navigates pages, clicks "show more" links, and scrolls to trigger dynamic content loading before saving results.
AI-generated descriptionMIT Licensed