Web scraping for Summer Teams
✓ Published0🌍 Public
CCliffordAnderson
Last edited Aug 6, 2020
Created on Aug 6, 2020
This example shows how to extract participant names from Vanderbilt University Library’s summer project listings and reformat them into CSV rows of project–participant pairs. It scrapes the HTML page using `fetch:text` and `html:parse`, then navigates the DOM with XPath to select project headings and member text. The code uses `fn:tokenize` and `fn:replace` to split and normalize names into a “last,first” format, with special handling for suffixes like Jr. or III. Output is generated via the XQuery `output:method` and `output:csv` declarations, producing a comma-separated table.
AI-generated description