Skip to content

Scrape a news page into rows

Reads the demo newspaper's front page and turns each story into a row: its headline, its link made absolute, its date as ISO-8601 and its section. The site is an api connection, not a URL typed into the node, so which sites the platform reads is decided where connections are. The page is ours, served by the demo stack: no third party is scraped.

Connects to

  • HTTP API

How you know it worked

Open the run on the canvas: the node's output is three rows, starting { "headline": "Ferry timetable changes from Monday", "url": "http://localhost:7779/news/ferry-timetable.html", "published": "2026-09-20", … }. The menu, the cookie banner and the "Most read" sidebar are not rows, because only article.story elements are. The run log says '"article.story" matched 3, 3 row(s) kept'.

Try this template

Used in: Retail & e-commerce

Stop moving data by hand.

Start free with three projects, or talk to us about running it across your organisation.