Scrape a news page into rows
Reads the demo newspaper's front page and turns each story into a row: its headline, its link made absolute, its date as ISO-8601 and its section. The site is an api connection, not a URL typed into the node, so which sites the platform reads is decided where connections are. The page is ours, served by the demo stack: no third party is scraped.
Connects to
- HTTP API
How you know it worked
Open the run on the canvas: the node's output is three rows, starting { "headline": "Ferry timetable changes from Monday", "url": "http://localhost:7779/news/ferry-timetable.html", "published": "2026-09-20", … }. The menu, the cookie banner and the "Most read" sidebar are not rows, because only article.story elements are. The run log says '"article.story" matched 3, 3 row(s) kept'.
Try this templateUsed in: Retail & e-commerce
Related templates
Webhook → database
A pipeline that starts when something posts to it. The webhook's JSON body is the run's starting payload, so the first node already has the data —…
Sequence of pipelinesThree pipelines run as one unit (sequence)
A sequence is orchestration above the run machinery: it runs pipelines you already have, in order, feeding each one's output to the next. This one…
Published appOne app, one URL namespace (published app)
A published app groups things you have already published under a single URL — /apps/demo-shop/v1/… — with its own OpenAPI document. This one binds…
Stop moving data by hand.
Start free with three projects, or talk to us about running it across your organisation.