To scrape a news article cleanly, work in order: read the headline, author and date from the page's JSON-LD, isolate the single node that holds the article prose, expand any "continue reading" control ...
Some of the entries in this section require checking a repository's commit history to notice the maintenance question. pyppeteer doesn't require that. Its own README opens with: "this repo is ...