Why the licensed route is also the better dataset
It is tempting to see licensing as the expensive alternative to scraping. For this source it is simply the better product, and the comparison is not close.
A crawl of a paywalled site, even where permitted, collects pages as they happen to render: current versions, whatever the page template shows, whatever the paywall lets through. A licensed archive delivers complete articles, with metadata applied consistently and corrections handled, going back as far as the archive does.
For the questions people actually bring - how a company was covered over years, what was reported before a market event, how coverage of a sector changed - completeness and consistency are the whole point. A partial, template-dependent crawl would answer them badly even if it were allowed.
And it is not allowed: the site asks for authentication, and the crawl rules exclude AI training crawlers by name. We do not treat 401 as something to engineer around. The honest route and the high-quality route are the same route here.