The JSTOR Database: Journal Archives, Books, Reports and Collections
A JSTOR scraper works on an archive rather than a news feed. JSTOR, short for Journal Storage, was conceived in 1994 by William G. Bowen, then president of the Mellon Foundation, to digitize the back runs of academic journals; it was founded in 1995 and has been part of the nonprofit ITHAKA, based in New York, since 2009. JSTOR is not a journal but a library of journals: most titles run from volume 1, issue 1 up to the JSTOR moving wall, a delay of 0 to 10 years, usually 3 to 5, which the publisher sets and which advances every January.
In late September 2026 JSTOR's own figures put the archive at more than 12 million journal articles from over 2,800 journals in 75-plus disciplines, next to more than 150,000 scholarly ebooks in Books at JSTOR, research reports from think tanks and millions of images and primary sources. The humanities and social sciences are the core, and the Life Sciences collection adds botany, ecology and ornithology.
Every item has a JSTOR stable URL, and for an article it rests on an integer ID: jstor.org/stable/2626876 is R. H. Coase's The Nature of the Firm, and its JSTOR DOI is 10.2307/2626876. Issue IDs start with i, books with j.ctt or j.ctv, research reports with resrep, images and shared collection items with community, while a chapter adds a sequence number to its book's ID. Browsing follows the same lines - by title, with separate lists of journals, books and research reports, by subject, by publisher and by collection - and that is how JSTOR content is cut into datasets.
Get a Quote