๐ What book discovery data can you collect from Goodreads search results?
This template collects structured records from search-results pages, including Keyword, Progress_result_count, Page_URL, Page, image_url, book_title, author, rating, rating_numbers, published_year, editions. It is useful for publishers, authors, librarians, literary researchers, and book marketers.
Data is collected from Goodreads. Goodreads is a book discovery and reading community with titles, ratings, editions, and reader reviews.
๐ฐ Pricing
This template is currently free of charge and has no per-line usage fee. Octoparse plan or resource limits may still apply.
๐ฆ Output
Goodreads Scraper returns these fields:
KeywordProgress_result_countPage_URLPageimage_urlbook_titleauthorratingrating_numberspublished_yeareditions
{
"Keyword": "education",
"Progress_result_count": null,
"Page_URL": null,
"Page": "1",
"image_url": "https://i.gr-assets.com/images/S/compressed.photo.goodreads.com/books/1375414895i/17851885._SY75_.jpg",
"book_title": "I Am Malala: The Story of the Girl Who Stood Up for Education and Was Shot by the Taliban",
"author": "Malala Yousafzai",
"rating": "4.14",
"rating_numbers": "504,883",
"published_year": "2012",
"editions": "170"
}
๐ฏ Use Cases
- Use rating, rating_numbers to benchmark ratings and reputation.
- Use image_url, book_title to build and compare product, content, or menu catalogs.
- Use author to research creators, reviewers, sellers, or stores.
- Use Page_URL to link records to source pages or downstream detail collection.
๐ Why Octoparse
- Built for Goodreads: Takes the configured Keyword (1 per line), Number of Pages inputs, submits each supplied value and iterates the returned records on Goodreads search-results pages, then emits structured rows containing fields such as Keyword, Progress_result_count, Page_URL, Page, image_url.
- Inputs match the workflow: The form uses Keyword (1 per line) and Number of Pages.
- Fields stay connected: Goodreads Scraper returns
book_title,Page_URL,rating,Keyword, andProgress_result_countin the same structured dataset. - Ready for repeat use: After Goodreads Scraper runs locally or in the cloud, schedule eligible tasks and export the rows for spreadsheets, dashboards, APIs, or AI workflows.
๐ Input
Complete the following fields:
- Keyword (1 per line) (Required) โ Enter book-category or subject keywords, one per line.
- Number of Pages (Optional) โ Set the number of Goodreads search-result pages to process.
๐ How to Use
- Open Goodreads Scraper and click Try it!.
- Complete Keyword (1 per line) and Number of Pages using the formats and limits shown in the Input section.
- Choose a local or cloud run in Octoparse and start the task.
- Review fields such as
book_title,Page_URL, andrating, then export the rows in the format you need.
๐ก Tips
- Start with a focused Keyword (1 per line) so the first run is easy to review before widening the search.
- Keep
Keywordin the export so every row can be traced to its originating input or page.
โ FAQ
What do I enter in Goodreads Scraper?
Complete Keyword (1 per line) and Number of Pages using the accepted values and limits listed in the Input section.
Which Goodreads pages does Goodreads Scraper process?
The workflow is configured for search-result pages and returns fields such as book_title, Page_URL, rating, and Keyword.
What can I use data from Goodreads Scraper for?
Use rating, rating_numbers to benchmark ratings and reputation.
๐ Related Templates
- Goodreads Comments Scraper โ Use Goodreads Comments Scraper when reader review text, ratings, dates, and likes are needed.
- Google Books Scraper โ Use Google Books Scraper for comparable book discovery plus ISBN, publisher, description, subject, and buy-link fields.