📌 What Xiaohongshu post results can you collect by keyword?
This template collects search-result records with Title, Image URL, Details page URL, Author URL, Datetime, Author, Input keywords, and Like count. It is useful for social-media researchers, content teams, brand analysts, and audience researchers.
Data is collected from Xiaohongshu. Xiaohongshu is a social discovery platform centered on user posts, products, lifestyle topics, and recommendations.
💰 Pricing
This template is currently free of charge and has no per-line usage fee. Octoparse plan or resource limits may still apply.
📦 Output
The current published implementation can return the following fields:
Input_keywordsTitleImage_URLDetails_page_URLAuthorDatetimeAuthor_URLLike_countRecommended_reason
{
"Input_keywords": "web scraping",
"Title": "推荐3款自动爬虫神器,再也不用手撸代码了",
"Image_URL": "https://sns-webpic-qc.xhscdn.com/202511101556/e6aa53f42ed33364e0afd76422cd9bc3/1040g008314t3depl14005npqjgqg8e7a9v66v7g!nc_n_webp_mw_1",
"Details_page_URL": "https://www.xiaohongshu.com/search_result/66889d19000000000a026420?xsec_token=ABNlEB76L3vQ5N6cIYO-nnMuxpO0YslweSJOt3wLL32-8=&xsec_source=",
"Author": "Python大数据分析",
"Datetime": "2024-07-06",
"Author_URL": "https://www.xiaohongshu.com/user/profile/5f3a9c3500000000010038ea?channel_type=web_search_result_notes&parent_page_channel_type=web_profile_board&xsec_token=AB6uZC_94FYmu431pd_4pk8yP8HtOlqXAyi_kjD4UVbiA=&xsec_source=pc_search",
"Like_count": "511",
"Recommended_reason": null
}
🎯 Use Cases
- Build a structured research dataset organized by Input keywords and Title.
- Analyze source content and timing through Author, Author URL, and Datetime.
- Measure audience or customer engagement using Like count.
- Use Image URL, Details page URL, and Author URL to audit source records or feed verified pages into downstream workflows.
🐙 Why Octoparse
- Ready-to-use workflow: The extraction steps for Xiaohongshu are already configured, so you do not need to build the scraper from scratch.
- Flexible execution: Choose the supported local or cloud run mode to fit one-off checks or repeatable collection work.
- Structured, repeatable output: Results are returned as consistent rows that are easier to compare, filter, deduplicate, and process than manually copied pages.
- Verified input guardrails: The form exposes the current inputs, selectable values, and meaningful limits configured for this template.
- Practical data handoff: Review results in Octoparse and export or process them using the options supported by your Octoparse environment.
📝 Input
Complete the following fields:
- Pre-login (Required) — Enable pre-login before searching Xiaohongshu.
- Keywords (Required) — Enter Xiaohongshu search keywords, one per line. Up to 100 entries per run.
🚀 How to Use
- Open the template and select Try it or Start.
- Complete the input fields listed above.
- Start the task using the supported local or cloud run mode.
- Review the output rows and export or process the structured data.
⚠️ Limitations
Results depend on what the source site exposes at run time. A listed output field can be empty when the source page does not provide that value. Input limits shown above are enforced by the template.
💡 Tips
Use specific, valid inputs and review a representative result before starting a large batch. Remove duplicate inputs when repeated records are not needed.
❓ FAQ
How does this scraper work?
The automation applies Pre-login and Keywords (up to 100 per run), processes the resulting search pages and their result items, and writes structured rows including Input keywords, Title, Image URL, and Details page URL.
What do I need to enter?
Use the fields and accepted values shown in the Input section. Only user-relevant limits and selectable options are listed.
What data does this template return?
The current output fields and a representative JSON Data Preview are listed in the Output section. Field availability can vary when the source page does not display a value.
Is this template free to use?
It is currently free of charge and has no per-line usage fee. Octoparse plan or resource limits may still apply.
Why can some output fields be empty?
Source pages do not always expose every value for every record, and layouts can vary by item, market, or current site response. The template returns a field when the current page provides it.
Can I export the collected data?
You can review the structured rows in Octoparse and export or process them using the options supported by your Octoparse environment.
🔗 Related Templates
- Xiaohongshu Post Details Scraper — Use Xiaohongshu Post Details Scraper to enrich discovered URLs with detail-page data.
- Xiaohongshu Search Results Scraper (by URL) — Use Xiaohongshu Search Results Scraper (by URL) for a complementary workflow on the same source site.
- Twitter (X) Comments Scraper — Use Twitter (X) Comments Scraper to collect comparable social media data from another source.