Xiaohongshu (REDnote) Datasets
Unlock data from China's largest lifestyle and consumer decision platform. This dataset collects high-quality UGC content from Xiaohongshu (Little Red Book), delivering deep insights into consumer trends and user behavior. With clean, well-structured data, it's perfectly suited for market research, recommendation algorithm training, and NLP sentiment analysis.
Covers Major Global Sites
Strict GDPR & CCPA Compliance
JSON/CSV Format Testing Available
Flexible Pricing, Pay as You Go
Trusted by over 200 clients worldwide
Available Xiaohongshu (Little Red Book) Datasets
Data is updated daily, structured and cleaned, and supports direct integration through API or file download.
Xiaohongshu Notes & Metadata
Note ID, Title, Description, Media URLs, Engagement Stats.
Xiaohongshu Comments & Sentiment
Comment Text, Likes, IP Location, Nested Replies.
Xiaohongshu User Profiles
User ID, Nickname, Avatar, Xsec Token.
Xiaohongshu Trending Tags
Tag ID, Tag Name, Topic Classification.
Available delivery methods
Maximize ROI on data investment through intelligent strategies
Incremental Update Model
Pay only for new or changed records—no need to repurchase the full database. Reduce acquisition costs with precision.
Multi-Source Data Bundling
Buy one or multiple datasets and unlock exclusive discounts. Get a full cross-platform view in a single purchase—better value, broader coverage.
Enterprise Volume Pricing
Built for high-volume demands. The more you buy, the lower the unit price. Deep discounts on bulk extractions and subscriptions—do more for less.
Data Cleaning & Enrichment
Receive pre-cleaned, deduplicated, and standardized data. No post-processing needed—ready for immediate business analysis, saving time and effort.
Xiaohongshu Notes Content & Engagement Data Sample
The Xiaohongshu Notes dataset captures core UGC content on the platform, including note ID, title, description, posting time, content type (video/image-text), and key engagement metrics (shares, comments, likes, saves). It also includes multimedia resource links (video streams, image lists), serving as foundational data for product seeding effect analysis, content trend mining, and multimodal AI training.
| Name | Description | Type | Example |
|---|---|---|---|
| id | unique to each company | AZ text | highgoal–capital |
| name | The name of the company | AZ text | Highgoal Capital |
| country_code | The country where the company is located | AZ text | GB,EE |
| locations | General information about the company's locations | [ ] array | ["London, GB", "Tallinn, EE"] |
| followers | The number of followers the company has | # number | 41 |
| employees_in_linkedin | The number of employees listed on LinkedIn | # number | 2 |
| about | A description or summary of the company | AZ text | xtHighgoal Capital is a technology focused in... |
No data set found? Start custom collection
Please let us know your specific project requirements, and we will match you with the appropriate data set to help your project land efficiently.
| Name | Description | Type | Example |
|---|---|---|---|
| note_id | Unique identifier for the note | AZ text | 686b353d000000001202f2fe |
| title | Title of the note | AZ text | I have exciting news about my vegetables! |
| desc | Main text content/caption of the note including hashtags | AZ text | I found two tomatoes, one chili pepper... #MyLifeVlog[Topic]# |
| type | Type of the note (video or image) | AZ text | video |
| liked_count | Total number of likes | AZ text | 12K |
| collected_count | Total number of collections/saves | AZ text | 584 |
| comment_count | Total number of comments | AZ text | 1430 |
| share_count | Total number of shares | AZ text | 132 |
| video_url | URL to the video stream (if type is video) | ∞ url | http://sns-video-zl.xhscdn.com/stream/... |
| image_list | List of image URLs (if type is image or video cover) | [ ] list | [{"url": "http://sns-webpic-qc.xhscdn.com/..."}] |
No data set found? Start custom collection
Please let us know your specific project requirements, and we will match you with the appropriate data set to help your project land efficiently.
Dataset Pricing
Buy from a provider with a large scale and high moral standards
Register now and receive a bonus on your first deposit, up to $25.
Starter Plan
Minimum 100K Records
Suited for small-scale validation and initial use
600K Records Included
$840.00 Monthly Plan
Suited for medium-scale monthly needs
2.5M Records Included
$2,800.00 Semi-Annual Plan
Suited for continuously growing data needs
13M Records Included
$10,400.00 Annual Plan
Suited for long-term data solutions at large enterprises
Do you need more than 10 million data or a custom collection solution?
Instantly Empower AI Agents & LLMs
Our datasets are deeply optimized for RAG and model fine-tuning. Clean structure, full documentation, and multi-language SDK examples—seamlessly integrate e-commerce insights into your AI workflows.
Structured Data
Pre-formatted data ready for training and inference with ChatGPT, Claude, and other AI models.
Multi-Language Code Samples
Code snippets in Python, Java, C#, Node.js, and more. No coding from scratch—copy, paste, and build data pipelines in seconds.
Developer Documentation
Comprehensive API references and field definitions that reduce prompt engineering costs for AI-powered data understanding.
CustomXiaohongshu (REDnote) Datasets Tailored to Your Needs
Easy-to-use, fully structured datasets built for diverse business scenarios.
High-Efficiency Data Extraction
Leverage clean residential proxy IPs to extract global site data in one click. 99%+ success rate, zero blocks, billion-scale collection capability.
Multiple Export Formats
Supports JSON, NDJSON, CSV, Parquet, JSON Lines, gzip compression, and more. Integrate seamlessly with your existing systems.
Flexible Payment Models
Flexible pricing, pay as you go. Covers major global sites. Fully GDPR & CCPA compliant—your data stays secure and compliant.
Unlimited Scaling Architecture
Handle massive concurrent requests via high-throughput proxy IPs. Integrates with Snowflake, Google Cloud, SFTP, and more—peak-ready.
Significant Cost Savings
Optimized proxy rotation and data extraction cut costs by 30%+. No self-hosted infrastructure required—focus on growing your business.
Fully Managed Service
We manage the entire data pipeline—including proxy IP maintenance and monitoring. Reduce operational overhead with guaranteed 24/7 uptime.
Seamless API Integration
Simple API interface with Webhook and S3 support. Quickly connect to your e-commerce system—extract ASINs, prices, reviews, and more.
24/7 Professional Support
Dedicated team on standby for custom guidance and troubleshooting. Combined with proxy optimization for worry-free, high-efficiency data collection.
Data Quality Assurance
AI-driven validation ensures accurate, complete, deduplicated data. Real-time monitoring and reporting included—ideal for product analysis, competitor tracking, and inventory management.
Popular Xiaohongshu Datasets
Xiaohongshu Notes Dataset
Includes note titles, body text, hashtags, and core engagement metrics (likes, saves, shares). Ideal for viral trend analysis, "seeding" effectiveness evaluation, and e-commerce product selection..
Xiaohongshu Comments Dataset
Captures comment bodies, reply hierarchies, and user feedback timestamps — essential data for NLP sentiment analysis, authentic consumer review mining, and competitive intelligence gathering.
Xiaohongshu User (KOL) Dataset
Covers creator usernames, IDs, bios, and follower interaction data — enabling influencer profiling, precise KOL screening, and brand campaign ROI estimation.
Xiaohongshu Hashtags Dataset
Aggregates trending topic tags, view counts, and associated note counts — helping brands precisely capture trending topics, optimize content keyword strategy (SEO), and claim high-traffic ground.
Focus on Your Core Business. Leave the Data Collection to Us.
Unlimited Web Scraping
Powered by dynamic residential IPs and intelligent unblocking. Bypass CAPTCHAs and geo-restrictions effortlessly—access data points from public web pages worldwide.
Ready-to-Use, Accurate Data
Every record goes through multi-stage validation and cleaning. Delivery-ready with no post-processing required—directly power your market analysis or AI model training.
Fully Automated Data Pipeline
Scheduled tasks and incremental updates supported. Data auto-delivers to your AWS S3 or database—zero manual intervention from start to finish.
How Companies Use Xiaohongshu Datasets
Viral Discovery & Competitor Analysis
Track "saves" and "likes" on trending notes to precisely capture viral trends in beauty, fashion, and beyond. Deconstruct competitor content strategies and high-frequency hashtags — optimizing your product selection and maximizing "seeding" conversion efficiency.
Deep Review & Pain Point Mining
Apply NLP sentiment analysis at scale to comments — listening to authentic user feedback and "avoidance" warnings. Rapidly identify negative sentiment and uncover unmet core needs to fuel product iteration with data-backed insights.
Precision High-Converting KOL Selection
Say no to fake metrics. Select high-fit creators based on real historical note engagement rates and follower personas. Forecast campaign ROI and ensure every dollar of budget maximizes brand "seeding" impact.