Preparing insights
Curating global datasets for you
Core services
Enterprise Data Extraction

Scalable web, app and AI-powered collection across 40+ countries.

All 58 services →
New 2026
AI Training Data

Corpus building with provenance and opt-out compliance.

Learn more →
Free pilot
24-hour sample

We run collection on your own sources before you commit.

Get a sample →
58Services
40+Countries
DEVELOPER

Ready-Made Scrapers

Pre-built for top platforms. Self-serve, no setup.

View All →
TRY FREE

API Playground

Test endpoints instantly. No credit card.

Start Free →
28Tools
2SDKs
icons Delivery & SDKs
Streaming Crawl API Scheduler Realtime Alerts Webhook Delivery 🐍 Python SDK 💚 Node.js SDK
Need it managed instead?

Fixed monthly retainer, named engineer, no per-request metering.

Managed Data API →
HOT

Case Studies

How brands use Actowiz, with named outcomes.

Read →
FREE

Sample Datasets

Real output, no signup.

Download →
NEW

ROI Calculator

Model the return on a data engagement.

Calculate →
Ready to buy · download immediately

Amazon India Daily Product, Pricing & Category Dataset

A daily file covering the Amazon.in catalogue with 25 columns per listing — price, MRP, discount, sold-out flag, shipping charges, promised arrival date, rating and rating count, search position, and a structured three-level category hierarchy.

The column worth knowing about is category_hierarchy. It arrives as JSON with l1, l2 and l3 rather than a breadcrumb string, so grouping by department or sub-category is a parse rather than a guess at where one level ends and the next begins.

This data is available for download immediately after purchase

The file already exists, because this pipeline runs every morning whether you buy it or not. Most vendors start collecting after you order — which is why they quote a lead time. Here the most recent file is in your account within minutes of payment, with an API key issued at the same time.

$499.00per month
All prices are in USD · GST or VAT added on the invoice
2.6M listings per run, across 21 categories
25 fields, all listed below
This dataset was last updated on 20-Sep-2026
A fresh file every morning · cancel any time
Excel and CSV included · JSON and Parquet on request
Please download the sample first and check the fields against what you need. Data files are not refundable once delivered.
VISA · MASTERCARD · AMEX · UPI · NET BANKING

What you get, in plain terms

Four things. If any of them is not what you expected, the sample will show you before you spend anything.

Category tree as structured JSON

category_hierarchy arrives as l1/l2/l3 in one field, so you can cut the file at department or sub-category without matching text against a path string.

Price, MRP and discount, plus their JSON forms

The flat columns are what you will query. The matching _json columns keep the raw captured structure, so nothing is lost if a listing prices unusually.

Search position and ratings

position records where the listing appeared, and rating with rating count sit alongside it — enough to watch visibility and demand move together.

However you want it

Direct download, REST API, Amazon S3, Google Cloud, Snowflake or SFTP. The API key comes with the dataset.

Fields included in this dataset

All 25 columns, in the exact order they appear in the file — this list is taken straight from the sample, not from a brochure. The free sample ships with a data dictionary giving an example value for each one.

product_idcatalog_namecatalog_idsource scraped_dateproduct_nameimage_urlcategory_hierarchy product_pricearrival_dateshipping_chargesis_sold_out discountmrppage_urlproduct_url number_of_ratingsavg_ratingpositioncountry_code othersshipping_charges_jsonproduct_price_jsonmrp_json discount_json

Sample rows from the real file

Real rows from the sample file, not an illustration. The free sample is 50 listings with all 25 columns.

Sample is free, no signup and no card. Check the fields before you buy — files are not refundable once delivered.

Coverage

2.6 million listings across 21 top-level categories, weighted towards the categories where price competition is most active.

Category Listings tracked Median discount Share with ratings
Electronics & accessories 486,000 31% 62%
Home & kitchen 412,000 44% 58%
Beauty & personal care 358,000 39% 54%
Grocery & gourmet 291,000 28% 47%
Fashion 264,000 55% 41%
Sports & fitness 186,000 61% 66%
Garden & outdoor 174,000 52% 38%
All other categories 443,800 36% 44%

Only tracking your own brand and its competitors? Send the ASIN or brand list and we will price a narrower file, which costs considerably less than the full catalogue. We can also move this to a twice-daily capture if you are watching a sale period.

Historical data

The last 30 days come with the dataset. Beyond that we hold Amazon India records from January 2025 onwards — enough to cover two Great Indian Festival cycles, which is usually what people want it for. Tell us the categories and we will confirm exactly what exists.

Ask about historical data

Need more data points?

We can extend this dataset beyond the standard 25 columns — Buy Box winner and seller count is the most requested addition, along with keyword rank, sponsored placement and your own ASIN list captured more often than once a day. Tell us what you would query and we will say whether it is collectable before anything else happens.

Request custom fields

“The category JSON saved us a fortnight. Every other Amazon feed we tried gave us a breadcrumb string and we were writing split rules for each department.”

PS
Priya Sharma · Head of MarketplaceIndian consumer electronics brand · downloaded the sample, bought the same day

What people use this dataset for

Brands selling on Amazon

Track price, discount depth and availability on your own listings and the competitive set, grouped by the same category tree Amazon uses.

Retail and pricing teams

Benchmark online pricing against your other channels, with MRP, discount and shipping charges as separate fields.

Category managers

Watch assortment and discount shift within a sub-category, and see where sold-out listings are concentrated.

Analysts and investors

Category-level price and rating series for demand estimation, share shift and festive-season analysis.

No procurement cycle. Card payment, invoice on the spot. Most buyers are querying data the same afternoon.
Maintenance is ours. When Amazon changes its layout, the file you receive does not change shape.
Analysis-ready. Loads straight into Excel, Power BI, Tableau, BigQuery or a pandas notebook.
Cancel whenever. The feed runs to the end of the period you paid for, then stops. No notice, no fee.

About Amazon India marketplace data

Amazon.in is the largest horizontal marketplace in India by catalogue depth, and the one where price competition is most visible. Its category tree is also unusually deep, which is why the shape of the category field matters more than it sounds.

Why the category hierarchy is JSON rather than a string

A breadcrumb string forces you to guess where one level ends and the next begins, and the separator changes between departments. Here category_hierarchy arrives as an object with l1, l2 and l3 keys, so a group-by on department is a field access rather than a regular expression you have to maintain.

Why empty discount fields are correct

When a listing is not discounted, discount is empty rather than zero. That distinction matters: a zero reads like a measured value, an empty field reads like an absent state. In the sample a fifth of listings sat at full price with the discount column blank, and treating those as zero-percent discounts would understate median discount across the category.

What the JSON price columns are for

product_price_json, mrp_json, discount_json and shipping_charges_json preserve the structure as captured, including cases where a listing prices by variant or bundles shipping oddly. The flat columns are what you will query day to day; the JSON columns are there so nothing is discarded when a listing does something unusual.

Is collecting this data legal?

Collecting publicly visible product and price information is generally lawful in most jurisdictions. Actowiz collects only public pages, respects robots.txt and platform terms, holds no personal data, and aligns with GDPR and CCPA. We are ISO 9001 and ISO 27001 certified, and this dataset carries documented provenance so your legal team can review the source before you buy.

Frequently asked questions

Not in the standard 25 columns. Buy Box winner, seller count and the gap to the next-best offer are the most requested addition on this dataset and we can add them on a custom feed, scoped free.
Where the listing appeared in the result set at the moment of capture. It is useful as a visibility series over time on the same product; it is not a keyword rank, because it is not tied to a specific search term.
The twice-weekly file uses this identical 25-column schema on a lighter schedule. The weekly file is a different, leaner 13-column extract keyed on ZIP code. If you need one schema across cadences, say so and we will align them.
Immediately. The file already exists because the pipeline runs every morning whether you buy it or not. Most vendors start collecting after you order, which is why they quote a lead time — here the current file is in your account within minutes, with an API key.
Yes. 100 real rows with all 25 fields and a data dictionary, no signup and no card. Please use it — checking the fields against your own requirement takes five minutes and prevents almost every problem we see.
Excel and CSV are included. JSON and Parquet are available on request at no extra cost. If you need GeoJSON, Shapefile or something unusual, ask and we will arrange it.
Yes, and the key comes with the dataset at no extra charge. It is unmetered for the data you have bought, with a fair-use rate limit documented in the API reference.
Data files are not refundable once delivered, which is why the sample matters. One exception: if a delivered file does not match the field list on this page, that is our error and we fix it or refund that month.
Any time from your account. The feed keeps running until the end of the period you have paid for, then stops. No notice period and no cancellation fee.
We email you 30 days before any field is added, renamed or removed, and we never remove one inside a period you have already paid for.
The last 30 days are included. Beyond that we hold records for this source from January 2025 onwards — ask and we will confirm exactly what exists for the categories you care about.
Often, yes — Buy Box winner and seller count, your own ASIN list prioritised, keyword rank and sponsored placement are common additions. That becomes a custom feed rather than this dataset, and we scope it free before quoting.
Internal use across your whole organisation with unlimited seats. Redistributing the data, reselling it, or embedding it in a product you sell needs a separate licence — tell us what you have in mind.
Of course, though nothing on this page requires it. If you would rather have a call before buying, use the contact form and ask for the datasets team.

Related datasets, same field names

These join to the file above without any reconciliation work, which is why most buyers take more than one.

Get in Touch
Let's Talk About
Your Data Needs
Tell us what data you need — we'll scope it for free and share a sample within hours.
  • Free Sample In 2 Hours
    Free Sample in 2 HoursShare your requirement, get 500 rows of real data — no commitment.
  • Plans From $500
    Plans from $500/monthFlexible pricing for startups, growing brands, and enterprises.
  • US Based Support
    US-Based SupportOffices in New York & California. Aligned with your timezone.
  • ISO 9001 & 27001 Certified
    ISO 9001 & 27001 CertifiedEnterprise-grade security and quality standards.
Request Free Sample Data
Fill the form below — our team will reach out within 2 hours.
+1
Free 500-row sample · No credit card · Response within 2 hours