Core services
Enterprise Data Extraction

Scalable web, app and AI-powered collection across 40+ countries.

All 58 services →
New 2026
AI Training Data

Corpus building with provenance and opt-out compliance.

Learn more →
Free pilot
24-hour sample

We run collection on your own sources before you commit.

Get a sample →
58Services
40+Countries
DEVELOPER

Ready-Made Scrapers

Pre-built for top platforms. Self-serve, no setup.

View All →
TRY FREE

API Playground

Test endpoints instantly. No credit card.

Start Free →
28Tools
2SDKs
icons Delivery & SDKs
Streaming Crawl API Scheduler Realtime Alerts Webhook Delivery 🐍 Python SDK 💚 Node.js SDK
Need it managed instead?

Fixed monthly retainer, named engineer, no per-request metering.

Managed Data API →

A client needing a deployable Apify Python/Playwright Actor to scrape all product categories from IGA's online grocery store — name, price, and real image URL — reliably on the Apify cloud.

Industry
Grocery • E-commerce Data
Region
Canada
Stack
Apify Actor • Python • Playwright
Output
CSV + Apify Dataset
All
Grocery Categories
Apify
Cloud-Deployable Actor
CSV +
Apify Dataset
Deduplicated
Clean Product Rows

Client Overview

The client needed a robust, production-grade scraper for IGA's online grocery store, built specifically as an Apify Actor in Python using Playwright. The goal was to extract every product from every grocery category and subcategory on the site — capturing product name (Nom), price (Prix), and the real product image URL (not a placeholder).

The deliverable was not just data but a deployable, reusable Actor — complete with the correct project files — that runs reliably on the Apify cloud so the client can re-run it whenever needed.

The Challenge

  • Dynamic, lazy-loaded content. Products load via scrolling and lazy-loaded images (src, data-src, srcset), so simple HTML fetches miss most of the catalog.
  • Deep category navigation. The full category and subcategory tree had to be discovered via the sidebar/flyout menu and traversed completely.
  • Long page loads. Some pages take 60–80 seconds to load, demanding robust waits and retries.
  • Pagination & duplicates. Category pages paginate, and products can repeat across pages — requiring pagination handling and de-duplication.
  • Real images only. Placeholder images had to be avoided; the genuine product image URL captured.
  • Production-grade delivery. It had to be a fully functional Apify Actor with correct project files, running reliably headless on Apify cloud.

The Solution by Actowiz Solutions

Actowiz built a resilient Apify Actor in Python with Playwright, engineered around IGA's dynamic rendering and delivered with the full Apify project structure.

What the Actor Does
  • Category discovery. Navigates the root grocery page and interacts with the sidebar/flyout menu to enumerate every category and subcategory.
  • Full product loading. Scrolls and waits on each category page until all products load, handling 60–80-second load times gracefully.
  • Robust extraction. Captures Nom, Prix, and the real image URL, resolving lazy-loaded images from src / data-src / srcset.
  • Pagination & dedup. Handles pagination at the bottom of category pages and removes duplicates across pages.
  • Structured output. Writes products_iga.csv and saves to the Apify Dataset; category name stored on every product row, with optional per-category export.
  • Progress logging. Logs category names and product counts for visibility during runs.
Deliverables & Technical Structure
Component Detail
main.py Playwright-based Actor logic — navigation, scrolling, extraction, pagination
requirements.txt Python dependencies for reproducible builds
.actor/actor.json Apify Actor configuration
products_iga.csv Primary output — all products across categories
Apify Dataset Structured dataset output on the Apify platform
Per-category export Optional separate CSV/JSON per category
Fields Extracted
Field Description
Nom (Name) Product name
Prix (Price) Product price
Image Real product image URL (not placeholder)
Category Category/subcategory the product belongs to

Quality Assurance

Validation Check Rule Applied
Category completeness Every category and subcategory traversed
Full product load Scroll/wait confirmed all products rendered before extraction
Real image capture Genuine image URL resolved from src/data-src/srcset; no placeholders
Deduplication No duplicate products across pages
Cloud reliability Actor validated to run reliably headless on Apify cloud
Output integrity products_iga.csv and Dataset aligned; category on every row

Results & Business Impact

  • Complete catalog capture. Every product across all IGA categories extracted, despite heavy dynamic loading.
  • Reusable, deployable Actor. A production-grade Apify Actor the client can re-run on demand from the cloud.
  • Clean, real-image data. Genuine product images and deduplicated rows produced trustworthy output.
  • Operational visibility. Progress logs made large runs transparent and easy to monitor.

Why the Client Chose Actowiz Solutions

  • Apify + Playwright expertise. A production Actor built to platform standards, not a one-off script.
  • Dynamic-site resilience. Scrolling, long waits, and lazy-image handling engineered in.
  • Correct project structure. main.py, requirements.txt, and .actor/actor.json delivered and cloud-ready.
  • Clean, deduplicated output. CSV and Dataset with category on every row.

Project at a Glance

Metric Value
Industry Grocery • E-commerce Data
Region Canada
Target IGA online grocery — all categories
Stack Apify Actor • Python • Playwright
Fields Nom, Prix, Image (real URL), Category
Edge Cases Lazy images, 60–80s loads, pagination, duplicates
Deliverables main.py, requirements.txt, .actor/actor.json, products_iga.csv, Dataset
Deployment Runs reliably on Apify cloud (headless)

Client Feedback

"It's a proper Apify Actor, not a fragile script — it walks every category, waits out the slow pages, grabs the real images, and drops a clean CSV plus the Dataset. I can re-run it from the cloud whenever I need fresh data."

— Developer / Data Lead, Client Team

Need a custom data pipeline for your platform?

Actowiz Solutions designs custom, large-scale scraping, extraction, and delivery pipelines with rigorous QA. Visit actowizsolutions.com to discuss your data requirement.

Social Proof That Converts

Trusted by Global Leaders Across Q-Commerce, Travel, Retail, and FoodTech

Our web scraping expertise is relied on by 4,000+ global enterprises including Zomato, Tata Consumer, Subway, and Expedia — helping them turn web data into growth.

4,000+ Enterprises Worldwide
50+ Countries Served
20+ Industries
Join 4,000+ companies growing with Actowiz →
Real Results from Real Clients

Hear It Directly from Our Clients

Watch how businesses like yours are using Actowiz data to drive growth.

1 min
★★★★★
"Actowiz Solutions offered exceptional support with transparency and guidance throughout. Anna and Saga made the process easy for a non-technical user like me. Great service, fair pricing!"
TG
Thomas Galido
Co-Founder / Head of Product at Upright Data Inc.
2 min
★★★★★
"Actowiz delivered impeccable results for our company. Their team ensured data accuracy and on-time delivery. The competitive intelligence completely transformed our pricing strategy."
II
Iulen Ibanez
CEO / Datacy.es
1:30
★★★★★
"What impressed me most was the speed — we went from requirement to production data in under 48 hours. The API integration was seamless and the support team is always responsive."
FC
Febbin Chacko
-Fin, Small Business Owner
icons 4.8/5 Average Rating
icons 50+ Video Testimonials
icons 92% Client Retention
icons 50+ Countries Served

Join 4,000+ Companies Growing with Actowiz

From Zomato to Expedia — see why global leaders trust us with their data.

Why Global Leaders Trust Actowiz

Backed by automation, data volume, and enterprise-grade scale — we help businesses from startups to Fortune 500s extract competitive insights across the USA, UK, UAE, and beyond.

icons
7+
Years of Experience
Proven track record delivering enterprise-grade web scraping and data intelligence solutions.
icons
4,000+
Projects Delivered
Serving startups to Fortune 500 companies across 50+ countries worldwide.
icons
200+
In-House Experts
Dedicated engineers across scrapers, AI/ML models, APIs, and data quality assurance.
icons
9.2M
Automated Workflows
Running weekly across eCommerce, Quick Commerce, Travel, Real Estate, and Food industries.
icons
270+ TB
Data Transferred
Real-time and batch data scraping at massive scale, across industries globally.
icons
380M+
Pages Crawled Weekly
Scaled infrastructure for comprehensive global data coverage with 99% accuracy.

AI Solutions Engineered
for Your Needs

LLM-Powered Attribute Extraction: High-precision product matching using large language models for accurate data classification.
Advanced Computer Vision: Fine-grained object detection for precise product classification using text and image embeddings.
GPT-Based Analytics Layer: Natural language query-based reporting and visualization for business intelligence.
Human-in-the-Loop AI: Continuous feedback loop to improve AI model accuracy over time.
icons Product Matching icons Attribute Tagging icons Content Optimization icons Sentiment Analysis icons Prompt-Based Reporting

Connect the Dots Across
Your Retail Ecosystem

We partner with agencies, system integrators, and technology platforms to deliver end-to-end solutions across the retail and digital shelf ecosystem.

icons
Analytics Services
icons
Ad Tech
icons
Price Optimization
icons
Business Consulting
icons
System Integration
icons
Market Research
Become a Partner →

Popular Datasets — Ready to Download

Browse All Datasets →
icons
Amazon
eCommerce
Free 100 rows
icons
Zillow
Real Estate
Free 100 rows
icons
DoorDash
Food Delivery
Free 100 rows
icons
Walmart
Retail
Free 100 rows
icons
Booking.com
Travel
Free 100 rows
icons
Indeed
Jobs
Free 100 rows

Latest Insights & Resources

View All Resources →
thumb
Blog

How to Overcome Competitor Price and Availability Gaps with Tyres Categories Data Collection from Lazada and Tuhu App

Tyres Categories data collection from Lazada and Tuhu App helps businesses track tyre prices, brands, availability, and assortment for market insights.

thumb
Case Study

How We Empowered a Leading Food Brand Using Scrape Ready-to-Cook Cut Veg Product Data from Blinkit TN for Smarter Product & Pricing Decisions

Track Scrape Ready-to-Cook Cut Veg Product Data from Blinkit TN to monitor prices, availability, SKUs, and trends for smarter retail insights.

thumb
Report

Brazil Car Rental Pricing Intelligence Report 2026

Brazil Car Rental Pricing Intelligence Report 2026 reveals rental price trends, market shifts, competitor rates, and opportunities for smarter pricing.

Start Where It Makes Sense for You

Whether you're a startup or a Fortune 500 — we have the right plan for your data needs.

icons
Enterprise
Book a Strategy Call
Custom solutions, dedicated support, volume pricing for large-scale needs.
icons
Growing Brand
Get Free Sample Data
Try before you buy — 500 rows of real data, delivered in 2 hours. No strings.
icons
Just Exploring
View Plans & Pricing
Transparent plans from $500/mo. Find the right fit for your budget and scale.

Request Free Sample Data

Our team will reach out within 2 hours with 500 rows of real data — no credit card required.

+1
Free 500-row sample · No credit card · Response within 2 hours