
Completed
Posted
Paid on delivery
I need data from Ikea website : [login to view URL] Anti bot : Datadome I have the URL of the each product. There will be 125 000 products to scrap. Here is JSON format : { "url": "string", // /ref/<id> page — PRIMARY MATCH KEY "ref": "string", // the number in the URL (/ref/1182158) "ean": "string", // gtin13 — image-gallery key + secondary match key "name": "string", "brand": "string", "description": "string", "price": number, // in EUR "currency": "string", // "EUR" "availability": "string", // "InStock" | "OutOfStock" "condition": "string", // "NewCondition" | "RefurbishedCondition" "rating": number|null, // out of 5, if present "reviewCount": number|null, "category_path": ["string"], // breadcrumb, general → specific "images": ["string"], // ALL scene7 gallery image URLs "weight_kg": number|null, // from "Poids Net" / "Poids Brut" "dimensions": "string|null", // from "Dimensions l x h x p" "packaging_dimensions": "string|null", // from "Dimensions emballé l x h x p" (parcel) "attributes": { // ALL technical spec name:value pairs "<spec name>": "<value>" } } Here is a JSON product example : { "url": "[login to view URL]", "ref": "1182158", "ean": "5028683082323", "name": "Piano de cuisson gaz FALCON Classic 110 Gaz Blanc", "brand": "FALCON", "description": "Professional-grade range cooker with multiple cooking modes thanks to its 3 ovens and 5-burner gas hob…", "price": 6433.64, "currency": "EUR", "availability": "InStock", "condition": "NewCondition", "rating": null, "reviewCount": null, "category_path": ["Accueil", "Gros électroménager", "Cuisson", "Piano de cuisson"], "images": [ "[login to view URL]", "[login to view URL]", "[login to view URL]" ], "weight_kg": 159, "dimensions": "109.2 x 93 x 66 cm", "packaging_dimensions": "120.5 x 115 x 73 cm", "attributes": { "Type de table": "gaz", "Puissance": "15.700 W", "Nombre de foyers total": "5.0", "Energie du four": "électrique", "Volume de la cavité": "79 L", "Dimensions de la cavité l x h x p": "46.7 x 43.8 x 38.5 cm", "Nettoyage": "catalyse", "Coloris": "Blanc", "Finition": "Chrome brossé", "Poids Net": "159 Kg", "Dimensions l x h x p": "109.2 x 93 x 66 cm", "Dimensions emballé l x h x p": "120.5 x 115 x 73 cm" } }
Project ID: 40535135
38 proposals
Remote project
Active 6 days ago
Set your budget and timeframe
Get paid for your work
Outline your proposal
It's free to sign up and bid on jobs

Hi, I can help with this project. I have experience building large-scale e-commerce scrapers, including Amazon and other anti-bot protected sites, using Playwright, Selenium, and Python. For your requirements, I can: ✔ Process 125,000 product URLs ✔ Extract all fields in the exact JSON structure provided ✔ Handle DataDome-protected pages with a scalable scraping architecture ✔ Capture all images, attributes, dimensions, pricing, availability, ratings, and category paths ✔ Implement retry logic, error handling, and deduplication ✔ Deliver clean JSON output and/or database storage Before confirming the final approach, I would like to review a small sample of product URLs to evaluate the current DataDome protection level and determine the most reliable extraction strategy. I have worked on high-volume scraping projects before and focus on data accuracy, stability, and maintainability. Looking forward to discussing the details. Best regards, Avinash
€10 EUR in 1 day
5.7
5.7
38 freelancers are bidding on average €51 EUR for this job

As a well-established firm with expertise primarily in web development and design, we have gained a thorough understanding of retrieving, manipulating, and parsing data in various formats -- including structured JSON. We've successfully tackled projects involving immense data scraping and our team's technical strength lies in precisely automating such processes at scale. With a firm grasp on PHP, we possess the skillset to leverage this high-level programming language for your benefit through strategic coding. Our team is not only proficient with tools like "DataDome", but our experience in scraping data from large e-commerce websites could significantly benefit your project. We have the resourcefulness and determination to navigate any anti-scraping mechanisms that may be put in place to retrieve the 125000 data points required from Ikea's specific website you've identified. The JSON example you've shared aligns perfectly with what we've done before -- accurately and efficiently extracting targeted information while considering each field's hierarchy and relationship. Given the scope of the project, choosing us will ensure a resourceful, quality-oriented approach that aims to meet your precise requirements within set deadlines. No wonder our clients vouch for us as "dream-catchers"! Would love to embark on this exciting project journey together and help you rise further heights with your ecommerce venture!
€70 EUR in 1 day
8.6
8.6

⭐⭐⭐⭐⭐ Efficient Data Scraping from Boulanger Website ❇️ Hi My Friend, I hope you're doing well. I've reviewed your project requirements and noticed you're looking for data scraping from the Boulanger website. Look no further; Zohaib is here to assist you! My team has successfully completed over 50 projects related to web scraping. I will use effective techniques to gather data from the 125,000 products you mentioned while overcoming the anti-bot measures in place. ➡️ Why Me? I can easily handle your data scraping needs as I have 5 years of experience in web scraping, specializing in handling large datasets and overcoming anti-bot systems. My expertise includes Python, data extraction, and JSON formatting. Additionally, I have a strong grip on technologies like Selenium and BeautifulSoup, ensuring a smooth scraping process. ➡️ Let's have a quick chat to discuss your project in detail and let me show you samples of my previous work. Looking forward to discussing this with you in our chat. ➡️ Skills & Experience: ✅ Web Scraping ✅ Python Programming ✅ Data Extraction ✅ JSON Formatting ✅ Selenium ✅ BeautifulSoup ✅ API Integration ✅ Data Cleaning ✅ Error Handling ✅ Automation ✅ Data Analysis ✅ Anti-Bot Solutions Waiting for your response! Best Regards, Zohaib
€12 EUR in 2 days
8.2
8.2

With extensive experience in web scraping, particularly from complex and anti-bot protected sites, I am fully equipped to undertake your project of Boulanger product data scraping from the diverse and challenging URLs you have provided. I am highly skilled in Python, using libraries such as BeautifulSoup, Scrapy, and Requests for efficient data extraction. Additionally, I can deliver your data in the precise format you require be it CSV, Excel or JSON. With my web-scraping expertise and adaptability in working with different platforms like WooCommerce and Shopify for e-commerce, I guarantee to extract and deliver all 125,000 products accurately. My services extend beyond data scraping - from internet research to lead generation, contact information collection to data entry and formatting - making me an ideal choice for a project of this scale. I also have experience handling large-scale projects and working with JSON formats from my work with batch importing on Magento, OpenCart, etc. For this particular project of scraping Ikea website, I am familiar with the DataDome hurdle that we would need to overcome.
€150 EUR in 4 days
7.3
7.3

<<<✔Consider it DONE✔>>> YO! I understand your project and I'm eager to help. With my extensive experience in web development and proficiency in PHP, I am well-equipped to handle your Boulanger product data scraping project. I have a great understanding of complex data structures and your JSON format will be no problem for me to work with. Additionally, my analytical skills and meticulous attention to detail ensure accurate and reliable data extraction. In regards to the scale of your project, my past experience with handling large volumes of data will prove invaluable. I understand that the job can be time-consuming and challenging, but I guarantee a timely deliverable without compromising on quality. Finally, my commitment to consider the bigger picture sets me apart. I am not just interested in carrying out the task as is; rather, I aim to ensure that the solution aligns with and contributes to your business goals. This means providing clean, usable data in a format that's ready for analysis or any other purpose you intend to use it for. Let me take on this project for you and undoubtedly exceed your expectations at every step! Looking forward to being part of your project! You will surely be impressed by my work! Not sure what the next step is? I offer free and professional consultation -- I'm just a text away. All the very best, Josh
€19 EUR in 2 days
5.5
5.5

Hi there, I see you need a way to scrape data from the Boulanger website for 125,000 products, while dealing with anti-bot measures like Datadome. With 4+ years of experience in web scraping and data extraction, I can create a tailored solution to gather the required data in the specified JSON format. My approach would involve developing a robust scraper that handles the anti-bot technology effectively, ensuring we retrieve all necessary product details like name, price, and specifications without getting blocked. I’ll also make sure the data is well-structured and ready for any further processing you might need. One question I have is whether you have specific requirements for handling the scraped data, such as storage or analysis, after it’s collected? Best regards, Arslan Shahid
€8 EUR in 1 day
5.7
5.7

Hello, I can build a Python scraper for Boulanger product pages and export the data in the exact JSON structure you provided. ✔ Supports 125,000 product URLs ✔ Extracts all product details, images, specifications, ratings, dimensions, and EAN ✔ JSON export matching your schema ✔ Retry logic, logging, and resume support ✔ Clean, documented source code I have experience with large-scale e-commerce scraping and handling protected websites. Delivery: 2–3 days. Please send a few sample URLs for testing.
€73 EUR in 3 days
5.6
5.6

Hi, I have reviewed your project description and this is exactly the kind of large-scale structured eCommerce data extraction and product catalog engineering work I enjoy working on, especially projects involving high-volume scraping, anti-bot handling, and clean structured JSON output pipelines. I understand the goal is to reliably extract 125,000+ product records with full technical attributes, images, pricing, and category hierarchies in a clean and consistent format. I have worked on similar large-scale product scraping, data engineering, and eCommerce catalog structuring projects and have 6+ years of experience in Python-based web scraping, proxy rotation systems, Datadome/anti-bot handling strategies, and building scalable extraction pipelines with clean JSON normalization and validation. You can check my portfolio for similar work. Two quick questions before we begin: Do you want this delivered as a one-time full dataset export or a continuously updating scraper pipeline? Do you already have access to proxies/CAPTCHA-solving infrastructure, or should I set up a complete anti-bot bypass architecture as part of the system? Regards, Solves Inn
€8 EUR in 1 day
5.3
5.3

Hi, I can scrape all 125,000 product URLs and deliver the data in your exact JSON structure, including specifications, breadcrumbs, Scene7 image gallery URLs, pricing, availability, ratings, dimensions, and all technical attributes. For a site protected by DataDome, I would use a robust scraping architecture with Playwright, rotating proxies, request throttling, retry handling, and structured data extraction to ensure stability at scale. Will you provide the complete list of 125,000 product URLs, and do you already have DataDome-compatible proxies available, or should I include that in the solution? Thanks.
€19 EUR in 1 day
5.0
5.0

I’ve built high-volume anti-bot scrapers against Datadome-protected ecommerce sites in Python with Playwright, so this is right in my wheelhouse. My approach: I’ll spin up a headless cluster with rotating residential proxies, bypass Datadome via Playwright stealth, fetch product pages at ~1 req/sec to avoid rate limits, parse JSON-LD and DOM for the schema fields, normalize attributes into key-value pairs, validate EAN/GTIN matches against scene7 image paths, and store raw JSON in S3 with SHA-256 hashes for reproducibility. I’ll log 429/403 responses, retry with exponential backoff, and include a lightweight SQLite checksum table to track progress and detect gaps. I can start immediately.
€19 EUR in 1 day
4.3
4.3

Hi there! You are building a 125k product extraction pipeline from Boulanger and the real challenge is Datadome protection with stable structured JSON mapping without data loss. I recently built a large scale scraper for an EU retail catalog with 80k+ SKUs using Python Playwright and proxy rotation, keeping 99% structured field accuracy across runs. I also optimized extraction pipelines to reduce blocking and improve throughput for long crawling sessions. I will design a resilient scraper with rotating sessions, retry logic, and clean field mapping into your exact JSON schema while handling anti bot layers safely. I will ensure incremental saving so the system never loses progress even on long runs. Check our work: https://www.freelancer.com/u/ayesha86664 How do you want updates handled for 125k items, batch JSON files or a live database pipeline? I am ready to start just say the word. Best Regards, Ayesha
€19 EUR in 4 days
4.3
4.3

Hi! I can build a robust scraper to extract all 125k products from Boulanger while bypassing DataDome using smart request handling, proxies, and browser automation if needed. I’ve worked on large-scale scraping projects with anti-bot protections, delivering clean structured data (JSON/DB) with high accuracy and stability. I’ll ensure efficient, scalable extraction with proper error handling and provide the data in your preferred format along with clear documentation.
€19 EUR in 1 day
3.9
3.9

Hello, I've carefully reviewed your requirements, and the real challenge is not scraping 125,000 products but doing it reliably despite DataDome protection while maintaining data completeness and extraction stability at scale. The biggest risks are incomplete product payloads, IP blocking, session invalidation, and long running job failures that create gaps in the dataset. The solution should be built as a distributed scraping pipeline using Playwright with session management, proxy rotation, retry logic, structured logging, duplicate detection, and incremental checkpoints. Experience with large scale e commerce extraction, anti bot protected websites, data validation, and cloud based scraping infrastructure helps ensure the process remains reliable across hundreds of thousands of product pages. Send over a sample of the product URLs and the exact fields you need extracted, and I will review the target structure and estimate the best extraction strategy. I can also assess the anti bot challenges and propose the most stable architecture for processing all 125,000 products efficiently. I look forward to discussing the project in more detail and defining the extraction workflow. Best regards
€15 EUR in 3 days
3.9
3.9

Hi, I can build an efficient Python scraper for Boulanger product pages and export the data in your exact JSON structure. Since you already have all product URLs, I would focus on a URL-driven pipeline rather than category crawling. The scraper would read your list, visit each `/ref/<id>` page, extract product JSON/structured data where available, then fall back to HTML parsing for technical specs, dimensions, weight, breadcrumbs, availability, reviews, and Scene7 images. For 125,000 products, I would build it with: • Resume support so failed runs continue safely • Rotating proxy/session handling for DataDome • Rate limiting and retry logic • Error logging for blocked/missing pages • JSONL/CSV export options • Exact fields: ref, EAN, name, brand, price, stock, condition, images, specs, dimensions, category path • Validation against your sample schema I would first run a small test batch of 50–100 URLs to confirm extraction accuracy and anti-bot handling, then scale the job in batches. One note: DataDome protection must be handled carefully, so I would need to confirm whether you already have proxies/cookies/session access available. Regards, Chaz C.
€19 EUR in 1 day
3.7
3.7

Hello! I am available immediately and excited to tackle the challenge of scraping product data from the Ikea website, specifically dealing with the anti-bot system Datadome for 125,000 products. I understand the complexity of gathering data in JSON format as outlined in your project description. Key Points: - Expertise in web scraping and data extraction - Proficient in handling large datasets efficiently - Experience with JSON data structuring and manipulation Capabilities: - Skilled in web scraping tools and techniques - Successfully completed similar projects with accurate results I am eager to discuss how I can assist you further with this project. Let's connect and explore the possibilities together.
€8 EUR in 7 days
3.9
3.9

Having successfully built a variety of automation systems, I am confident in my ability to deliver exactly what you need from this product data scraping project. I have extensive experience in web scraping and working with JSON data, which aligns perfectly with the requirements of this task. Plus, with a deep understanding of API integrations, I can navigate through complex authentication processes like Datadome proficiently. Over the course of my 6+ years as a full stack developer, I've become adept at creating efficient and scalable systems to handle tasks just like these. Given the large volume of products you require (125 000!), my proficiency in performance-driven coding methodologies would prove extremely valuable in producing accurate and reliable results in a timely manner. Additionally, my broad experience working with databases including PostgreSQL, MySQL, MongoDB, and Redis ensures I can deliver the data exactly how you need it. What sets me apart is not only my technical skills but also my strong emphasis on quality and client satisfaction. I'm focused on delivering production-ready solutions that truly save time and increase workflow efficiency - that's precisely why I'm so fit for your project. Choose me and rest easy knowing that your project will be handled by a seasoned professional who won't just get things done, but will deliver them exceptionally well. Let's bring your product data requirements to life!
€19 EUR in 2 days
3.3
3.3

Hi there, I understand that you need to scrape data from the Boulanger website, specifically targeting 125,000 products while overcoming the Datadome anti-bot measures. I have the skills required to effectively handle such complex web scraping projects. I am Abdul Haseeb Siddiqui, with over 6 years of experience in PHP, Python, Data Processing, Web Scraping, Software Architecture, and Data Analysis. My expertise ensures accurate data extraction formatted as per your JSON requirements. You can find my portfolio here: https://www.freelancer.com/u/haseebsidd07 I am confident that I can deliver high-quality results for your project, turning your data requirements into a powerful resource. Thank you for considering my proposal. Regards, Abdul Haseeb Siddiqui
€8 EUR in 7 days
3.4
3.4

GIVE ME 30 SECONDS TO SHOW YOU WHY I'M THE RIGHT FIT FOR THIS PROJECT. I successfully completed a similar project scraping product data for an e-commerce platform, delivering accurate and structured results that exceeded expectations. Relevant experience includes expertise in web scraping, JSON data handling, and working with large datasets, aligning perfectly with the requirements of this project. Understanding your goal of efficiently scraping 125,000 product data from the Boulanger website, I will meticulously extract the specified fields in JSON format, ensuring data accuracy and completeness. The difference between an average result and an exceptional one is usually decided before the work even begins. Regards, Patrick
€15 EUR in 7 days
2.5
2.5

Greetings, I see you’re looking for a reliable way to scrape data from the Boulanger website, specifically targeting around 125,000 products. This sounds like a significant project, and I can help you extract the detailed information you need while navigating the anti-bot measures like Datadome. To tackle this, I would utilize efficient web scraping techniques, ensuring that each product’s data is captured accurately in the specified JSON format. My experience with Python and data processing will help me create a robust solution that respects the website's terms while delivering the required information. With a focus on clean and efficient code, I can ensure that the scraper runs smoothly and delivers reliable data. I’m committed to providing you with a comprehensive dataset that meets your needs.
€19 EUR in 7 days
2.6
2.6

Hello, The real challenge isn't extracting 125,000 fields—it's navigating Datadome's anti‑bot protections at scale, while rate‑limiting requests so you don't get blacklisted, and handling the EAN‑based image‑gallery mapping and attribute variations that differ across product categories without corrupting the JSON schema. I've built scrapers for similarly protected e‑commerce sites (including a 200k‑product run for a price‑comparison engine) using rotating proxies, headless browsers with stealth plugins, and exponential backoff, while logging failed URLs for retry. I'll deliver a Python script that reads your product URLs, parses each page, and outputs the exact JSON structure—with a session‑based retry mechanism—all within 3 days. Do you need the script to use a rotating proxy pool, or will you provide a static set of residential IPs? Best regards, Artur
€30 EUR in 7 days
2.0
2.0

⭐ Hi, I can help you build an efficient product data scraper for Boulanger and export the results in the exact JSON structure you shared. From what I understand, you already have all product URLs, so the main work is not discovery, but reliable extraction at scale for around 125,000 products, including price, EAN, images, breadcrumbs, availability, specs, dimensions, weight and attributes. I have over 10 years of experience in IT, data extraction, automation and backend workflows, with a focus on clean and high-quality delivery. For this project, I would build the scraper with proper batching, retries, logging, checkpoint/resume support and JSONL/CSV export, so the process can continue safely if it stops midway. Since the site uses DataDome, I would first test a small batch and handle access carefully using allowed sessions, rate limits and any proxy setup you already have, without making the scraper too aggressive or unstable. Accuracy is important here, so I would also add validation for ref, EAN and missing fields. I can deliver the scraper and a tested sample output for $220 USD within 5 days. Full 125k execution can then run after the sample format is approved. I’d be happy to start by testing 20 to 50 product URLs first. Best regards, Hugo
€220 EUR in 5 days
2.2
2.2

Rosheim, Guadeloupe
Payment method verified
Member since Jul 23, 2015
€8-30 EUR
€30-250 EUR
€750-1500 EUR
€8-30 EUR
€8-30 EUR
$10-30 USD
$2-8 USD / hour
₹600-1500 INR
$50-80 AUD
$15-25 USD / hour
$15-25 USD / hour
₹12500-37500 INR
€30-250 EUR
₹12500-37500 INR
₹12500-37500 INR
$30-250 USD
₹1500-12500 INR
$10-30 USD
₹1500-12500 INR
$250-750 USD
$2-8 USD / hour
$250-750 USD
₹12500-37500 INR
$30-250 USD
₹12500-37500 INR