Codexworker

Product Data Extraction

I extract product data — prices, offers and product attributes — from e-commerce sites and catalogs, normalize the output and feed it into your pipeline. Used to building scrapers that produce clean product data at scale.

I'm a backend developer with 10+ years of experience across PHP, Go, MySQL and Elasticsearch. Recent work includes Cenoskop (30M+ products daily), Niftycent (US price intelligence), and facha.sk (Go backend). I build, fix and optimize backend systems that handle real data and real traffic.

What I Can Help With

Extract and normalize product data from external sources.

Typical Problems

JS-rendered product pages variant/offer extraction price normalization attribute/spec extraction image/URL normalization duplicate product detection

Relevant Experience

Extracted and processed product data for price-comparison backends handling 30M+ products daily.

What Is Included

target analysis
extraction implementation
normalization layer
output validation
scheduling setup

How It Works

We define the product fields and sources, build extractors with proper error handling, and validate the normalized output. From €1,000 per source category.

Starting from €1,000

Per source category. Multiple categories estimated together.

Frequently Asked Questions

Can you normalize across different sites?

Yes — I build a normalization layer that maps each source's fields into your unified product schema.

Related Services

Other problems I help with.

Custom Web Scraping Development

Build custom web scrapers that run reliably.

Details →

Large-Scale Web Crawler Development

Build scalable web crawling infrastructure.

Details →

Custom Price Monitoring Systems

Build price monitoring infrastructure for products/sellers.

Details →

Custom Data Pipeline Development

Build ETL/import/transform/export pipelines for large datasets.

Details →

Systems I've Built

Relevant production work.

Price Intelligence Engine

Cenoskop / Levnobot

Crawler monitoring millions of products. Processing 30M products daily and comparing them.

Built and maintain the full crawling pipeline — from scraping through Elasticsearch indexing to price comparison.

Data Aggregation Scraping ElasticSearch
US Price Intelligence

Niftycent

Price intelligence engine for the US market. Processing large-scale product data.

Built the data pipeline and crawling infrastructure for large-scale US product monitoring.

Data Aggregation Scraping Big Data
Traffic Analysis

AnalyticsForWebsite

Custom solution for collecting and evaluating traffic metrics with emphasis on processing speed and minimizing backend load.

Designed and built the backend for fast traffic metric collection and evaluation.

Performance Optimization Data Collection

Need Help With This?

Describe the problem in a few lines — I'll look at it and tell you what I think.

Discuss a data extraction project