Resources · Blog

Blog about data collection

Guides, research, and case studies on web scraping, price monitoring, and data delivery for analytics, marketing, and product needs.

Guides & Basics
What Is Selenium WebDriver? The Selenium Ecosystem Explained

One name covers six things moving at six different speeds. Current versions, W3C status, Grid sizing, cloud grid list prices and detection signals, all read from primary sources on 13 August 2026.

31 May · 17 min
Guides & Basics
What Is Web Scraping? The Complete Guide

What web scraping is, how a scraper works step by step, and what it costs: proxy, rendering and CAPTCHA prices read from vendor pages on 13 August 2026, library versions current to that date, and where the law actually stands after hiQ v. LinkedIn.

29 May · 22 min
Tools & Reviews
Wireshark as an HTTP Sniffer: Review and Setup

Wireshark 4.6.8 as an HTTP sniffer: capture filters that do not drop QUIC, key-log HTTPS decryption, JA3/JA4 fingerprint comparison, and priced alternatives, checked August 2026.

27 May · 18 min
Tools & Reviews
WordPress Scraper: Plugins and Custom Content Parsers

Six WordPress scraper plugins checked against WordPress 7.0.4 and WooCommerce 11.0.1 on 13 August 2026, and the custom PHP that replaces them when they run out.

25 May · 31 min
Data & Formats
XML Parsing: Libraries, XPath, and Examples in Five Languages

XML parsing in Python, Node, PHP, Java and Go, rechecked August 2026: namespace traps that return nothing, DOM versus streaming memory measured on a 32 MB feed, and which parsers still open external entities by default.

23 May · 14 min
Techniques
XPath for Web Scraping: A Practical Guide

XPath for web scraping, rechecked in August 2026: the seven node types, all thirteen axes, the XPath 1.0 comparison rules that silently drop rows, and why a path copied from Chrome returns nothing in lxml.

21 May · 32 min