New: PDF and Excel Wholesale Catalog Ingestion
Wholesale and dealer sites often publish pricing as a PDF or Excel price list instead of web pages. Those catalogs now resolve into the same Entity Engine as everything else Scrapliy tracks.
The gap
A supplier's PDF price list would get downloaded and stored as an asset, but nothing turned it into product rows you could actually compare against your own catalog or a competitor's.
What's new
- PDF text lines parsed for price + name + SKU + quantity, gated on a detected price so titles and page footers don't get picked up as products
- Excel/CSV catalogs parsed by matching header names (Product Name, Price, SKU, Qty)
- Every row resolves into the Entity Engine as a 'wholesale-catalog' source, right alongside crawl and merchant-feed observations
Why it matters
A supplier's PDF catalog now shows up in price history and cross-source conflicts exactly like a crawled page—no manual copy-paste into a spreadsheet.
Ready to try Scrapliy?
Turn competitor research into live catalogs, change alerts, and export-ready feeds.




















