Koca Ventures Ltd
71-75 Shelton Street
Covent Garden, London
WC2H 9JQ, United Kingdom
Registered in England & Wales — 16231043
A price-intelligence pipeline you own —your sources, your catalogue, your systems.
Not a per-SKU SaaS seat you rent. We build a private pipeline that scrapes the public sources you care about, matches them to your catalogue with confidence scores, and wires the result into your own systems — public-data-only by design.
Our legal & ethical stance is the feature
Most price-monitoring marketing glosses over how the data is collected. We put it first — here is exactly what we will and won't do.
✓We collect public, logged-out prices and availability.
We don't scrape behind logins or paywalls.
✓We respect robots.txt and rate-limit politely.
We don't defeat anti-bot, CAPTCHA, or access-control systems.
✓No fake accounts, no credential use.
We don't impersonate users or platforms.
✓We minimise and avoid personal data; GDPR/UK DPA-aware design.
This isn't legal advice — you get counsel for your jurisdiction and use.
✓Auditable collection and repricing decision logs.
You own your pricing strategy and competition-law compliance (MAP/RPM, algorithmic-pricing antitrust).
✓A pipeline and dataset you own — on your infrastructure if you want.
Coverage of any specific site is best-effort; hostile anti-bot targets may be declined on principle.
The bright line is public, logged-out data versus everything else: staying logged out, respecting robots.txt, never circumventing access controls. This is not legal advice — for any sizeable program we recommend jurisdiction-specific counsel.
Different buyers want different things
Retailer repricing
Daily-to-near-real-time competitor price and availability tracking, fed into repricing logic with hard margin floors. The goal is never the lowest price — it's the highest price that still wins, with a guardrail at your true landed cost.
Brand MAP & channel-price monitoring
For brands and distributors: spot retailers advertising below your agreed floor and watch channel pricing across markets to protect brand equity. We monitor and report — the commercial decision stays yours.
Marketplace Buy-Box & assortment-gap
Stock and promotion visibility plus assortment-gap analysis — what competitors stock that you don't — and the signals that move Buy-Box wins. Structured so a category team can act on it, not just admire a chart.
Data-team structured feeds
If you have your own pricing or data team, you want the pipeline, not a dashboard: clean schemas, stable structured feeds, and APIs that drop straight into your ERP, PIM, BI, or your own models.
The five-layer pipeline
Collection
Headless Playwright drivers read public, logged-out pages — politely, with driver-level stealth that never defeats login walls or anti-bot challenges. A site that actively blocks automated reading is a “do not collect” signal we respect.
Network & proxies
Residential and ISP proxies geo-targeted to the markets that matter, sized to volume and target difficulty. A real pass-through cost that scales with what you ask for — we're transparent about it.
Anti-bot-aware reliability
Where a do-it-yourself setup hits a reliability ceiling on permissible public targets, we use managed collection APIs — for uptime on pages we're allowed to read, never to beat an access-control system.
Product matching — confidence + human QA
The make-or-break layer. With a correct GTIN/EAN/UPC barcode, matching is a clean join; the engineering value is the rest — fuzzy text, image, and LLM semantic matching, each producing a per-pair confidence score, with low-confidence pairs escalated to human QA.
Intelligence, repricing & dashboards
Change detection is not intelligence. We turn “the number moved” into “what it means” — vs your price, who leads the category, promo vs error — and feed hybrid repricing: transparent rules at the boundaries, ML inside them, every decision logged.
Be honest about the cat-and-mouse: scrapers are a living system — targets change layouts and defences constantly. Part of the work is perpetual maintenance, which is why the operate phase is a retainer, not a one-time deliverable.
A real multi-source scraper, retargeted at retail
We already run a production multi-source market-intelligence scraper for a client in another vertical — six public property portals across the EU, with market rollups and an anti-hallucination guardrail that blocks overclaiming before it leaves the system. The same engine, retargeted at retail prices, is exactly what this service is.
Straight answers
Is scraping competitor prices legal in the UK?
Collected the way we do it — public, logged-out pages, robots.txt respected, polite rate limits, no fake accounts, no defeated access controls, personal data minimised — it's the defensible posture. It isn't legal advice: for a sizeable program, get jurisdiction-specific counsel.
Can you guarantee you'll scrape site X?
No honest shop can. Coverage is best-effort — targets change layouts and defences constantly. A hostile anti-bot wall means the operator doesn't want automated collection; we decline rather than try to beat it, and we tell you up front what we can reliably read.
What accuracy do you promise?
Per-pair confidence scores and a human-QA loop on low-confidence matches — not a universal accuracy percentage. Headline figures are catalogue-conditional: a number measured on a vendor's reference catalogue doesn't carry over to yours. We'd rather show you a calibrated score than sell a number we can't stand behind.
Do we own it?
Yes. The codebase, the schemas, and the matched dataset are your asset — on your own infrastructure if you want. No per-SKU seat tax, no data lock-in.
Do you set our prices, or handle MAP enforcement?
No. We build the engine, the matching, and the audit trail; pricing strategy stays yours. In the US, “MAP enforcement” is a recognised concept; in the EU/UK, resale-price maintenance is restricted, so we frame brand-side work as monitoring and reporting — the pricing rules stay your commercial and legal responsibility.
How do you price it, and can it run on our infrastructure?
Per engagement — there's no list price. The typical shape: a fixed-fee pilot on your own catalogue, a project build, then a retainer to keep the scrapers healthy. Yes, it can run on your infrastructure — and proxy costs are a real pass-through we're transparent about.
How is this different from Prisync or Price2Spy?
Those are good SaaS products — fast to start if your competitors are already in their coverage. We build the other thing: an owned pipeline on your exact sources, matched to your catalogue, wired into your systems. You keep the code and the data.
Should we build or buy price monitoring?
Buy if a SaaS tool already covers your competitors and the per-SKU cost stays sane — we'll say so. Build when you need bespoke sources, native ERP/PIM integration, ownership of the matched dataset, or insulation from a vendor's roadmap. A pilot on your catalogue settles it.
Last reviewed:
Scope a pilot on your own catalogue
The honest first step isn't a demo — it's a pilot on your own SKUs, because live data is where the complexity hides. Tell us the sources and the catalogue, and we'll scope coverage, confidence, and a build.
