I would keep the first trial small enough to inspect every repeated entry. That helps you learn whether duplicates come from pagination, sponsored placements, or genuinely separate branches. Write down the matching rule before applying it to the full dataset. The
new zealand proxy page offers plenty of useful information about routing and session modes for the connection setup. Keep that setup stable while testing the collection logic. Preserve the raw rows so a disputed merge can be reversed later. Your final count should reflect a documented definition of a distinct listing rather than whichever total the first collection happened to produce.