300K Polish LLC Email Database
Budget / Salary$10–30
TypeFreelance project
LocationRemote
Posted1 hour ago
I need a single, well-structured file containing at least 300,000 limited-liability companies registered in Poland, each paired with a working email address.
Company names must be taken directly from the official Statistics Poland (GUS) and CEIDG registers, then matched with a valid email that you find through Google search or other open-web sources.
To keep the file easy to use, I only want two columns:
1. Company Name (exactly as it appears in GUS or CEIDG)
2. Email Address (one address per company, no catch-all “info@” unless it is the only address publicly available)
Deliver the finished dataset in whichever tabular format you prefer—CSV, XLSX or Google Sheets are all fine—as long as it opens cleanly and every row is consistently formatted.
Quality is more important than speed:
• The count must reach or exceed 300,000 complete rows.
• Obvious duplicates, dead domains, or syntactically invalid emails should be removed before delivery.
• Please maintain the original Polish diacritics in company names.
When you respond, briefly outline:
• Your approach for extracting names from GUS/CEIDG.
• The tools or scripts you’ll use to harvest and validate email addresses.
• An estimated turnaround time for the full dataset.
A small sample of 200–300 rows will be required for review before I release the first milestone, and I will verify a random subset of emails before final acceptance.
Budget 15 USD.
Company names must be taken directly from the official Statistics Poland (GUS) and CEIDG registers, then matched with a valid email that you find through Google search or other open-web sources.
To keep the file easy to use, I only want two columns:
1. Company Name (exactly as it appears in GUS or CEIDG)
2. Email Address (one address per company, no catch-all “info@” unless it is the only address publicly available)
Deliver the finished dataset in whichever tabular format you prefer—CSV, XLSX or Google Sheets are all fine—as long as it opens cleanly and every row is consistently formatted.
Quality is more important than speed:
• The count must reach or exceed 300,000 complete rows.
• Obvious duplicates, dead domains, or syntactically invalid emails should be removed before delivery.
• Please maintain the original Polish diacritics in company names.
When you respond, briefly outline:
• Your approach for extracting names from GUS/CEIDG.
• The tools or scripts you’ll use to harvest and validate email addresses.
• An estimated turnaround time for the full dataset.
A small sample of 200–300 rows will be required for review before I release the first milestone, and I will verify a random subset of emails before final acceptance.
Budget 15 USD.
Apply on Freelancer →
Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.