Web Scraping & Data Extraction Specialist: Healthcare Directory (Excel / Google Sheets)

via Freelancer ·

Budget / Salary₹1,500–12,500
TypeFreelance project
LocationRemote
Posted1 hour ago
Project Overview:
​We are looking for a reliable, detail-oriented freelancer to extract and compile structured healthcare-related business data from 2–3 specified directory websites into an organized Excel / Google Sheets workbook.
​The directory covers healthcare entities categorized by location (City, State) and category, including images/photographs where available.
​Key Responsibilities:
​Target Categories:
​Hospitals & Clinics
​Doctors / Specialists
​Medical Stores / Pharmacies
​Diagnostic Centers & Pathology Labs
​Required Data Fields (Columns):
​Location: State, City, Area/Locality, Full Address, Pincode
​Entity Info: Business/Entity Name, Category, Sub-Category/Specialization (e.g., Cardiologist, Dental Clinic, 24x7 Pharmacy)
​Contact Details: Phone Number(s), Email Address, Official Website URL
​Operating Details: Working Hours/Timings, Emergency Services Available (Yes/No, if applicable)
​Media/Photographs: Direct image URLs of the facility/storefront/doctor profile (or locally downloaded and organized in a cloud folder matching the Record ID)
​Source: Original profile/listing URL
​Data Quality & Hygiene:
​Ensure 100% duplicate-free records.
​Standardize phone number formats, text casing (Title Case), and addresses.
​Validate that all image URLs are active and accessible.
​Deliverables:
​Clean, well-structured Microsoft Excel (.xlsx) and/or Google Sheets file.
​Separate tabs/sheets organized by State/City or filtered by Category as agreed.
​Google Drive folder containing images named systematically (if image downloads are requested).
​Required Skills & Qualifications:
​Proven experience in Web Scraping, Data Mining, or Data Entry.
​Proficiency with tools like Python (BeautifulSoup, Scrapy, Selenium), Octoparse, ParseHub, or manual curation tools.
​Advanced Excel/Google Sheets skills (data formatting, data validation, deduplication).
​Strong attention to detail and zero tolerance for missing/corrupted fields.
​To Apply, Please Provide:
​Brief summary of your experience with similar data extraction/directory projects.
​Estimated turnaround time for a sample batch (e.g., 5000–1,0000 records).
​Preferred tools/methods (automated scraper vs. manual extraction).
​Any relevant portfolio or sample sheets from past web scraping jobs.
python data entry excel web scraping web search data mining scrapy data extraction beautifulsoup google sheets
Apply on Freelancer →

Project sourced from Freelancer.com. Applications happen directly on the original platform — we never collect your data.