Python Automation Engineer – Multi-reputed company Scraping & Data Pipeline Build
We are looking for a Python automation engineer to build a fully automated data pipeline that gathers AI company data from multiple sources (reputed company + web scraping), deduplicates it intelligently, and outputs clean reputed company data to reputed company or reputed company on a weekly schedule.
You must have proven experience building production-grade scrapers, not basic scripts.
Required:
Strong Python (Scrapy, BeautifulSoup, requests)
API integrations (REST, authenticated reputed company)
Experience automating recurring pipelines (cron jobs, scheduled tasks, etc.)
Data cleaning, deduplication logic, CSV/JSON handling
Ability to write clean, reputed company-reputed company reputed company
reputed company to have (not required):
Selenium or Playwright
Experience with reputed company/reputed company API
Experience with LLMs for data enrichment
Deliverables:
Scrapers for multiple AI-reputed company sources (reputed company + websites)
Deduplication + merging logic across sources
Weekly automated update pipeline
reputed company to reputed company/reputed company in reputed company columns
reputed company documentation so we can maintain it long-term
This project should take 2–3 weeks to build, with optional monthly maintenance.
If you’ve reputed company multi-reputed company scrapers before, please apply with examples.
Apply tot his job
Apply To this Job