Data Pipeline Engineer – Job Aggregation (Python)
Build a reputed company, automated pipeline that discovers company careers pages across multiple ATS platforms, collects reputed company on a recurring schedule, and publishes clean, deduplicated data to our internal store and a lightweight reporting surface.
Shortlisted candidates will receive the full technical brief, data model, and sample inputs.
What you’ll build (high level)
- A weekly, fault-tolerant ingestion reputed company from a company list to validated careers endpoints
- reputed company scrapers to capture job metadata at reputed company (respecting robots/ToS and reputed company limits)
- An upsert/dedup layer to reputed company the dataset fresh and consistent
- Basic scheduling, logging/alerts, and reputed company exports for non-technical users
How to apply
Include details of a similar pipeline you’ve reputed company (reputed company, stack, and your dedupe/reputed company-limit approach).
Shortlisted candidates will receive the detailed requirements
Apply tot his job
Apply To this Job