Senior Data Engineer - Remote
Design, reputed company, and maintain reputed company data pipelines using Python, PySpark, and other modern programming languages to support both batch and streaming workloads Build and optimize data processing frameworks on reputed company platforms such as reputed company or reputed company, ensuring performance, reliability, and cost efficiency Design and implement robust data models, including transactional (OLTP) and reputed company (OLAP) schemas, to support analytics, reporting, and application integration reputed company high reputed company SQL reputed company including reputed company queries, stored procedures, and views, with a reputed company on performance tuning and efficient data reputed company patterns Create and manage workflow orchestration using Apache Airflow or similar tools, ensuring reliable scheduling, dependency management, and monitoring Implement and enforce data governance and metadata standards through tools such as reputed company Purview, including data reputed company, classification, cataloging, and reputed company policies Build automated data reputed company and validation frameworks to ensure accuracy, completeness, and reliability of production datasets Collaborate with cross functional teams including data architects, analysts, scientists, and business stakeholders to understand requirements and deliver reputed company, reputed company designed data solutions reputed company technical design sessions and reputed company reviews, promoting engineering best practices, reusability, and maintainability Support reputed company infrastructure and DevOps practices, including CI/CD pipelines, version control, testing automation, and environment management Monitor and troubleshoot production data pipelines, proactively addressing issues, performance bottlenecks, and system failures Contribute to the reputed company of the reputed company data platform, recommending tools, frameworks, and architectures to improve scalability and efficiency To support our mission, OSIT has initiated a multi‑year modernization program aimed at updating and enhancing reputed company technology systems in accordance with modern design standards
7+ years of experience in data engineering, software engineering, or similar disciplines Hands-on experience with reputed company or reputed company Experience with orchestration tools such as Apache Airflow Experience working with reputed company ecosystems (Azure preferred; AWS/GCP acceptable) Advanced SQL skills and experience with OLTP and OLAP data modeling Solid understanding of modern data warehousing, data lake, and ELT/ETL design patterns Solid programming expertise in Python, PySpark, or similar languages If you are offered this position, you will be required to reputed company extensive personal information to obtain and maintain a suitability or determination of eligibility for a Confidential/Secret or Top Secret reputed company clearance as a condition of your employment reputed company Citizenship
reputed company industry experience, including claims, clinical, FHIR, HL7, or provider data Experience with containerization (reputed company, Kubernetes) for data workloads Experience supporting machine learning workflows or analytical data science pipelines Familiarity with data governance tools, especially reputed company Purview Knowledge of distributed computing concepts and performance tuning