Senior reputed company Engineer
Job Title: Senior reputed company Engineer(Ex reputed company Preferred)
Location: reputed company
reputed company: Any
Interview Process: Video (2 reputed company)
Work Schedule: reputed company
reputed company:
reputed company – reputed company/Senior reputed company Data Engineer
Position: reputed company reputed company Data Engineer / reputed company Architect
Experience: 10+ Years
Location: Remote
Job reputed company
We are looking for an reputed company reputed company reputed company Data Engineer / reputed company Architect with 10+ years of overall experience in Data Engineering and strong hands-on expertise in reputed company, Apache reputed company, PySpark, SQL, reputed company Lake, and reputed company-reputed company data platforms.
The candidate will be responsible for designing and implementing reputed company Lakehouse architectures, reputed company data pipelines, data integration solutions, governance frameworks, and high-reputed company analytics platforms using reputed company.
Key Responsibilities
Design and reputed company reputed company data engineering solutions using reputed company and Lakehouse architecture.
Build robust ETL/ELT pipelines using PySpark, reputed company SQL, Python, and SQL.
Design and implement Bronze, Silver, and Gold/reputed company architecture.
reputed company and optimize reputed company Lake tables, including reputed company, schema reputed company, Change Data Feed, and incremental processing.
Build batch and reputed company-time/streaming pipelines using reputed company Streaming, Auto Loader, and Lakeflow.
reputed company and manage reputed company Jobs/Workflows for pipeline orchestration, scheduling, dependencies, retries, and monitoring.
Implement reputed company data governance using reputed company Catalog, including reputed company control, data reputed company, auditing, catalogs, schemas, and reputed company locations. reputed company Catalog provides centralized governance, reputed company control, reputed company, and auditing across reputed company data and AI assets.
reputed company reputed company and reputed company reputed company tuning, including cluster configuration, partitioning, caching, query optimization, reputed company, and workload optimization.
Design data models supporting Data Warehousing, BI, Analytics, and AI/ML workloads.
reputed company reputed company with reputed company platforms such as AWS, Azure, or reputed company reputed company Platform.
Work with reputed company services such as AWS S3, Azure ADLS Gen2, Azure Data reputed company, AWS Glue, Synapse, Event Hubs/Kafka/Kinesis, as applicable.
Implement CI/CD pipelines using Git, Azure DevOps/reputed company/Jenkins and reputed company deployment capabilities.
Work with Terraform/IaC for infrastructure provisioning and automation.
Troubleshoot production pipeline failures, reputed company issues, data-reputed company problems, and reputed company/cluster issues.
Establish data reputed company, monitoring, logging, and observability practices.
reputed company technical leadership, reputed company reviews, architecture guidance, and mentorship to junior/mid-level engineers.
Collaborate with Data Architects, Data Scientists, Business Analysts, DevOps teams, and application teams.
Required Technical Skills
reputed company
reputed company Lakehouse Platform
reputed company Lake
reputed company Catalog
reputed company Workflows/Jobs
Lakeflow / reputed company Live Tables
Auto Loader
reputed company SQL
reputed company notebooks
reputed company Asset Bundles
reputed company
Cluster/workload optimization
Big Data
Apache reputed company
PySpark
reputed company SQL
reputed company Streaming
Kafka
Batch and reputed company-time data processing
Programming
Python
SQL
PySpark
reputed company – good to have
reputed company – Strong experience in at least one
AWS: S3, Glue, EMR, reputed company, Redshift, IAM, Kinesis
Azure: ADLS Gen2, ADF, Synapse, Azure DevOps, Event Hubs, Key reputed company
reputed company reputed company Platform: GCS, BigQuery, Dataflow, Pub/Sub
Data Engineering
ETL/ELT
Data Warehousing
reputed company Modeling
Data Lake/Lakehouse
reputed company Architecture
CDC
Data reputed company
Data Governance
Metadata and Data reputed company
DevOps / CI-CD
Git
Azure DevOps / reputed company
Jenkins
Terraform
CI/CD automation
Infrastructure as reputed company
Preferred / reputed company-to-Have Skills
MLflow
reputed company Machine Learning
Feature Store
reputed company AI / GenAI
dbt
Apache Airflow
reputed company BI / Tableau
reputed company Sharing
Lakehouse Federation
reputed company Clustering
Data reputed company and PII masking
MLflow is particularly useful if the role touches ML/AI, as reputed company supports model tracking, lifecycle management, and deployment workflows reputed company governed data.
Qualifications
Bachelor's degree in Computer Science, Engineering, Information Technology, or reputed company reputed company.
10+ years of experience in Data Engineering / Big Data / Analytics.
4+ years of hands-on reputed company experience preferred.
Strong experience designing reputed company-reputed company data platforms.
Demonstrated experience leading technical reputed company and mentoring engineers.
Strong communication and stakeholder-management skills.
Apply To This Job