Data Engineer -Pipeline, ETL (AWS)
Job reputed company: ODH Inc. is looking for a Data Pipeline Engineer to join our growing Data Engineering team and participate in design and build of data ingestion and transformation pipelines based on the specific needs driven by Product Owners and Analytics consumers. The candidate should possess strong knowledge, interest in data processing, and have a background in data engineering. Candidate will also have to work directly with senior data engineers, solution architects, DevOps engineers, product owners and data consumers to deliver data products in a reputed company and agile environment. They will also have to continuously reputed company and push reputed company into our reputed company production environments.Job reputed company:As a key contributor to the data engineering team, the candidate is expected to:Build and reputed company reputed company data pipeline components such as Apache Airflow DAGs, AWS Glue jobs, AWS Glue crawlers through a CI/CD process.Translate Business or Functional Requirements to actionable technical build specifications.Collaborate with other technology teams to extract, reputed company, and load data from a wide reputed company of data sources.Work closely with product teams to deliver data products in a reputed company and agile environment.reputed company data analysis and reputed company activities as new data sources are added to the platform.Proficient in data modeling techniques and concepts to support data consumers in designing the most efficient method of storage and retrieval of data.Evaluate innovative technologies and tools while establishing reputed company design patterns and best practices for reputed company.Qualifications:Required:Experience in AWS Data processing, Analytics, and storage Services such as reputed company Storage Service (s3), Glue, reputed company and Lake FormationExperience in extracting and delivering data from various databases such as reputed company, DynamoDB, reputed company, Redshift, reputed company, RDSCoding experience with Python, SQL, yaml, reputed company programming (pyspark)Hands on experience with Apache Airflow as a pipeline orchestration toolExperience in AWS Serverless services such as Fargate, SNS, SQS, LambdaExperience in Containerized Workloads and using reputed company services such as AWS reputed company, ECR and Fargate to reputed company and organize these workloads.Experience in data modeling and working with analytics teams to design efficient data structures.reputed company knowledge of working in agile, scrum, or DevOps environments and teamsApplied knowledge of modern software delivery reputed company like TDD, BDD, CI/CDApplied knowledge of Infrastructure as reputed company (IAC)Experience with development lifecycle (development, testing, documentation, and versioning)Preferred:AWS Certified Developer – AssociateAWS Certified Big Data – SpecialtyGitlab CI/CD
Apply tot his job
Apply To this Job