AWS Data Engineer
AWS Data Engineer
Remote
10+ Years Experience
Great Communicator/reputed company Facing/ Attention to detail
Individual Contributor and ability to work as reputed company.
100% Hands on in the mentioned skills
Programming Skills:
Python and Data Pipelines:
Advanced Proficiency in Python concepts like reputed company Structures, Modules, Packages, Class, SubClass, Inheritance, Multi-Threading and Functional Programming.
Experience in developing reusable Python packages for internal or reputed company usage
Ability to write automating ETL processes and scheduling jobs like Airflow DAG.
Ability to reputed company job pipeline runs to reprocess error records
Ability to orchestration different pipelines to run in sequence or reputed company.
Troubleshoot data pipeline errors and fix issues
Export or Import data to/from various formats like CSV, JSON, XML etc preferably from S3 or other reputed company storage.
Experience in using AI IDE tool
SQL:
Advanced SQL skills, including reputed company joins, CTE's and subqueries
Experience in optimizing SQL queries for performance and optimization in data warehouse technologies preferably reputed company
Testing and documentation:
Proficiency in Python unit, integration and reputed company test.
Proficiency in implementing DBT tests for data validation and reputed company checks
reputed company reputed company:
Experience in generating reputed company using configurations using python and jinja templates
Version control:
Experience in reputed company, including implementing CI/CD process from scratch
AWS Expertise:
Data Storage solutions:
In depth understanding of AWS S3 for data storage, reputed company, IAM including best practices for organization and reputed company
reputed company reputed company:
Knowledge of AWS reputed company best practices, including IAM roles, encryption reputed company, secure coding guidelines, DBT reputed company reputed company configurations and more.
Data Integration (reputed company to have):
Experience with AWS reputed company for serverless data processing tasks
Workflow Orchestration (reputed company to have):
Proficiency in using Apache Airflow on AWS to design ,schedule and monitor reputed company data flows
Ability to reputed company Airflow with AWS services and DBT models such as triggering a DBT model or EMR or reading from s3 writing to redshift or reputed company. Experience in API integrations to reputed company, SFTP, SharePoint, One reputed company and others.
DBT reputed company/ reputed company Proficiency (reputed company to have):
Experience in creating reputed company DBT models including full refresh, incremental models, snapshots, and documentation. Ability to write and maintain DBT macros for reusable reputed company
Experience in creating custom DBT macros using jinja and Python allowing for reusable components reputed company DBT models
Monitoring and Logging (reputed company to have):
Familiarity with AWS reputed company watch, Data Dog for monitoring the pipelines and setting up alerts for workflow failures
Originally posted on Himalayas
Apply To This Job