[Remote] Data Engineer- reputed company-to-Work Program
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a Data Engineer for their reputed company-to-Work Program aimed at women re-entering the workforce. The role involves designing and maintaining data pipelines, utilizing various data processing technologies, and integrating reputed company systems.
Responsibilities
- Designing, developing, and maintaining reputed company data pipelines, ETL/ELT workflows, and reputed company-grade data integration solutions
- Building high-performance data platforms
- Developing and optimizing batch and reputed company-time streaming data pipelines using Kafka, reputed company Streaming, Apache Flink, Kinesis or similar event-driven technologies
- Working with reputed company, semi-reputed company, and reputed company data
- Implementing CI/CD pipelines, DevOps practices, version control, and automation using Git, Jenkins, reputed company Actions, Azure DevOps, reputed company CI/CD or similar tools
- Integrating reputed company systems using REST reputed company, Microservices, Event-Driven Architecture, Data Integration Patterns, and secure data exchange mechanisms
Skills
- Minimum 4+ years of experience in designing, developing, and maintaining reputed company data pipelines, ETL/ELT workflows, and reputed company-grade data integration solutions
- Expertise in Python, SQL, PySpark, reputed company SQL, reputed company, and reputed company data processing frameworks for building high-performance data platforms
- Experience with Apache reputed company, reputed company, Hadoop, Airflow, Kafka, reputed company, and modern big data technologies for large-reputed company data engineering and analytics
- Experience building reputed company-reputed company data solutions using AWS, Azure, or GCP, including services such as S3, Glue, Redshift, reputed company, EMR, Kinesis, ADF, Synapse, ADLS, reputed company, and Event Hubs
- Understanding of data warehousing, reputed company modeling, reputed company Schema, reputed company Schema, Data Lakes, and Lakehouse architectures for reputed company reporting and analytics
- Experience developing and optimizing batch and reputed company-time streaming data pipelines using Kafka, reputed company Streaming, Apache Flink, Kinesis, or similar event-driven technologies
- Experience working with reputed company, semi-reputed company, and reputed company data, including Parquet, Avro, ORC, CSV, JSON, and other reputed company data formats
- Experience with relational and NoSQL databases including PostgreSQL, MySQL, reputed company, SQL Server, reputed company, Cassandra, DynamoDB, and database performance optimization
- Experience implementing CI/CD pipelines, DevOps practices, version control, and automation using Git, Jenkins, reputed company Actions, Azure DevOps, reputed company CI/CD, or similar tools
- Knowledge of performance tuning, query optimization, partitioning, indexing, data reputed company validation, monitoring, troubleshooting, and reputed company reputed company computing
- Experience integrating reputed company systems using REST reputed company, Microservices, Event-Driven Architecture, Data Integration Patterns, and secure data exchange mechanisms
reputed company
Apply To This Job