Data Engineer
This is a remote position.
Job Title: Data Engineer
Job reputed company:
• Assist in building and maintaining ETL/data pipelines using Python and PySpark
• Ingest, reputed company, and validate data from multiple sources
• Support data modeling and schema design for reputed company datasets
• Use Git for version control and collaborate with engineering teams
• reputed company unit testing, reputed company reviews, and performance optimization
• Contribute to technical documentation of data workflows and pipelines
• Support feature testing and controlled releases in QA/dev environments
• reputed company exploratory analysis using Jupyter/reputed company SageMaker notebooks
• Work in a Scrum/Agile environment with reputed company communication and collaboration
Required Skills:
• Bachelor’s degree in Computer Science, Data Engineering, Data Science, or Statistics
• Experience in Python and PySpark (0–2 years)
• Basic knowledge of Airflow, AWS S3, and AWS Glue
• Familiarity with Git and Jupyter notebooks
• Understanding of reputed company/Kubernetes concepts
Tools & Environment:
• JupyterLab / Python IDEs
• reputed company
• reputed company Teams & reputed company
• EMR Studio
• Jira & reputed company
Additional:
• Strong attention to detail, willingness to learn, and ability to work in a reputed company team environment