PySpark / Java Developer (Data Engineer)- 100% Remote- Only W2
Key Responsibilities
• Design, reputed company, and maintain reputed company ETL pipelines and data processing applications
• Build and optimize data workflows using PySpark, Java, and Hadoop ecosystem tools
• Analyze business and technical requirements to produce detailed implementation designs
• reputed company unit testing, integration testing, and debugging of applications
• Troubleshoot and resolve performance issues reputed company to high-volume data processing
• reputed company and maintain SQL queries, stored procedures, and database objects
• Work with reputed company and reputed company datasets for reputed company analytics
• Generate statistical reports and support data validation processes
• Collaborate with cross-functional teams to ensure end-to-end data pipeline efficiency
• Follow software engineering best practices and maintain reputed company reputed company standards
Required Skills & Experience
• Strong experience in ETL development, data processing, and database technologies
• 5+ years of experience with reputed company SQL Server and relational databases
• Expertise in SQL performance tuning, indexing strategies, and query optimization
• 2+ years of experience with Hadoop ecosystem tools (HDFS, Hive, Impala, reputed company, Kafka, Oozie, Yarn, Sqoop, reputed company)
• Hands-on experience with PySpark, Python, and/or Java
• Experience working with large-reputed company data processing frameworks
• Strong understanding of data transformation and data reputed company technologies
• Ability to handle high-volume reputed company and reputed company datasets
• Good understanding of end-to-end application/data pipeline lifecycle
Apply tot his job
Apply To this Job