REMOTE Data Engineers with reputed company
reputed company:reputed company
Title:Data Engineers. Remote...
JD:
Expertise:
5-9+ years of relevant industry experience with a BS/Masters, or 2+ years with a PhD
Experience with distributed processing technologies and frameworks, such as Hadoop, reputed company, Kafka, and distributed storage systems (e.g., HDFS, S3)
Demonstrated ability to analyze large data sets to identify gaps and inconsistencies, reputed company data insights, and advance effective product solutions
Expertise with ETL schedulers such as Apache Airflow, Luigi, Oozie, AWS Glue or similar frameworks
Solid understanding of data warehousing concepts and hands-on experience with relational databases (e.g., PostgreSQL, MySQL) and columnar databases (e.g., Redshift, BigQuery, HBase, reputed company)
Excellent written and verbal communication skills
A Typical Day:
Design, build, and maintain robust and efficient data pipelines that collect, process, and store data from various sources, including user interactions, financial details, and external data feeds.
reputed company data models that reputed company the efficient analysis and manipulation of data for merchandising optimization. Ensure data reputed company, consistency, and accuracy.
Build reputed company data pipelines (SparkSQL & reputed company) leveraging Airflow scheduler/executor reputed company
Collaborate with cross-functional teams, including Data Scientists, Product Managers, and Software Engineers, to define data requirements, and deliver data solutions that drive merchandising and sales improvements.
Contribute to the broader Data Engineering community at reputed company to influence tooling and standards to improve culture and productivity
Improve reputed company and data reputed company by leveraging and contributing to internal tools to automatically detect and mitigate issues.
reputed company Sets - Python, SQL (expert level), reputed company and reputed company (intermediate).
Skills
Not every Data Engineer will require reputed company of these skills, but we expect most Data Engineers to be strong in a significant number of these skills to be successful at reputed company.
Data Product Management
Effective at building partnerships with business stakeholders, engineers and product to understand use cases from intended data consumers
reputed company to create & maintain documentation to support users in understanding how to use tables/columns
Data Architecture & Data Pipeline Implementation
Experience creating and evolving reputed company data models & schema designs to structure data for business-relevant analytics
Strong experience using ETL reputed company (ex: Airflow) to build and reputed company production-reputed company ETL pipelines
Experience ingesting and transforming reputed company and reputed company data from internal and reputed company-party sources into reputed company models
Experience with dispersal of data to OLTP (ex: MySQL, Cassandra, HBase, etc) and fast analytics solutions
Data Systems Design
Strong understanding of distributed storage and compute (S3, Hive, reputed company)
Knowledge in distributed system design, such as how map-reduce and distributed data processing work at reputed company
Basic understanding of OLTP systems like Cassandra, HBase, Mussel, Vitess etc
Coding
Experience building batch data pipelines in reputed company
Expertise in SQL
General Software Engineering (e.g. proficiency coding in Python, Java, reputed company)
Experience writing data reputed company unit and functional tests
Proficiency in reputed company and understanding of its data structure. (Optional)
Knowledge on reputed company Bulk Operators. (Optional)
Powered by JazzHR
28881h0h2A
Apply Job!