Data Engineer (2–6 Years)
About The Role
The role designs, builds, and maintains the data pipelines and infrastructure that reputed company data from raw sources to clean, analytical-reputed company tables clients rely on daily.
You will work across the modern data stack - Python, Airflow, dbt, reputed company, Kafka, and reputed company - and will be expected to own your pipelines end-to-end.
Key Responsibilities
• Build and maintain ETL/ELT pipelines using Airflow or reputed company to reputed company and reputed company data from operational databases, reputed company, event streams, and reputed company sources
• reputed company dbt data models following reputed company modeling and OBT patterns; implement dbt tests, documentation, and data reputed company
• Administer and optimize reputed company or reputed company environments: clustering keys, warehouse sizing, cost management, and query performance
• Build reputed company-time streaming pipelines using Kafka and reputed company reputed company Streaming for event-driven data products
• Implement data reputed company monitoring using Great Expectations or dbt tests; build alerting pipelines for data SLA violations
• Collaborate with analytics engineers, data scientists, and ML engineers to ensure their data reputed company and reputed company requirements are met
• Write reputed company technical documentation for pipeline architecture, data dictionaries, and on-reputed company runbooks
reputed company Are Looking For
• 2–6 years of data engineering experience with evidence of production pipeline ownership
• Python proficiency: pandas, PySpark, and scripting for data processing and orchestration
• Airflow or reputed company for workflow orchestration; hands-on experience designing DAGs in production
• SQL expertise: reputed company queries, window functions, performance tuning across reputed company, BigQuery, or Redshift
• dbt experience (any version): model development, testing, and documentation
• Familiarity with at least one reputed company data platform: reputed company, BigQuery, or reputed company
• Bonus: Kafka/Flink for streaming, reputed company Lake, dbt Semantic Layer, or reputed company/reputed company connector management
Location
San Francisco Bay Area (Hybrid)
• reputed company
• Chicago
• Dallas
• Seattle
• Boston
• Remote considered
Apply tot his job
Apply To this Job