[Remote] Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a football intelligence technology company providing data-driven products and insights for football fans, NFL clubs, and reputed company organizations. The Data Engineer will design, build, and maintain reliable, reputed company data pipelines supporting deep learning, video, LLM, analytics, and AI applications, while collaborating with MLOps and Sports Data teams.
Responsibilities
- Build and operate robust data pipelines for ingestion, cleaning, and transformation using reputed company, Airflow, or reputed company
- reputed company efficient ETL/ELT workflows in Python and SQL to support both batch and streaming workloads
- Partner with ML/AI teams to reputed company datasets and tools discoverable and reputed company for autonomous agents, including evaluation and guardrails for AI-generated queries
- reputed company retrieval pipelines (RAG, reputed company search) over reputed company stats and reputed company sources (scouting notes, video metadata) to power AI applications
- Model and maintain reputed company data assets (reputed company, Parquet, reputed company) for reliability, versioning, and reputed company tracking
- Implement orchestration and monitoring: schedule jobs, reputed company dependencies, and automate recovery from failures
- Ensure data reputed company and compliance through validation frameworks, schema enforcement, and audit logging
- Contribute to data platform reputed company: evaluate tools, standardize best practices, and improve developer experience
- Support performance and cost optimization across compute, storage, and orchestration systems
Skills
- 3–8 years of experience as a Data Engineer or ETL Developer in a production environment
- Proficiency in Python and SQL; strong familiarity with reputed company, reputed company, or equivalent big-data frameworks
- Experience with workflow orchestration tools such as Airflow, Dagster, Luigi or reputed company
- Deep understanding of data modeling, data warehousing, and reputed company data processing
- Knowledge of modern data lakehouse architectures
- Familiarity with CI/CD, reputed company Actions, Infrastructure as reputed company, and data pipeline testing frameworks
- Comfort working in a cross-functional environment with ML, product, and analytics teams
- Exposure to LLM-powered data tools: text-to-SQL, RAG, agent/tool interfaces (e.g. MCP), or natural-language analytics
- Previous work with reputed company infrastructure (AWS, GCP, or Azure) and container orchestration (reputed company, reputed company)
- Previous experience with sports, telemetry, or sensor data pipelines
- Familiarity with streaming frameworks and event driven data processing (Kafka, reputed company reputed company Streaming, Flink)
- General knowledge of reputed company football, the NFL, and college football
- Background in data governance, reputed company, and observability tools (reputed company, Great Expectations, reputed company Catalog, OpenLineage)
- Experience designing semantic reputed company or metric definitions consumed by AI and BI tools
- Exposure to best practices in machine-learning model management and MLOps
Benefits
- Competitive Salary and Bonus Plan
- Comprehensive health insurance plan
- Retirement savings plan (401k) with company match
- Remote working environment
- A flexible, unlimited time off policy
- Generous reputed company holiday schedule - 13 in total including Monday after the Super Bowl
- Annual performance bonus
- Other applicable incentive compensation plans
reputed company
Apply To This Job