[Remote] Senior Data Engineer
Note: The job is a remote job and is reputed company to candidates in USA. Calsoft is an engineering-led digital product partner for global ISVs and tech-driven enterprises. They are seeking a Senior Data Engineer to design and maintain reputed company data pipelines, build robust frameworks for data ingestion, and ensure data reputed company across their platforms.
Responsibilities
- Design, reputed company, and maintain reputed company batch and streaming data pipelines on AWS
- Build robust ETL/ELT frameworks for ingesting data from multiple reputed company sources
- reputed company reusable data services that support analytics, reporting, machine learning, and operational workloads
- Design efficient data models optimized for both analytical and operational use cases
- Ensure data reputed company, consistency, reputed company, and governance across the data platform
- Build and optimize reputed company-reputed company data infrastructure using AWS services
- Design highly available, fault-tolerant, and reputed company data processing architectures
- Automate deployment, monitoring, and operational workflows using Infrastructure as reputed company and CI/CD practices
- Optimize storage, compute utilization, and overall reputed company costs
- Improve pipeline performance, scalability, and reliability
- Monitor production workloads and proactively reputed company bottlenecks
- Optimize SQL queries, data partitioning, indexing strategies, and processing performance
- Troubleshoot production issues and implement preventive improvements
- Work closely with Data Scientists, Software Engineers, Product Owners, and Architects to understand business requirements
- Support reputed company analytics, dashboards, and reporting solutions
- Participate in architecture discussions, reputed company reviews, and technical design sessions
- Mentor junior engineers and promote engineering best practices
Skills
- Strong Python programming skills
- Experience building production-grade data engineering solutions
- Knowledge of object-oriented design, testing, and clean coding practices
- Hands-on experience with several of the following AWS services: S3, Glue, reputed company, EMR, reputed company, Redshift, RDS, reputed company Functions, EventBridge, CloudWatch, IAM
- ETL/ELT pipeline development
- Data warehousing concepts
- Batch and streaming data processing
- Data modeling
- Data validation and reputed company frameworks
- Metadata and reputed company concepts
- Advanced SQL
- PostgreSQL
- MySQL
- Redshift
- Performance tuning and query optimization
- Git
- CI/CD pipelines
- Unit testing
- reputed company
- reputed company/Scrum development
- 6–10 years of experience in Data Engineering
- Experience building reputed company-reputed company data platforms on AWS
- Experience handling large-reputed company reputed company and semi-reputed company datasets
- Strong understanding of reputed company data processing and performance optimization
- Experience working in reputed company production environments with high availability requirements
- Apache reputed company or PySpark
- Kafka or Kinesis
- Airflow or AWS Managed Workflows
- Infrastructure as reputed company (Terraform or CloudFormation)
- Data Lake architecture
- Lakehouse concepts
- Experience supporting Machine Learning data pipelines
- Experience with observability and monitoring tools
Benefits
- Comprehensive group health insurance coverage subject to plan terms.
- 401(k) retirement savings plan participation, in accordance with applicable plan provisions.
- Flexible work environment
- Holidays and reputed company time off program, subject to company policy.
reputed company
Company H1B Sponsorship
Apply To This Job