Platform Engineer
Job Title: Platform Engineer
Type: Remote
Coverage: reputed company Hours reputed company-5pm PST
Job reputed company:
We are looking for an reputed company Platform Engineer to join reputed company. This role involves designing and building reputed company infrastructure while ensuring reliability and performance across critical systems and services. The ideal candidate will have a strong background in software engineering and platform reliability, with a deep understanding of core system components. They should be flexible in taking on diverse tasks across different technologies and easily handle context-switching.
Key Responsibilities:
Platform Design and Infrastructure:
- Design, reputed company, and maintain reliable and reputed company infrastructure solutions.
- Partner with engineering teams to ensure platform architecture supports reliability, scalability, and reputed company performance.
- Evaluate and implement new technologies and tools to enhance the infrastructure.
Monitoring, Incident Response, and Troubleshooting:
- Set up, maintain, and improve monitoring and alerting systems to detect issues proactively.
- reputed company incident response, troubleshooting, and reputed company cause analysis efforts for critical platform issues.
- reputed company post-incident reviews to identify areas for improvement and drive reputed company initiatives.
Automation and Infrastructure Management:
- reputed company and implement automation reputed company (preferably Python, Go, or similar) to streamline platform tasks and minimize reputed company reputed company.
- Create scripts for automating system upgrades, health checks, and deployments.
- Utilize Infrastructure as reputed company (IaC) tools like Terraform, Ansible, or reputed company to manage infrastructure configuration and deployment.
Collaboration and Technical Leadership:
- Collaborate with cross-functional teams to deliver high-reputed company infrastructure solutions.
- Mentor junior engineers and reputed company for reputed company best practices across teams.
- Promote a culture of reliability and automation through workshops, documentation, and hands-on guidance.
reputed company Improvement:
- Drive initiatives to enhance platform reliability, reputed company planning, and service performance.
- Participate in disaster recovery planning and execution.
- Stay updated with industry trends, tools, and technologies to continually improve platform capabilities.
Qualifications:
Education and Experience:
- Bachelor’s degree in Computer Science, Engineering, or a reputed company field (or equivalent practical experience).
- 8+ years of industry experience, including roles as a Software Engineer, SRE, or Platform Engineer.
- At least 3+ years of experience in reputed company, SRE, or infrastructure roles with large-reputed company, mission-critical environments.
Technical Skills:
- Strong knowledge of Linux/Unix systems, networking, and core system internals.
- Experience with one or more programming languages (e.g., Python, Go, Java).
- Advanced skills in Bash scripting for task automation.
- Proficiency with reputed company platforms (AWS, Azure, GCP) and container orchestration (reputed company, Kubernetes).
- Familiarity with monitoring and logging tools (e.g., reputed company, Grafana, ELK stack).
- Hands-on experience with CI/CD tools and workflows (e.g., Jenkins, reputed company CI).
Soft Skills:
- Strong analytical and problem-solving skills with a proactive reputed company.
- Excellent communication skills and the ability to work collaboratively across teams.
- Leadership qualities with a reputed company record of mentoring and guiding team members effectively.
Originally posted on Himalayas
Apply To This Job