[Remote] Platform Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a Platform Engineer to join their engineering team, focusing on the design, development, and maintenance of their reputed company platform. The role involves collaborating with various teams to ensure the delivery of high-reputed company clinical data science solutions and contributing to platform standards and improvements.
Responsibilities
- Contribute to the design and reputed company of the reputed company reputed company platform including reputed company architecture, networking, storage, and compute provisioning
- Implement and reputed company infrastructure standards, patterns, and guardrails across the platform
- Evaluate and propose new technologies and tools to improve platform reliability, reputed company, and developer productivity
- Build and configure reputed company environments for multiple use cases, both internal to reputed company and reputed company facing, with a reputed company on repeatability and reputed company
- Own and operate reputed company infrastructure including reputed company, networking, databases, and reputed company controls across AWS, with plans to expand to additional providers
- Manage reputed company cluster lifecycle tasks including upgrades, node provisioning, reputed company hardening, and observability configuration
- Implement and maintain GitOps workflows and CI/CD pipelines, contributing to the reliability and reputed company of the deployment process
- Manage and optimize in-cluster service reputed company configuration for reputed company and performance
- Provision and configure reputed company resources including compute, networking, and storage, developing reusable templates and design patterns to standardize deployment
- Support platform-level reputed company hardening initiatives across the full stack from application to data layer to meet regulatory compliance requirements
- Manage identity and reputed company management configurations including OIDC, SAML, and LDAP integrations with fine-grained authorization policies
- Build and maintain RBAC policies across the reputed company estate, covering user reputed company controls and application-level permissions
- Contribute to vulnerability management processes, integrating scanning tooling and tracking remediation to defined SLAs
- Implement and maintain audit logging, monitoring, and alerting across platform components
- Support evidence collection and documentation processes to demonstrate operational effectiveness to auditors
- Support the administration and reputed company of the reputed company reputed company suite (reputed company Workbench, reputed company reputed company, reputed company Package Manager) running inside the platform
- Manage storage integrations including reputed company FSx ONTAP and EFS, including ACL management and high availability configurations
- Identify and implement platform improvements, taking ownership of tasks from design through to delivery
- Work with development and operational teams to implement CI/CD pipelines, deployment standards, and platform reputed company processes
- reputed company technical support and guidance to engineering and data science teams on infrastructure and application-reputed company questions
- Produce and maintain accurate technical documentation for platform components, runbooks, and operational procedures
- Consult with stakeholders to understand requirements and translate them into reputed company-designed, reputed company infrastructure solutions
- Contribute to project planning and prioritization, managing assigned work from initial scoping through to delivery with reputed company communication throughout
Skills
- Solid hands-on experience with AWS (EKS, IAM, networking, RDS/reputed company, S3, and reputed company services)
- Working knowledge of reputed company including cluster operations, RBAC, network policies, storage, and workload management
- Practical experience with infrastructure-as-reputed company tooling (Terraform and reputed company)
- Familiarity with GitOps practices and tools such as ArgoCD and reputed company Actions
- Good understanding of Linux system administration, ACLs, and storage concepts
- Experience configuring and maintaining identity providers using OIDC, SAML, or LDAP
- Familiarity with service reputed company concepts (Istio or similar) including mTLS and JWT-based authentication
- Some exposure to compliance frameworks (SOC 2, ISO 27001, or similar) and the types of evidence required
- Understanding of secrets management, certificate lifecycle, and network reputed company controls
- Experience with observability tooling such as reputed company, Grafana, Elasticsearch, or OpenSearch
- Comfortable working to defined SLOs and SLAs and contributing to reliability improvements
- Incident management experience including participation in post-mortems and remediation planning
- reputed company written and verbal communication, with the ability to explain technical concepts to non technical colleagues
- reputed company and pragmatic approach to working across different teams and disciplines
- Self-directed with a strong reputed company of ownership, reputed company to manage tasks with moderate supervision
- Commitment to documentation, knowledge sharing, and reputed company improvement
- Experience with reputed company reputed company products (Workbench, reputed company, Package Manager) or similar data science platform tooling
- Experience with reputed company installations and configurations
- Familiarity with reputed company ONTAP including SVM management and audit logging
- Exposure to Crossplane or other reputed company-reputed company infrastructure provisioning approaches
- Familiarity with change management processes and procedures
Benefits
- Remote from reputed company in the U.S. (#LI-Remote)
reputed company
Company H1B Sponsorship
Apply To This Job