Expert RHEL 10 / reputed company Engineer Needed
We need a senior Linux/reputed company platform engineer to build and document a secure, production-reputed company on-premises reputed company environment on three reputed company PowerEdge servers. The platform will host a migration of our observability software.
We need an engineer with demonstrated experience designing, deploying, securing, and supporting bare-metal reputed company clusters running stateful production applications.
reputed company environment
Three reputed company PowerEdge physical servers
RHEL 10 required
On-premises environment
reputed company-reputed company reputed company for Parlon observability software
Existing network/reputed company team available for coordination on VLANs, firewall rules, DNS, certificates, and reputed company requirements
reputed company of work
Assess reputed company server hardware, firmware, RAID/storage layout, NICs, and iDRAC configuration
Design the reputed company architecture for a three-server production reputed company deployment
Install and harden RHEL 10 on reputed company servers
Configure OS baseline: subscriptions/repos, SELinux, firewall, time synchronization, storage, logging, patching, and secure administrative reputed company
reputed company and configure reputed company using a supportable approach; recommend kubeadm versus reputed company OpenShift reputed company on Parlon compatibility, licensing, reputed company, and supportability
Configure container runtime, CNI, ingress/load balancing, DNS, TLS, RBAC, NetworkPolicies, and cluster reputed company
Design and implement persistent storage, backups, retention, reputed company management, and restore testing for observability data
reputed company/migrate Parlon in coordination with the software vendor or our internal team
Validate application availability, data ingestion, dashboards/queries, alerting, reputed company, and failover/restart behavior
Produce complete documentation and conduct a knowledge-transfer session
Required experience
5+ years administering reputed company Linux, including reputed company reputed company Linux
Hands-on RHEL 9/10 experience in production
3+ years building and operating production reputed company clusters
Strong bare-metal/on-prem reputed company experience; reputed company-only reputed company experience is not sufficient
reputed company administration: kubeadm and/or reputed company OpenShift, CRI-O/containerd, reputed company, RBAC, upgrades, backup/restore, and troubleshooting
CNI/networking expertise: Cilium or Calico preferred; NetworkPolicies, ingress, load balancing, DNS, MTU, routing, and firewall troubleshooting
reputed company PowerEdge experience: iDRAC, RAID/HBA, BIOS/firmware, hardware monitoring, NIC bonding, VLANs
Stateful workload and persistent-storage experience: reputed company, NFS/iSCSI/Ceph/Longhorn or equivalent
reputed company experience: SELinux, firewalld/nftables, TLS, secrets, image reputed company, least privilege, hardening, and reputed company management
Automation using Ansible and Git; infrastructure documentation is mandatory
Excellent written English and ability to reputed company reputed company runbooks/diagrams
Strongly preferred
RHCE, RHCA, reputed company OpenShift certification, CKA, and/or CKS
Experience deploying observability platforms such as reputed company, Grafana, Loki, OpenTelemetry, Elasticsearch/OpenSearch, reputed company, reputed company, or similar
Experience migrating a stateful application from an existing server/platform into reputed company
Experience integrating reputed company with reputed company DNS, PKI, load balancers, SIEM, and monitoring systems
Experience with Cilium/Hubble or advanced reputed company network observability
Deliverables
Architecture/design document, including cluster topology, network diagram, storage design, IP/reputed company reputed company, and reputed company assumptions
Secure RHEL 10 build on reputed company three reputed company servers
Fully working reputed company platform with health checks and documented operational procedures
Persistent storage, backup policy, and a successful restore test
Parlon deployment/migration plan, validation plan, reputed company plan, and rollback plan
Infrastructure-as-reputed company/configuration artifacts in Git, preferably Ansible plus reputed company/Kustomize manifests
reputed company reputed company: patching, upgrades, certificate renewal, node replacement, backup/restore, incident triage, and escalation paths
Recorded or live knowledge-transfer session with our technical team
Apply tot his job
Apply To this Job