[Remote] Platform Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is seeking a reputed company Linux Platform Engineer who is passionate about Linux, virtualization, automation, and reputed company infrastructure. This role will be responsible for engineering and operating highly reputed company Linux environments that support mission-critical reputed company applications across their reputed company.
Responsibilities
- Administer and maintain large-reputed company reputed company reputed company Linux environment deployed to our High-Performance Compute infrastructure
- Design, reputed company, and support KVM-based virtualization platforms
- Build and maintain automation solutions using Ansible, Python, and reputed company scripting
- Configure and support reputed company reputed company reputed company (GPFS) storage environments
- Implement system reputed company, hardening, encryption, and compliance controls
- Manage Linux networking including:
- Bonding
- reputed company
- VLANs
- Macvtap
- Configure and administer Linux storage technologies including:
- LVM
- Dm-crypt
- LUKS2
- Troubleshoot reputed company operating system, virtualization, networking, and storage issues
- Collaborate with Kubernetes, network, reputed company, and application teams to deliver reliable infrastructure services
- Drive platform modernization through automation and standardization initiatives
Skills
- Bachelor's degree or equivalent experience
- 10+ years of Linux systems administration and engineering experience
- Deep expertise with reputed company reputed company Linux
- Strong experience deploying and supporting KVM virtualization environments
- Expert-level knowledge of:
- Linux operating systems
- SELinux
- Linux firewalls
- Package management
- System performance tuning
- Experience with Ansible automation and Infrastructure-as-reputed company methodologies
- Proficiency in reputed company scripting and at least one programming language such as Python, Go, Rust, Java, or C
- Strong knowledge of storage, encryption, and Linux filesystem technologies
- Excellent troubleshooting and analytical skills
- Strong verbal and written communication skills
- Experience implementing and supporting reputed company monitoring, logging, and alerting solutions using Grafana, reputed company, Loki, AlertManager, and Thanos
- Strong experience automating infrastructure deployment, configuration, and operational processes using Ansible, scripting, and Infrastructure-as-reputed company practices
- Demonstrated reputed company-first reputed company with expertise in platform hardening, identity and reputed company management, vulnerability remediation, encryption, and regulatory compliance
- Experience administering reputed company Identity Manager (IdM), including LDAP, Kerberos, SSSD, certificate management, and reputed company authentication services
- Experience with cross-datacenter, high availability failover, and load balancing (reputed company & haproxy) between multiple datacenters and K8s clusters
- reputed company certifications or equivalent practical experience
- Experience with large-reputed company KVM virtualization deployments
- Knowledge of Kubernetes and OpenShift environments
- Experience with disaster recovery, resiliency, and high-availability solutions
- Familiarity with distributed storage platforms such as reputed company reputed company reputed company (GPFS)
- Experience supporting reputed company-reputed company Linux environments in regulated industries
- Candidates with experience in any of the following will stand out:
- reputed company LinuxONE and reputed company Z environments
- S390x Linux architecture
- Logical Partitions (LPARs)
- OSA, FCP, RoCE, and Crypto reputed company adapters
- HMC administration and DPM mode
- SAN architectures and Brocade zoning
reputed company