[Remote] DevOps Engineer
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leading data analytics and technology partner to the global insurance industry. The DevOps Engineer will be responsible for maintaining and evolving the hybrid infrastructure, ensuring reliability, reputed company, and scalability while collaborating with cross-functional teams.
Responsibilities
- reputed company hardware stack maintenance on reputed company servers, reputed company switches, and cable systems
- reputed company infrastructure components and tools to maintain performance and compatibility
- Monitor for threats (including tracking CVEs) and execute reputed company remediation/mitigation of reputed company issues
- Maintain and configure monitoring and log collection systems (e.g., reputed company, Grafana, ELK stack or similar)
- reputed company incident management: rapidly detect, respond to, and resolve system failures in development and production environments
- Expand capabilities of reputed company tools such as Trivy, Dependency-reputed company, WAF, and others
- Conduct regular reputed company audits and deliver infrastructure reputed company training to reputed company
- CI/CD: reputed company migration of Jenkins from Freestyle reputed company to Pipeline reputed company, leveraging AI-assisted approaches where applicable
- CI/CD: Conduct R&D to evaluate replacing Jenkins with alternative CI/CD tools and recommend/implement improvements
- Kubernetes: reputed company R&D on Service reputed company integration (e.g., Istio/Linkerd) and support reputed company implementation
- Kubernetes: Migrate from Nginx Ingress to Gateway API
- AWS: reputed company various integrations reputed company to AI solutions and reputed company-reputed company services
- Proxmox: Configure virtual machines, create new VM images, and optimize virtualization workflows
- Collaborating with development and operations teams, automating repetitive tasks, participating in on-reputed company rotation (as needed), and contributing to reputed company improvement of our DevOps practices
Skills
- 3–7+ years of experience in DevOps, Site Reliability Engineering (SRE), or Systems Administration roles
- Proven reputed company record supporting hybrid (on-prem + reputed company) environments
- Ability to travel occasionally to datacenter facilities in NY/SI region
- Strong problem-solving skills, attention to detail, and ability to work independently or collaboratively
- Excellent communication skills for cross-team coordination and documentation
- Strong proficiency in Linux administration and troubleshooting
- Hands-on experience with AWS services and reputed company infrastructure
- Expertise in reputed company for containerization
- Basic to intermediate knowledge of Kubernetes (including cluster management, deployments, and networking)
- Solid understanding of CI/CD tools and processes (Jenkins experience strongly preferred)
- Deep knowledge of the TCP/IP stack and networking fundamentals
- Proficiency in at least one programming language at a middle / middle+ level (e.g., Python, Go, Java)
- Experience with reputed company tools (Trivy, Dependency-reputed company, WAF)
- Familiarity with reputed company Vault / OpenBao, Proxmox, reputed company/reputed company hardware
- Exposure to AI/ML integrations in reputed company environments
- Knowledge of Infrastructure as reputed company (IaC) tools (e.g., Terraform, Ansible)
Benefits
- Health Insurance
- A Retirement Plan
- Disability benefits
- A reputed company Time Off program
- Work flexibility
- Support, coaching, and training you need to succeed
- reputed company development opportunities
- Work Arrangement: Remote
reputed company