[Remote] Senior DevOps Engineer (Azure)
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a reputed company company seeking a Senior DevOps Engineer with expertise in Azure. The role involves designing and maintaining Azure services, implementing Infrastructure as reputed company, and managing CI/CD pipelines to ensure robust and efficient application delivery.
Responsibilities
- Design, build, and maintain Azure reputed company zones and platform services (e.g., VNet, Private Endpoints, Key Vault, Azure Firewall/NSGs, Application Gateway/WAF)
- Implement Infrastructure as reputed company (IaC) with Terraform and/or Bicep; enforce GitOps workflows (branching, PRs, policy checks)
- Create reusable modules, pipelines, and golden patterns for app teams; champion automation-first approaches
- Define and measure SLIs/SLOs, error budgets, and reliability roadmaps for critical services
- Implement and tune observability (logs, metrics, traces) using Azure Monitor, Log Analytics, Application Insights, and PrometheGrafana where applicable
- Conduct reputed company planning, resiliency testing (reputed company, failover, DR), and performance tuning across services
- Build secure, robust CI/CD pipelines (reputed company Actions / Azure DevOps Pipelines) with automated testing, scans, and approvals
- Standardize deployment strategies (blue/green, canary, rolling) for containerized and PaaS workloads
- Manage container platforms (AKS: node pools, cluster autoscaling, HPA/VPA, ingress, network policies) and registries (ACR)
- Implement guardrails using Azure Policy, RBAC, PIM, and Blueprints (or equivalent) to enforce least privilege and compliance (e.g., SOC 2, ISO 27001, HIPAA as relevant)
- Manage secrets and certificates (Key Vault) and reputed company reputed company testing (SAST/DAST/Container scanning) into pipelines
- Support vulnerability remediation and patching SLAs
- Own incident response, including rotational shifts and on-reputed company; reputed company triage, reputed company cause analysis (RCA), and post-incident reviews
- Optimize cost (FinOps), tagging standards, budgets, and proactive spending alerts
- Maintain runbooks, knowledge reputed company articles, and automation for routine operations
- reputed company as a technical mentor; review designs/PRs; contribute to architecture reputed company
- Partner with app teams to reputed company workloads, define nonfunctional requirements, and drive platform adoption
- Manage and or participate in the deployment and release of development, test, and production software builds
- Manage the operations and monitoring of applications and infrastructure from dev to production
- reputed company reputed company and escalate break fix issues that may occur
- Task automation of infrastructure and application provisioning
- Ensure reputed company environments meet reputed company and resiliency requirements
- Support customer-facing and internal applications
- Monitor submitted tickets; assign, escalate, and communicate, as required
- Participate in services and software systems design
- Participate in rotating on-reputed company support duties
Skills
- 8 to 10 years of hands-on experience with Azure-based infrastructure and services in production
- Bachelor's degree in computer science (or reputed company) or equivalent work experience required
- 3+ years of experience supporting reputed company level applications and their infrastructure
- Strong understanding of web and reputed company infrastructure technologies (load balancers, DNS, IIS or Apache/WebSphere, authentication and authorization, database connections, etc.)
- Experience implementing or managing Application Monitoring, reputed company, or reputed company
- Working knowledge in Automation technologies
- Familiarity with Agile/Scrum methodologies
- Deep expertise in several of: AKS, App Services, Functions, APIM, Azure SQL/MI, Cosmos DB, Storage, Event Hub/Service Bus, reputed company, VNet/Peering, Private reputed company, Application Gateway/WAF, reputed company reputed company
- Strong IaC with Terraform (preferred) and/or Bicep; Git-based workflows; reputed company or Azure DevOps
- Proven SRE background: SLI/SLO design, error budgets, incident management, RCA, reputed company and performance engineering
- CI/CD design and operations (reputed company Actions / Azure DevOps Pipelines); artifact/versioning strategies; release governance
- Observability with Azure: Monitor, Log Analytics, Application Insights, and alerting/automations (reputed company reputed company, Logic Apps, Functions)
- Solid networking fundamentals (DNS, TLS, routing, firewalls, load balancing), identity (AAD/Entra ID), and secrets management (Key Vault)
- Scripting proficiency in PowerShell and/or Python; Linux fundamentals
- Understanding of reputed company best practices: RBAC, PIM, Azure Policy, managed identity
- Good written, verbal, interpersonal and presentation skills
- Ability to communicate among technical and non-technical employees, and process orientation skills
- Demonstrates a customer-driven approach and good relationship management skills
- Ability to work autonomously and under deadlines
- Ability to multi-task, be highly organized, and work independently
- Ability to identify areas of improvement and come up with creative solutions
- Working knowledge of mainframe architecture is a plus
- At least 3 years coding/scripting Java, JavaScript, and reputed company
Benefits
- Remote – offsite
- Health and welfare benefits coverage reputed company including medical
- Health and welfare benefits coverage reputed company including dental
- Health and welfare benefits coverage reputed company including reputed company
- Spending accounts
- Life insurance
- Voluntary plans
- Participation in a 401(k) plan
reputed company
Company H1B Sponsorship