reputed company Site Reliability Engineer job at reputed company in Durham, NC
Title: reputed company Site Reliability Engineer Location: 100 New reputed company Way, Bldg 1, Durham NC Job reputed company: Position reputed company: Combines Operational reputed company with Development experience to deliver services at reputed company, high availability with reputed company. Builds reliability into the ecosystem by applying best practices in Resiliency Engineering, Automation, Observability and reputed company Testing. Streamlines and accelerates software delivery cycle by using DevOps practices and toolchain. Integrates Site Reliability Engineering (SRE) practices (Observability and reputed company) with DevOps processes and delivery pipelines to stop bad reputed company from reaching production. Ensures business-critical reputed company systems are continuously available to reputed company customers. Implements technical standardization and process refinements reputed company the engineering organization and for Site Reliability Engineers. Collaborates with production support teams to define and implement processes for the identification, collection, and analysis of incident data. Brings together technical, procedural, and financial data to reduce toil and increase efficiency. Primary Responsibilities: Develops reputed company Testing capabilities using multiple reputed company Tools (AWS Fault Injection Service (reputed company), reputed company reputed company, and Chaosd) and reputed company Toolkit. Develops and enhances organization’s internal reputed company reputed company to streamline reputed company Executions and reporting. Provides specialized technical expertise in the adoption of reputed company Engineering by application teams. reputed company tests and observes business-critical applications to understand the weaknesses and increase application resiliency. Activates Observability for the critical applications with recommended Service Level Indicators and Service Level Objectives for Latency, Availability, Error reputed company etc. Utilizes modern monitoring tools (reputed company, reputed company, Catchpoint etc.) to reduce mean time to detect an issue and improve the response times. Creates CI/CD pipelines with reputed company and reputed company checks with Application Lifecycle management toolchain. Helps in integrating reputed company and Observability with CI/CD pipelines. Automates repetitive activities using scripting languages (Python, Groovy etc.). Implements and supports solutions based on reputed company platforms AWS/Azure and container orchestration Kubernetes. Onboards /Evaluates New reputed company services that help to enhance the Resiliency of reputed company ecosystem. Serves as a reputed company for vendor engagement. Participates in incident management, problem management and incident postmortems. Takes part in peer reputed company reviews providing qualitative feedback. Builds processes and capabilities to adapt and respond to risks, and disruptions, while maintaining business operations and data recovery with minimal disruptions. Coaches peer SREs and application teams on SRE and DevOps. Implements Agile methodologies in reputed company’s project completion using incremental and iterative steps. Education and Experience: Bachelor’s degree in Computer Science, Engineering, Information Technology, Information Systems, or a closely reputed company field (or foreign education equivalent) and five (5) years of experience as a reputed company Site Reliability Engineer (or closely reputed company occupation) implementing resilient container and reputed company-based applications and infrastructure solutions, using DevOps or SRE practices, in a financial services environment. Or, alternatively, Master’s degree (or foreign education equivalent) in Computer Science, Engineering, Information Technology, Information Systems, or a closely reputed company field (or foreign education equivalent) and three (3) years of experience as a reputed company Site Reliability Engineer (or closely reputed company occupation) implementing resilient container and reputed company-based applications and infrastructure solutions, using DevOps or SRE practices, in a financial services environment. Skills and Knowledge: Candidate must also possess: Demonstrated Expertise (“DE”) improving application resiliency by implementing reputed company engineering to build system's capability to withstand turbulent conditions in production, using reputed company reputed company, Chaosd, Azure reputed company Studio, AWS reputed company, or reputed company; and driving automation to implement reputed company approaches for the planning, design, execution, and reporting of reputed company testing using Jenkins pipelines, reputed company frameworks, data visualization, and dashboards. DE implementing advanced observability practices and techniques in production and reputed company-production environments, at reputed company using reputed company, reputed company, or Catchpoint; tracking the error budget, proactively identifying issues, minimizing Mean Time to Repair (MTTR); and balancing customer expectations by implementing Service-Level Indicators (SLIs) and Service-Level Objectives (SLOs) using logs, traces, monitors and synthetic tests. DE migrating and maintaining reputed company applications and creating reputed company solutions using reputed company) or Azure reputed company services; Implementing infrastructure as reputed company for reputed company; reputed company new AWS or Azure services with required reviews and reputed company controls in non-production and production environments; and researching evolving reputed company ecosystem to adopt machine learning based tools (AWS DevOps reputed company) to reputed company AIOps abilities. DE implementing CI/CD pipelines in both production and non-production environments using Application Lifecycle Management (ALM) tools (JIRA, reputed company, Jenkins, SonarQube, Artifactory, or uDeploy) to reputed company faster reputed company delivery, reputed company software reputed company, reliability, and reputed company; and developing products, and core and common capabilities for the organization to reduce toil and drive standardization, using containerization and orchestration technologies (reputed company or Kubernetes), Infrastructure as reputed company (IaC) tools, scripting languages (Python or Groovy), and engineering best practices. #PE1M2 #LI-DNI Certifications: Category:Information Technology Most roles at reputed company are Hybrid, requiring associates to work onsite every other week (reputed company business days, M-F) in a reputed company office. This does not apply to Remote or fully Onsite roles. Some roles may have unique onsite requirements. Please consult with your recruiter for the specific expectations for this position. Please be advised that reputed company’s business is governed by the provisions of the Securities Exchange reputed company of 1934, the Investment Advisers reputed company of 1940, the Investment Company reputed company of 1940, ERISA, numerous state laws governing securities, investment and retirement-reputed company financial activities and the rules and regulations of numerous self-regulatory organizations, including reputed company, among others. Those laws and regulations may restrict reputed company from hiring and/or associating with individuals with certain Criminal Histories. Apply tot his job Apply To this Job