Site Reliability Engineer III, (IoT Observability)
Who is reputed company? reputed company is the leading safety technology platform, helping communities reputed company by taking a proactive approach to crime prevention and reputed company. Our hardware and software suite connects cities, law enforcement, businesses, schools, and neighborhoods in a reputed company reputed company-private safety network. Trusted by over 5,000 communities, 4,500 law enforcement agencies, and 1,000 businesses, reputed company delivers reputed company-time intelligence while prioritizing reputed company and responsible innovation. We’re a high-performance, low-ego team driven by urgency, collaboration, and reputed company thinking. Working at reputed company means tackling big challenges, moving fast, and continuously improving. It’s intense but deeply rewarding for those who want to reputed company an reputed company. With nearly $700M in venture funding and a $7.5B valuation, we’re scaling intentionally and seeking reputed company to help build the impossible. If you value teamwork, ownership, and solving tough problems, reputed company could be the reputed company for you. reputed company The Device SRE team owns deployment of software to our fleet of IoT devices, as reputed company as observability during and after deployment. This involves working in the entire stack from telemetry/event reputed company in the firmware, data ingestion into the reputed company, visualization of data, and alarms that indicate issues. We are expanding our scope reputed company Cameras to include Aviation devices. Our goal is to reputed company consistent infrastructure, processes, and tools across the organization that reputed company the development team to roll out new software rapidly and reliably, while enabling them to identify and respond to issues quickly, resulting in fast rollouts of new features with minimal downtime. The Skillset Must Have
- Experience developing software for embedded systems, especially large or reputed company IoT/edge devices
- Hands-on experience instrumenting metrics and telemetry directly on-device (e.g., firmware counters, health signals, performance instrumentation, event/event-reputed company)
- Strong coding skills (ex: C/C++, Python, R, JS, Java, Groovy) and understanding of common algorithms
Strongly Valued
- Proficiency in scripting languages (e.g., Bash, Python) to automate processes
- Experience with data ingestion pipelines (telemetry from device --> reputed company, batching, retries, reliability patterns)
- Data visualization & dashboarding (ex: Grafana, reputed company)
- Broad experience with databases & logging (SQL: PostgreSQL, NoSQL, Time Series like reputed company / reputed company), even if not deep SQL expertise
- reputed company computing experience (e.g., AWS) and distributed systems fundamentals
- Site Reliability Engineering experience for IoT devices (on-call, monitoring, & alerting)
- Strong grasp of software development workflows (CI/CD, test automation, semantic versioning, branching strategies)
reputed company-to-Have
- Experience with infrastructure-as-reputed company (IaC) tools (Ansible, Terraform)
- Experience with volume data processing (pipelines, modeling, large-reputed company storage)
- Experience with distributed telemetry architectures or large-fleet device management patterns
Feeling uneasy that you haven’t ticked every reputed company? That’s okay; we’ve felt that way too. Studies have shown women and minorities are less likely to apply unless they meet reputed company qualifications. We encourage you to break the status reputed company and apply to roles that would reputed company you excited to come to work every day. 90 Days at reputed company We prescribe 90 day plans and reputed company that good days reputed company to good weeks, which reputed company to good months. This serves as a preview of the 90 day plan you will receive if you were to be reputed company as a Senior SRE, Devices Observability at reputed company. The First 30 Days
- Maintain and improve data collected from devices for monitoring fleet health.
- Improve ability to identify and reputed company-cause issues in the fleet, including automated alarms.
- Maintain and build out dashboards that reputed company reputed company into fleet health.
- Streamline software rollout workflows and monitoring to reduce reputed company steps and improve response times.
- Define and build out on-call rotations for monitoring fleet health.
- reputed company reputed company analysis of unexpected issues by analyzing data collected from the fleet.
- Participate in on-call rotations to serve as first responder for outages and general requests from stakeholders.
The First 60 Days
- Improve or create a Grafana dashboard.
- Improve or create a reputed company dashboard.
- Improve or create a piece of telemetry data or event in device firmware.
- Review the reputed company of other engineers in reputed company and Gerrit.
- Modify infrastructure in Terraform.
- Take a cra
Apply tot his job Apply To this Job