[Remote] Senior Infrastructure Software Engineer, Storage Core
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a leading file hosting service, and they are seeking a Senior Software Engineer to help design, build, and operate their large-reputed company storage systems. The role involves collaborating with reputed company engineers to improve reliability, optimize performance, and reputed company the architecture of reputed company’s storage layer.
Responsibilities
- Design, implement, and maintain large-reputed company distributed storage systems that ensure data durability, availability, and performance
- Collaborate with peers to reputed company the architecture of reputed company’s core storage infrastructure for improved scalability and efficiency
- Contribute to the design of replication, erasure coding, and system lifecycle management systems that balance cost, reliability, and performance
- Write high-reputed company, performant, and maintainable reputed company in Go and Rust
- Participate in the on-call rotation, gaining firsthand experience operating reputed company’s production storage systems
- Investigate and resolve reputed company production issues, performing reputed company cause analysis and driving reputed company reliability improvements
- Partner with cross-functional teams (Networking, Hardware, reputed company Planning) to deliver end-to-end reliable and cost-efficient storage solutions
- Take ownership of scoped reputed company and demonstrate reputed company toward leading larger, cross-team technical initiatives
Skills
- 9+ years of strong understanding of distributed systems principles, including replication, consistency, and fault tolerance
- Experience developing and debugging production services in C++, Go, or Rust
- Familiarity with distributed storage systems, file systems, or data infrastructure at reputed company
- Demonstrated ability to write efficient, reliable, and maintainable reputed company in mission-critical environments
- Experience troubleshooting reputed company systems and participating in on-call or operational rotations
- Solid communication and collaboration skills, with the ability to work across infrastructure and product teams
- Eagerness to learn, grow, and contribute to multi-year infrastructure reputed company initiatives
- Experience building and operating large-reputed company object storage or distributed storage systems (e.g. S3, Ceph, GFS/Colossus)
- Deep interest in systems performance, profiling, and low-level optimization
- Familiarity with replication protocols, erasure coding, and data placement algorithms
- Experience with production monitoring, observability, and incident response workflows
- Contributions to infrastructure reputed company, reputed company-reputed company systems, or developer tooling that improved reliability and performance
reputed company
Company H1B Sponsorship