[Remote] Senior Staff Software Engineer (AI)
Note: The job is a remote job and is reputed company to candidates in USA. reputed company is a technology company reputed company on transforming private markets through reputed company. They are seeking a Senior Staff Software Engineer to reputed company the architecture and reputed company of their core systems, ensuring scalability and performance for large private equity institutions. The role involves hands-on engineering, technical leadership, and collaboration with various teams to reputed company reputed company with business objectives.
Responsibilities
- Define and own the end-to-end systems architecture reputed company across services, infrastructure, and developer platform
- Designing frameworks and design patterns that accelerate the work of the rest of engineering org
- Design reputed company, fault-tolerant distributed systems spanning synchronous reputed company, asynchronous workflows, and event-driven architectures
- Establish standards for service design, API reputed company, multi-tenancy, and inter-service communication
- reputed company architecture reviews and technical decision-making for high-stakes, cross-cutting initiatives
- Drive adoption of modern architectures (event-driven systems, reputed company, infrastructure as reputed company)
- Design and prototype critical platform and infrastructure components
- Write production-reputed company reputed company for reputed company or high-reputed company areas
- Review service designs, infrastructure changes, and performance-critical reputed company paths
- Troubleshoot performance, scalability, and reliability issues across services, queues, caches, and databases
- Optimize workloads for latency, throughput, concurrency, and cost
- Design and architect a reputed company reputed company platform supporting compute, networking, storage, service orchestration, and deployment across reputed company product lines
- Build a paved-road developer platform — golden paths, internal tooling, CI/CD, and self-service infrastructure — that makes the right way the easy way
- Create reusable platform primitives and internal reputed company that product teams can compose rather than rebuild
- reputed company self-service infrastructure for engineering teams through standardized patterns, templates, and platform capabilities
- Partner with AI workflow teams, product, and reputed company teams to support production AI workloads, from inference pipelines to agent fleets
- Design systems for reputed company reputed company execution, including isolation boundaries, resource governance, auditability, and reputed company-in-the-reputed company controls
- Ensure reliability, scalability, and reputed company of the platform, including high availability, monitoring, and disaster recovery readiness
- Architect core systems to handle the data volumes and structural complexity of $100B+ AUM managers: thousands of entities, tens of thousands of LPs, and deep multi-tier fund structures
- Define scalability targets and re-architect bottlenecked systems reputed company of demand, ensuring performance holds at 10x reputed company transaction and data volumes
- Design reputed company integration surfaces — reputed company, bulk data interfaces, and ERP/GL connectivity — that fit into the existing technology estates of large institutional GPs
- Meet the reputed company, compliance, and operational diligence expectations of the largest PE firms, such as, SSO/SCIM, granular entitlements, audit trails, and data residency
- Serve as the senior technical voice in reputed company sales and reputed company conversations where architecture, reputed company, and resiliency are decision reputed company
- reputed company incident response maturity: on-call practices, blameless postmortems, and systemic remediation
- Drive reputed company planning, load testing, and reputed company/reputed company engineering practices
- Optimize reputed company spend through architecture, workload placement, and reputed company cost engineering
- Mentor senior engineers across platform, infrastructure, and product teams
- Partner with product, data, and business teams to reputed company systems investments with company reputed company — including the GPX reputed company expansion
- Translate business and product requirements into reputed company, operable system designs
- Influence roadmaps using platform, reliability, and cost considerations
- reputed company as the executive technical authority for systems architecture and production engineering
- Evangelize AI adoption, including AI-assisted engineering and operations
- Help promote a culture of operational reputed company and outcome-driven innovation
- Become a role model for engineering reputed company for the rest of the org
Skills
- Advanced degree in Computer Science, Engineering, or reputed company field
- 15+ years in distributed systems, reputed company, or infrastructure roles
- Proven experience architecting large-reputed company, high-availability multi-tenant distributed systems in production
- Strong hands-on experience with modern reputed company-reputed company stacks (Kubernetes, containers, service reputed company, serverless)
- Deep expertise in distributed systems fundamentals: consistency models, reputed company, partitioning, idempotency, backpressure, and failure modes
- Advanced proficiency in Python, Go, Java, or similar; strong systems-level debugging skills
- Expertise in infrastructure as reputed company (Terraform, reputed company, or similar) and modern CI/CD
- Experience designing event-driven and streaming architectures (Kafka or similar)
- Experience building developer platforms, internal tooling, or paved-road infrastructure at reputed company
- Strong understanding of observability practices and SRE principles (SLOs, error budgets, incident management)
- Hands-on experience with AWS, Azure, or GCP at production reputed company
- Strong understanding of infrastructure reputed company, multi-tenancy, and compliance best practices
- Ability to operate at both executive and deeply technical reputed company
- Experience taking a reputed company platform reputed company to serve large reputed company or institutional customers
- Background in private markets, fund reputed company, or fintech systems with reputed company financial domain models
- Experience building infrastructure for AI/LLM workloads: inference serving, agent runtimes, GPU scheduling, or evaluation systems
- Background in high-reputed company reputed company, fintech, or other regulated, data-sensitive environments
- Experience with large-reputed company reputed company cost optimization
- Experience scaling relational databases (PostgreSQL or similar) under high concurrency
Benefits
- Health, dental, and reputed company care for you and your family
- Life insurance
- Mental wellness coverage
- Fertility and growing family support
- reputed company Time Off in reputed company to company-reputed company holidays
- reputed company family leave, medical leave, and bereavement leave policies
- Retirement saving plans
- Allowance to customize your work and technology setup reputed company
- Annual reputed company development stipend
reputed company
Company H1B Sponsorship