Senior / Staff Site Reliability, Platform Engineering

saviyntMidway, TX

yesterday

Occupations

Computer Systems Engineers/ArchitectsSoftware DevelopersNetwork and Computer Systems Administrators

Industries

Computer Systems Design ServicesSoftware PublishersComputing Infrastructure Providers, Data Processing, Web Hosting, and Related Services
APPLY NOW

About the role

Overview In this Staff Platform Engineer role, you will own reliable, scalable, and secure shared infrastructure for a cloud-native SaaS platform. You’ll build and operate core platform components across multi-cloud environments, enabling product teams to ship features faster. Expect hands-on development, platform leadership, and collaboration with cross-functional teams to drive reliability at scale. This role offers impact through shaping platform architecture and improving observability, automation, and deployment practices. Compensation / Benefitscompetitive compensationbenefitscareer growthsecurity traininghybrid work optionlarge-scale impact Responsibilities Design and maintain shared infrastructure services used by product teams Build scalable, reusable platform components that abstract complexity for internal developers Operate Kubernetes-as-a-service and multi-region cloud infrastructure Create internal tooling and automation for provisioning and management (Go-focused)Develop and optimize Event-Driven Architecture components and messaging (Kafka, Google Pub/Sub)Maintain CI/CD pipelines as a service (Git Lab CI, ArgoCD)Design reliable distributed systems and resilient data platforms Ensure global availability and performance across multi-region environments Enhance centralized observability and monitoring (Prometheus, Grafana, ELK, Datadog)Provide clear RESTful APIs for infrastructure services and support service mesh capabilities (Envoy, Istio)Collaborate with product teams to address infrastructure needs and participate in on-call rotations Manage relational databases as services (MySQL, PostgreSQL) Key requirements 6+ years in Infra Development, Platform Engineering, or SREDeep Kubernetes production experience, including multi-tenant setups Strong Go and Python programming for backend services and automation Hands-on cloud experience (AWS, Azure, or GCP); multi-cloud experience a plus Experience with Event-Driven Architecture and message queues (Kafka, RMQ, NATS)CI/CD experience with Git Lab CI; automated delivery for teams Distributed systems design and operation experience Familiarity with multi-region cloud strategies Observability/monitoring proficiency (Prometheus, Grafana, ELK, Datadog)RESTful API design and documentation Service Mesh knowledge (Istio)Relational databases management (MySQL, PostgreSQL)Excellent communication and customer-centric mindset Bachelor’s degree in CS/Engineering or equivalent Security and privacy policy adherence Excellent communication Collaborative mindset Customer-centric focus Kubernetes (platform as a service, multi-tenant)Go (Golang)Python

Matching similar jobs

JOB OVERVIEW

Experience level

Manager

Location

Midway, TX

Occupation

Computer Systems Engineers/Architects

Industry

Computer Systems Design Services

Posted

yesterday

Tired of running searches?

Rank the roles you'd take once, and matches like these arrive on their own.

CREATE PROFILE
Senior / Staff Site Reliability, Platform Engineering at saviynt |...