Senior / Staff Site Reliability, Platform Engineering
saviyntMidway, TX
yesterday
Occupations
Computer Systems Engineers/ArchitectsSoftware DevelopersNetwork and Computer Systems AdministratorsIndustries
Computer Systems Design ServicesSoftware PublishersComputing Infrastructure Providers, Data Processing, Web Hosting, and Related ServicesAbout the role
Overview
In this Staff Platform Engineer role, you will own reliable, scalable, and secure shared infrastructure for a cloud-native SaaS platform. You’ll build and operate core platform components across multi-cloud environments, enabling product teams to ship features faster. Expect hands-on development, platform leadership, and collaboration with cross-functional teams to drive reliability at scale. This role offers impact through shaping platform architecture and improving observability, automation, and deployment practices.
Compensation / Benefitscompetitive compensationbenefitscareer growthsecurity traininghybrid work optionlarge-scale impact
Responsibilities Design and maintain shared infrastructure services used by product teams Build scalable, reusable platform components that abstract complexity for internal developers Operate Kubernetes-as-a-service and multi-region cloud infrastructure Create internal tooling and automation for provisioning and management (Go-focused)Develop and optimize Event-Driven Architecture components and messaging (Kafka, Google Pub/Sub)Maintain CI/CD pipelines as a service (Git Lab CI, ArgoCD)Design reliable distributed systems and resilient data platforms Ensure global availability and performance across multi-region environments Enhance centralized observability and monitoring (Prometheus, Grafana, ELK, Datadog)Provide clear RESTful APIs for infrastructure services and support service mesh capabilities (Envoy, Istio)Collaborate with product teams to address infrastructure needs and participate in on-call rotations Manage relational databases as services (MySQL, PostgreSQL)
Key requirements 6+ years in Infra Development, Platform Engineering, or SREDeep Kubernetes production experience, including multi-tenant setups Strong Go and Python programming for backend services and automation Hands-on cloud experience (AWS, Azure, or GCP); multi-cloud experience a plus Experience with Event-Driven Architecture and message queues (Kafka, RMQ, NATS)CI/CD experience with Git Lab CI; automated delivery for teams Distributed systems design and operation experience Familiarity with multi-region cloud strategies Observability/monitoring proficiency (Prometheus, Grafana, ELK, Datadog)RESTful API design and documentation Service Mesh knowledge (Istio)Relational databases management (MySQL, PostgreSQL)Excellent communication and customer-centric mindset Bachelor’s degree in CS/Engineering or equivalent Security and privacy policy adherence Excellent communication Collaborative mindset Customer-centric focus Kubernetes (platform as a service, multi-tenant)Go (Golang)Python
Matching similar jobs
JOB OVERVIEW
Experience level
Manager
Location
Midway, TX
Occupation
Computer Systems Engineers/Architects
Industry
Computer Systems Design Services
Posted
yesterday
Tired of running searches?
Rank the roles you'd take once, and matches like these arrive on their own.
CREATE PROFILE