Senior SRE - Platform - Kubernetes Engineer at Meraki TalentWorks | Torre

Senior SRE - Platform - Kubernetes Engineer

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Full-time

Legal agreement: Employment

Provide your expected compensation while applying
location_on
Remote (for United States residents)
Shared by
Juan Fernando Domínguez
5 days ago

Responsibilities


About the RoleWe’re looking for a Site Reliability Engineer to support the development and operation of our Kubernetes-based platform in regulated environments. In this role, you will work closely with senior engineers and technical leaders to improve reliability, scalability, and compliance across the platform.This is a hands-on engineering role where you’ll contribute to key systems, build infrastructure to support production operations, and help implement solutions that improve the overall health and performance of the platform.What You Will DoContribute to the design, implementation, and operation of Kubernetes platforms in FedRAMP High / IL5 environmentsSupport day-to-day reliability and performance of platform services, including monitoring and alertingImplement automation and tooling to improve operational efficiency and reduce manual effortWork with senior engineers to define and track SLIs, SLOs, and error budgetsAssist in maintaining compliance and security requirements, including support for audits and continuous monitoringContribute to infrastructure as code and CI/CD pipeline improvementsCollaborate with cross-functional teams (Security, Platform, Application teams) to resolve issues and deliver platform capabilitiesParticipate in on-call rotations supporting customer requests and paging alertsWhat You Bring4–6 years of experience in SRE, DevOps, or platform engineering rolesExperience with Kubernetes in production environmentsFamiliarity with cloud platforms (AWS, Azure, or similar; GovCloud experience a plus)Solid understanding of Linux systems, networking, and containerizationExperience with Infrastructure as Code (e.g., Terraform)Proficiency in scripting or programming (e.g., Python, Go)Exposure to observability tools (Prometheus, Grafana, logging systems)