Site Reliability Engineer at Orion Groups | Torre

Site Reliability Engineer

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Freelance
Recurrent
Compensation
USD170k - 190k/year
location_on
Remote (for United States residents)
Shared by
Julien PETIT
4 days ago

Responsibilities


OVERVIEWAt Orion, we are seeking a Site Reliability Engineer to design, build, and operate scalable, secure, and reliable cloud infrastructure and automation to ensure seamless, efficient delivery of healthcare technology that supports clinicians and patients.OPPORTUNITY6-month W2 contract with extension; Fully Remote (U.S. – All Time Zones)KEY RESPONSIBILITIESDesign, implement, and maintaure, and highly available cloud infrastructure using Infrastructure-as-Code (IaC) tools and modern DevOps practices.Build and optimize CI/CD pipelines using tools such as Jenkins, GitHub Actions, Maven, and JFrog to ensure fast, reliable, and repeatable software delivery.Develop and manage containerized application environments with Kubernetes, ensuring optimal deployment, scalability, and service reliability.Automate infrastructure provisioning, configuration, and deployment processes to improve operational efficiency and reduce manual interventions.Design and implement comprehensive monitoring, alerting, and observability solutions to ensure system health, performance, and reliability.Collaborate closely with development, product, QA, and security teams to design robust platform solutions aligned with business and technical requirements.Participate actively int planning, demos, and retrospectives, driving continuous improvement and team collaboration.Contribute to operational support processes, including incident response, root cause analysis, capacity planning, and performance optimization for large-scale distributed systems.KEY EXPERIENCE5+ years of hands-on experience with tools such as Terraform and Ansible for building and managing infrastructure as code, and CI/CD automation tools like Jenkins, GitHub, Maven, and JFrog.Proven hands-on experience deploying and managing infrastructure and services on Google Cloud Platform (GCP) is required. Strong knowledge of GCP networking, IAM, security, cost optimization, and core services (e.g., Cloud Run, Pub/Sub, GKE, Firestore) is essential. Experience with AWS or Azure is a plus but not a substitute.Strong experience deploying, scaling, and managing containerized applications using Kubernetes, including service mesh, auto-scaling, and rolling update strategies.Proficiency , Java, C++, Perl, Ruby, or SQL, with the ability to automate workflows and optimize infrastructure operations.Experience designing and implementing operational support models, including incident response, root cause analysis, monitoring, and alerting strategies.Demonstrated experience with reliability engineering practices such as capacity planning, fault tolerance, and resilience design.Experience working ly stand-ups, planning sessions, and retrospectives.EDUCATIONBachelor’s degree ormation Technology, Computer Science, Engineering, or a related field — or equivalent work experience.INTERVIEW PROCESSQuick turnaround: preliminary interview followed by team interview on phone and Google Meets.Typically, one round with the hiring manager, joined by the team (virtual interview), possibly two rounds if necessary.ORION GROUPSis a professional services firm as that provides experienced consultants to key organizations across the United States. We focus our talent expertise in both the financial and healthcare industries. Our capabilities erprise software systems, data strategies, integration, and IT infrastructure. We work diligently to listen and provide value to your unique needs.