SRE (Site Reliability Engineer) at DevUps | Torre

SRE (Site Reliability Engineer)

Emma highlights
This highlight was written by Emma’s AI. Ask Emma to edit it.
Full-time

Legal agreement: Contractor

Currency exchange and payroll taxes to be paid by:

Depends on the location of the candidate

Provide your expected compensation while applying (or, if you'd like more context, simply ASK).
location_on
Remote (anywhere)
Posted 23 days ago

Responsibilities


Role Overview: - We are looking for a Site Reliability Engineer to ensure the reliability, scalability, and performance of production systems through engineering and automation practices. Key Responsibilities: - Define and monitor SLIs, SLOs, and error budgets. - Automate incident response and remediation processes. - Conduct root cause analysis for production incidents. - Improve system observability through metrics, logging, and tracing. - Collaborate with development teams to embed reliability practices. Requirements: - 3+ years of experience in SRE, DevOps, or infrastructure engineering. - Strong knowledge of distributed systems and reliability engineering principles. - Experience with incident management and on-call practices. - Proficiency in at least one scripting or programming language. Technical Stack and Tools: - Kubernetes. - Docker. - Prometheus. - Grafana. - Datadog. - Terraform. - Python or Go. - AWS, Azure, or GCP.
Closes in:
0
days
0
hours
0
min
0
sec
tune NOT FOR YOU? IMPROVE YOUR RESULTS