Site Reliability Engineer (SRE)
Infosys · Santiago
Job description
About the role
The Site Reliability Engineer (SRE) ensures the reliability, availability, and performance of production digital services. You will balance service stability with rapid delivery, working closely with application and platform teams to embed resilience into every change.
Key responsibilities
- Define, implement and monitor Service Level Indicators (SLIs) and Service Level Objectives (SLOs) for availability, latency and error rates.
- Lead incident response (L2/L3), conduct blameless post‑mortems and turn recurring issues into engineering backlog items.
- Identify repetitive operational tasks and build automation for deployments, monitoring, health‑checks and self‑healing mechanisms.
- Support change governance by ensuring traceability, rollback strategies and production readiness.
- Maintain near‑real‑time operational KPIs, assign clear ownership for deviations and drive data‑driven continuous improvement.
Required profile
- Proven experience as an SRE, Production Engineer or similar role.
- Strong background in production systems and reliability engineering.
- Experience working with cloud environments, CI/CD pipelines and integration points.
Required skills
- Cloud platforms (e.g., AWS, Azure, GCP)
- CI/CD pipeline tools
- Monitoring and alerting systems
- Automation scripting (e.g., Python, Bash)
- Health‑check and self‑healing design
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Chile.
Salaries by job title
Apply in 30 seconds
Enter your email to apply. An account will be created automatically.
By continuing, you accept our terms of use.
Already have an account? Login
Published 1 month ago
Expires 2 weeks from now
22 views · 0 interested
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Infosys
Santiago