Senior Site Reliability Engineer
Hybrid in Argentina, & 4 others
Site Reliability Engineering& 3 others
Looking for something else?
Find a vacancy that works for you. Send us your CV to receive a personalized offer.
Find me a jobChoose an option
We are seeking a skilled and results-driven Senior Site Reliability Engineer to join our engineering team and bridge the gap between software development and systems operations. In this role, you will apply engineering principles to automate operations, scale infrastructure, and keep systems highly available, resilient, and performant. Your mission is to build, run, and safeguard the production environments that power our applications, reducing downtime and enabling fast, safe software delivery.
Responsibilities
- Architect, build, and maintain cloud infrastructure using modern IaC practices such as Terraform and CloudFormation
- Develop and refine CI/CD pipelines to streamline software deployments, configuration management, and routine operational tasks
- Implement comprehensive logging, monitoring, and alerting systems using tools like Prometheus, Grafana, and Datadog
- Define clear Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
- Lead incident response efforts and troubleshoot production issues to restore services promptly
- Facilitate blameless post-mortems to uncover root causes and prevent future occurrences
- Collaborate with software developers to enhance system performance and plan for capacity needs
- Guarantee services scale effectively to accommodate growth and traffic surges
Requirements
- 3+ years of experience in systems administration, DevOps, or systems-oriented software development
- Competency in at least one scripting or programming language such as Python, Bash, Go, or Rust
- Hands-on experience with public cloud platforms such as AWS, Azure, or GCP, along with containerization tools like Docker and Kubernetes
- Solid understanding of Linux/Unix administration and networking fundamentals such as TCP/IP, DNS, and HTTP/SSL/TLS
- Demonstrated reliability mindset with a strong drive for automation, reducing toil, and designing resilient systems that fail gracefully
- English proficiency at B2 level or higher
