Senior Site Reliability Engineer (SRE)
Hybrid in Argentina, & 3 others
Site Reliability Engineering& 3 others
Looking for something else?
Find a vacancy that works for you. Send us your CV to receive a personalized offer.
Find me a jobChoose an option
We are seeking a skilled and proactive Senior Site Reliability Engineer (SRE) to join our engineering team. In this role, you will bridge the gap between software development and systems operations, applying software engineering principles to automate operations, scale infrastructure, and ensure systems remain highly available, resilient, and performant. Your mission is to build, run, and protect the production environments that power our applications, minimizing downtime and helping us deploy software rapidly and safely.
Responsibilities
- Design, build, and maintain cloud infrastructure using modern IaC practices such as Terraform and CloudFormation
- Build and optimize CI/CD pipelines to automate software deployments, configuration management, and repetitive operational tasks
- Design and implement robust logging, monitoring, and alerting systems to establish clear Service Level Objectives (SLOs) and Service Level Indicators (SLIs)
- Respond to production incidents and lead troubleshooting efforts to restore services quickly
- Conduct blameless post-mortems to identify root causes and prevent recurrence
- Partner with software developers to optimize system performance and plan capacity
- Ensure services can scale to handle growth and traffic spikes
Requirements
- 3+ years of experience in systems administration, DevOps, or systems-focused software development
- Proficiency in at least one scripting or programming language such as Python, Bash, Go, or Rust
- Experience with public cloud providers including AWS, Azure, or GCP, and containerization tools such as Docker and Kubernetes
- Understanding of Linux/Unix administration and networking fundamentals including TCP/IP, DNS, and HTTP/SSL/TLS
- Familiarity with monitoring and observability tools such as Prometheus, Grafana, and Datadog
- Passion for automation, eliminating toil, and building resilient systems that fail gracefully
- English proficiency at B2 level or higher
