Looking for something else?
Find a vacancy that works for you. Send us your CV to receive a personalized offer.
Find me a jobChoose an option
We are building a Lead DevOps Engineer role to replace brittle operations with automated, AI-assisted workflows while owning AWS infrastructure and CI/CD end-to-end. You will architect resilient AWS systems, expand Jenkins pipelines, and use coding agents such as Claude Code as a primary delivery mechanism to boost team throughput and reliability.
Responsibilities
- Spot manual, repetitive, or error-prone operational work and replace it with automated, AI-assisted workflows, then quantify recovered time and reduced toil
- Build and sustain infrastructure as code in Terraform, including modules, state strategy, environment promotion, drift detection, and remediation
- Own and enhance Jenkins pipelines for build, test, security scan, artifact, and deploy stages, including shared libraries and reusable pipeline patterns
- Integrate and calibrate SonarQube quality gates so code-quality and security signals are actionable without generating noise
- Drive implementation using coding agents such as Claude Code by breaking down problems, setting specs and guardrails, directing execution, and rigorously reviewing outcomes
- Design and debug AWS systems end to end across compute, networking, IAM, and storage, including expected behavior during service failures
- Respond to and resolve production and pipeline incidents, then codify learnings in runbooks and automation to prevent repeat issues
- Assess whether existing tooling is still the right fit and, when change is needed, plan and execute migrations methodically while keeping developers supported and productive
- Document delivered systems so the automation you create remains clear, shareable, and maintainable across the team
Requirements
- 5+ years of experience with AWS core services (EC2, ECS/EKS, S3, IAM, VPC, RDS, CloudWatch), including an applied view of how they behave during failures
- Proven ownership of infrastructure as code using a mainstream tool (Terraform, OpenTofu, Pulumi, CloudFormation/CDK, or similar), including authoring and maintaining definitions rather than only consuming existing modules
- Production track record running CI/CD pipelines with established platforms (Jenkins, GitHub Actions, GitLab CI, CircleCI, Azure DevOps, or similar)
- Strong system design fundamentals, including the ability to weigh tradeoffs around failure domains, state, idempotency, and blast radius before implementation
- Advanced troubleshooting ability, using a structured, evidence-based approach to debugging distributed systems and CI/CD pipelines under time pressure
- Solid software engineering discipline applied to infrastructure and tooling (testing, version control, modularity, code review, readable code), with proficiency in Python, Go, or a comparable language
- Practical experience using coding agents for meaningful implementation work, plus the judgment to validate and review agent output to the same bar as your own code
- Developer-as-customer mindset, with the judgment to assess tooling changes objectively and the discipline to migrate in well-communicated phases instead of disruptive cutovers
Nice to have
- Direct experience with Terraform, Jenkins, and SonarQube, including shared libraries and reusable pipeline patterns
- Familiarity with static analysis and code-quality gating tools such as Snyk, Semgrep, Checkmarx, or CodeQL
- Experience running containers and orchestration in production, including Docker and Kubernetes/EKS
- Knowledge of observability tooling (CloudWatch, Datadog, Prometheus/Grafana, OpenTelemetry) and experience creating alerting that teams trust
- Experience with security and compliance automation, including secrets management, policy as code, and SOC 2 or similar audit evidence collection
- Experience building internal developer platforms or self-service tooling
- Background in security or vulnerability management products
