This job is no longer available
This job expired on 15/08/2026. It no longer accepts applications.
Senior Site Reliability Engineer (SRE)
Jobgether
Job description
About the role
We are looking for a Senior Site Reliability Engineer to join a fast‑paced, engineering‑driven team that builds and operates large‑scale cloud infrastructure for AI and cloud‑native platforms. You will ensure that highly distributed systems remain reliable, scalable, and performant under demanding production workloads.
Key responsibilities
- Maintain high system availability through fault tolerance, monitoring, and rapid incident response.
- Design, implement, and optimise scalable infrastructure using modern cloud‑native technologies.
- Improve CI/CD pipelines to enable safe, efficient, and automated software delivery.
- Collaborate with software, infrastructure, and platform teams to troubleshoot complex issues across compute, networking, and storage layers.
- Apply infrastructure‑as‑code practices (Terraform, Ansible, etc.) to manage and standardise environments.
- Support containerised workloads and orchestration platforms such as Docker, Kubernetes, and Helm.
- Contribute to operational best practices, including observability, alerting, and performance tuning.
Required profile
- Strong programming skills in Go, Python, or C++ with solid algorithmic foundations.
- Deep understanding of Unix/Linux systems, networking fundamentals, and distributed system behaviour.
- Hands‑on experience with containerisation and orchestration tools (Docker, Kubernetes).
- Practical experience with infrastructure‑as‑code and configuration‑management tools (Terraform, Ansible, Salt, etc.).
- Familiarity with CI/CD systems and modern DevOps practices.
- Proven experience supporting high‑load distributed systems in production.
- Excellent problem‑solving, communication, and collaboration skills.
Required skills
- Go
- Python
- C++
- Unix/Linux
- Docker
- Kubernetes
- Helm
- Terraform
- Ansible
- Salt
- CI/CD pipelines
- Infrastructure‑as‑code
- Observability and monitoring
- Networking fundamentals
- Distributed systems
What we offer
- Competitive compensation package.
- Career growth and continuous learning opportunities.
- High degree of autonomy, flexibility, and ownership.
- Collaborative, innovation‑focused engineering culture.
- Opportunity to work on large‑scale, impactful cloud and AI infrastructure.
- International environment with highly skilled engineering teams.
Questions fréquentes
Why are you reporting this job?
Explore further
Salaries, guides and searches in Portugal.
Salaries by job title
Boost your chances
Upload your CV — we will match you with relevant openings.
Analyzing your CV...
Jobgether