Senior Site Reliability Engineer
Software Engineering
Lisbon, Portugal
Posted on Jul 29, 2026
As a Site Reliability Engineer (SRE) at Fountain, you’ll own the reliability, scalability, and operational excellence of the systems that power our platform. You’ll partner closely with Engineering, Security, and Product to build resilient infrastructure, improve developer experience, and raise the bar on observability and incident response. What you'll be doing: Own and improve platform reliability (SLOs/SLIs), capacity planning, and production readiness; design, build, and maintain Kubernetes-based infrastructure and deployment workflows; operate and evolve our AWS footprint; improve CD/GitOps practices (ArgoCD) and deployment safety; build autoscaling strategies (KEDA); lead incident response and postmortems; strengthen observability (OpenTelemetry + dashboards); partner with application teams; improve IaC and maintain Terraform modules; evaluate and integrate AI tooling and MCP tools. What you should bring: 5+ years in SRE; experience owning reliability (SLOs/error budgets); deep hands-on Kubernetes/Helm on EKS; experience with Cloudflare; CI/CD or GitOps; led incident response; working knowledge of AWS core services. Nice to have: AI-assisted ops tooling; OpenTelemetry/observability pipeline design; ArgoCD, KEDA; Terraform.