Jobiglo

No results.

Senior Site Reliability Engineer (SRE)

Jobgether

Senior 🇬🇧 English
Kubernetes VPC IAM load balancing firewalls WAF CDN DNS Infrastructure-as-code CI/CD GitOps Observability Monitoring Alerting Incident response Disaster recovery Distributed data systems

Job description

About the role

We are looking for a Senior Engineer, SRE to design, build, and operate the infrastructure that powers advanced AI‑driven products. Based in Ireland, you will work across cloud platforms, Kubernetes, networking, and automation to deliver reliable, secure, and scalable systems.

Key responsibilities

  • Design, build, and operate reliable infrastructure for high‑scale AI workloads.
  • Own and improve Kubernetes clusters, including reliability, networking, autoscaling, and deployment patterns.
  • Develop and maintain cloud infrastructure (networking, security, identity, automation).
  • Enhance production reliability through observability, monitoring, alerting, incident response, and disaster recovery.
  • Create infrastructure‑as‑code and automation to reduce manual effort.
  • Improve CI/CD, GitOps, progressive delivery, and deployment safety mechanisms.
  • Optimize infrastructure costs via capacity planning and performance tuning.
  • Operate and enhance distributed data systems such as search and database platforms.
  • Investigate and resolve complex production issues across application, infrastructure, networking, and data layers.

Required profile

  • Strong software engineering background with production‑grade code experience.
  • Proven experience designing, building, and operating infrastructure on major cloud platforms (GCP, AWS, or similar).
  • Hands‑on experience managing Kubernetes environments in production.
  • Deep understanding of cloud networking and security concepts (VPC, IAM, load balancing, firewalls, WAF, CDN, DNS).
  • Experience with infrastructure‑as‑code tools and automated deployment pipelines.
  • Ability to debug complex, multi‑layer issues in large‑scale systems.

Required skills

  • Kubernetes
  • Google Cloud Platform (GCP)
  • Amazon Web Services (AWS)
  • Cloud networking (VPC, IAM, load balancing, firewalls, WAF, CDN, DNS)
  • Infrastructure‑as‑code (e.g., Terraform, CloudFormation)
  • CI/CD and GitOps practices
  • Observability, monitoring, and alerting
  • Incident response and disaster recovery
  • Cost optimization and capacity planning
  • Distributed data systems (search, databases)

Questions fréquentes

Le salaire n'est pas communiqué publiquement par le recruteur. Vous pouvez postuler et négocier directement avec Jobgether.
Cliquez sur "Postuler maintenant" en haut de la page. Vous pouvez importer votre CV en 1 clic — Jobiglo extrait automatiquement vos informations et postule pour vous.

Why are you reporting this job?

Thank you for your report. We will review this job.

Apply in 30 seconds

Enter your email to apply. An account will be created automatically.

By continuing, you accept our terms of use.

Already have an account? Login

💬 Chat with us on Telegram Chat on WhatsApp

Published 1 month ago

Expires 20 hours from now

61 views · 1 interested

Boost your chances

Upload your CV — we will match you with relevant openings.

Analyzing your CV...

Jobgether