Senior SRE (Native English)
Exciting opportunity for a Native English Senior Site Reliability Engineer to join a pioneering AI startup based in Greater Barcelona. This role is perfect for those who are passionate about building robust backend systems and thrive at the intersection of artificial intelligence and cybersecurity.
Are you ready to take on a critical role in advancing secure AI adoption at scale? As a Senior Site Reliability Engineer, you will be at the forefront of designing and maintaining robust cloud infrastructure that powers enterprise-grade AI solutions. This is an opportunity to tackle complex challenges, drive innovation, and make a tangible impact on the reliability and scalability of cutting-edge platforms.
In this role, you’ll architect resilient systems, automate deployment pipelines, and ensure platform stability through proactive monitoring and rapid incident response. You’ll collaborate closely with developers to streamline release cycles and contribute to defining operational standards that set the bar for excellence. By staying ahead of emerging technologies, you’ll help shape the future of our infrastructure while fostering a culture of continuous learning within the team.
Key Responsibilities:
- Design and maintain highly available, scalable, and secure cloud infrastructure tailored for enterprise AI applications.
- Automate deployment processes, CI/CD pipelines, and monitoring systems to ensure seamless operations across environments.
- Proactively monitor platform performance, ensuring optimal uptime and swift resolution of incidents to minimise disruption.
- Lead troubleshooting efforts for complex production issues while driving effective incident management practices.
- Optimise infrastructure for cost efficiency, security, and observability using modern best practices.
- Collaborate with software developers to enable smooth release cycles and ensure production readiness for new features.
- Enhance internal tooling, documentation, and operational workflows to improve team efficiency and knowledge sharing.
- Stay informed about advancements in cloud-native technologies and bring innovative ideas to elevate team capabilities.
What You Bring:
To thrive in this role, you’ll need a strong technical foundation combined with a collaborative mindset. Your expertise in managing large-scale cloud infrastructures will be complemented by your ability to embed security into every layer of design. Beyond technical skills, your commitment to teamwork and knowledge sharing will help foster a supportive environment where everyone can excel.
Required Skills & Experience:
- Proven experience managing production environments on major cloud platforms such as AWS, GCP (preferred), or Azure.
- Deep understanding of Linux systems administration, networking concepts, and distributed system architectures.
- Hands-on expertise with Infrastructure-as-Code tools like Terraform or similar technologies.
- Proficiency in Kubernetes orchestration for containerised workloads using tools like Helm or Kubectl.
- Familiarity with modern monitoring frameworks such as Prometheus and Grafana for observability.
- Strong scripting or automation skills using Python, Bash, or similar languages to streamline operations.
- Solid grasp of CI/CD pipelines and GitOps workflows for efficient code delivery.
- A security-first approach with practical experience implementing infrastructure security best practices.
- Native English or, at least, a fluent enough level of English to keep technical conversations with UK-based clients.
Preferred Attributes:
- A passion for solving complex problems in fast-evolving domains like AI or cybersecurity.
- A collaborative attitude focused on supporting teammates through shared learning and collective problem-solving.
This is more than just a technical role—it’s an opportunity to influence how organisations adopt secure AI at scale while working alongside talented peers who share your drive for innovation and excellence. If you’re ready to take on this challenge, we’d love to hear from you!
Sobre la posición
Tipo de contrato: Perm
Especialización: IT & Telecomunicaciones
Área: Infraestructura y sistemas
Sector: Tecnología de la información
Banda salarial: Competitive Salary
Tipo de trabajo: Híbrido
Nivel de experiencia: Mando intermedio
Idioma principal: Inglés - Bilingüe
Idioma secundario: Español - Trabajo profesional
Ubicación: Barcelona
FULL_TIMEReferencia: AP0LZH-C6C36BE3
Fecha de publicación: 10 de julio de 2026
Consultor/a: Luís Cespedes
barcelona information-technology/infrastructure-and-systems 2026-07-10 2026-09-08 it Barcelona Barcelona Barcelona ES 08019 Robert Walters https://www.robertwalters.es https://www.robertwalters.es/content/dam/robert-walters/global/images/logos/web-logos/square-logo.png true