
Senior Site Reliability Engineer (SRE) (Spain)
- Azure
- Java
- Micronaut
- CI/CD
- Incident Management
- Docker
- Kubernetes
- IaC
- Terraform
- Configuration Management
- Azure Cloud
- AKS
- Spring Boot
- Microservices
- NoSQL
- Couchbase
- Python
- Bash
- Azure DevOps
- Jenkins
- Azure Monitor
- Prometheus
- Grafana
- Elastic Stack
- Splunk
- Health insurance
Senior Site Reliability Engineer (SRE)
We are seeking a highly skilled and passionate Senior Site Reliability Engineer to join ourEngineering Enablement team. This is a critical role within a large, complex, and high-impactinitiative focused on deconstructing our monolithic architecture, revitalising our technologystack, and embedding quality and resilience into every stage of our development lifecycle.
You will play a pivotal role in shaping our future-state platform, driving operationalexcellence, and fostering a culture of continuous improvement.
What You'll Do:
As a Senior SRE Engineer in our Engineering Enablement team, you will:
• Architect and Implement Reliability: Design, build, and maintain highly scalable,resilient, and performant systems on Azure, focusing on our Java, Kafka, andCouchbase stack.
• Drive Modernisation: Work hands-on as part of the team spearheading the adoptionof Micronaut, standardising application templates, and transitioning to managed cloudservices.
• Enhance Operational Excellence: Develop and implement strategies for improvingsystem observability (standardised logging, metrics, tracing), alerting, and on-callpractices.
• Automate Everything: Champion automation across the software developmentlifecycle (SDLC), from CI/CD pipelines to infrastructure provisioning, focusing onaccelerating delivery and de-risking deployments.
• Incident Management & Learning: Contribute to our mature, blameless post- incident review process, identifying root causes and implementing preventativemeasures to reduce incident hours.
• Tooling & Standards: Develop, maintain, and drive the adoption of shared,standardised SRE tooling and best practices across engineering teams, includingcontainerisation (e.g., Docker, Kubernetes on Azure), infrastructure as code (e.g.,Terraform), and configuration management.
• Mentorship & Collaboration: Provide technical leadership and mentorship to juniorengineers, fostering a culture of SRE principles and operational excellence across thewider engineering organisation.
• Strategic Input: Contribute to the overall technical strategy and roadmap for ourSRE and platform initiatives, ensuring alignment with business objectives.
What You'll Bring:
• Deep SRE Expertise: Proven experience as a Senior Site Reliability Engineer or asimilar role, with a strong understanding of SRE principles (error budgets,SLOs/SLIs, toil reduction).
• Azure Cloud Proficiency: Extensive hands-on experience designing, deploying, andoperating highly available and scalable applications on Microsoft Azure.
• Azure Kubernetes Service (AKS) Expertise: Mandatory extensive hands-onexperience with AKS for container orchestration, including deployment, scaling,monitoring, and troubleshooting.
• Java Ecosystem Mastery: Expert-level proficiency with Java, including experiencewith modern frameworks (ideally Micronaut, Spring Boot, or similar) and JVMperformance tuning.
• Distributed Systems Knowledge: Solid understanding and practical experience withdistributed systems, microservices architecture, and associated challenges (e.g.,consistency, fault tolerance).
• Messaging & Database Expertise: Hands-on experience with an event streamingplatform (ideally Kafka) and NoSQL data storage (ideally Couchbase), includingoperational best practices.
• Automation First Mindset: Strong scripting skills (e.g., Python, Bash) andexperience with Infrastructure as Code tools (e.g., Terraform, ARM templates) andCI/CD pipelines (e.g., Azure DevOps, Jenkins).
• Observability Tools: Experience with monitoring, logging, and alerting tools (e.g.,Azure Monitor, Prometheus, Grafana, ELK Stack, Splunk).
• Problem-Solving Acumen: Exceptional analytical and troubleshooting skills, with amethodical approach to diagnosing and resolving complex production issues.
• Communication & Collaboration: Excellent communication skills, with the abilityto articulate complex technical concepts to diverse audiences and collaborateeffectively with cross-functional teams.
• Continuous Improvement: A proactive and innovative mindset, always seekingways to improve systems, processes, and team efficiency.
Some of the benefits you’ll enjoy working with us:
• The chance to join an organization with triple-digit growth that is changing the paradigm on how software products are built.
• The opportunity to form part of an amazing, multicultural community of tech expert
• A highly competitive compensation package.
• Medical insurance.
• English lessons.
Come and join our #ParserCommunity.
Follow us on Linkedin
Senior Site Reliability Engineer (SRE) (Spain) · Parser Limited