Senior DevOps / SRE Engineer
- Devops
- CI/CD
- GitHub Actions
- Jenkins
- IaC
- DevSecOps
- Groovy
- Python
- Bash
- GitHub
- Nexus
- SonarQube
- cosign
- Argo
- Jira
- Confluence
- GitOps
- Configuration Management
- Elastic Stack
- Prometheus
- Grafana
- Incident Management
- Incident Response
- Bash scripting
Role Overview
We are looking for an experienced and highly skilledSenior DevOps / SRE Engineer to design, implement, automate, and manage modern DevOps and Site Reliability Engineering practices across enterprise technology environments.
The ideal candidate will have strong hands-on experience inCI/CD, GitHub Actions, Jenkins, container technologies, Infrastructure as Code (IaC), DevSecOps, monitoring, logging, automation, and cloud-native technologies. The role requires someone who can work across development, infrastructure, security, and operations teams to improve deployment efficiency, system reliability, security, scalability, and operational excellence.
The successful candidate will be expected to take ownership of projects frominception through delivery and ongoing operational support, while continuously identifying opportunities to automate and improve existing processes.
Key Responsibilities
DevOps & CI/CD
- Design, implement, maintain, and optimizeCI/CD pipelines for enterprise applications and services.
- Develop and manage deployment pipelines usingJenkins and GitHub Actions.
- Build and maintain Jenkins pipelines using:
- Jenkins Shared Libraries
- Declarative and Scripted Pipelines
- Groovy scripting
- Python and Bash automation
- Implement automated build, test, security scanning, packaging, and deployment processes.
- Develop reusable CI/CD components and standards to improve consistency across development teams.
- Support automated deployments across containerized and cloud-native environments.
- Identify and eliminate manual deployment activities through automation.
- Establish best practices around source control, branching, release management, artifact management, and deployment automation.
DevOps Tooling
Install, configure, administer, upgrade, integrate, and troubleshoot enterprise DevOps tooling, including:
- Jenkins
- GitHub / GitHub Actions
- Nexus / Nexus IQ
- SonarQube
- Checkmarx
- Sysdig
- Cosign
- Argo CD
- JIRA
- Confluence
- Manage integrations between DevOps tools to establish an efficient and secure software delivery lifecycle.
- Monitor tool availability, performance, capacity, and security.
- Troubleshoot issues related to CI/CD tools and their integrations.
- Establish standards for tool configuration, access management, security, and governance.
Containerization & Deployment
- Implement and manage CI/CD pipelines for applications deployed oncontainer technologies.
- Support containerized application build, packaging, deployment, and lifecycle management.
- Work closely with development and infrastructure teams to improve container-based application delivery.
- Implement automated deployment strategies using modern DevOps and GitOps practices.
- SupportArgo CD-based GitOps workflows and ensure reliable application deployments.
- Troubleshoot container, deployment, configuration, and runtime-related issues.
Infrastructure Automation & Configuration Management
- Automate infrastructure provisioning, deployment, and configuration management activities.
- Develop reusable automation to createconsistent, scalable, and repeatable environments.
- Apply Infrastructure as Code principles to infrastructure and application configuration.
- Reduce operational dependency on manual activities through automation.
- Implement configuration management standards and ensure consistency across environments.
- Support infrastructure changes through controlled, automated, and auditable processes.
DevSecOps & Security
- Integrate security controls into CI/CD pipelines followingDevSecOps principles.
- Implement and maintain automated security and quality gates using tools such as:
- SonarQube
- Checkmarx
- Nexus IQ
- Sysdig
- Cosign
- Support vulnerability scanning, code quality analysis, dependency analysis, container security, and artifact verification.
- Implement secure software supply-chain practices, including artifact signing and verification.
- Work with security teams to address vulnerabilities and improve application and infrastructure security.
- Ensure DevOps processes comply with organizational security and governance standards.
Monitoring, Logging & Observability
- Implement and maintain enterprise logging, monitoring, and observability solutions.
- Work with technologies such as:
- ELK Stack
- Prometheus
- Grafana
- Develop dashboards, alerts, and monitoring mechanisms for infrastructure and application environments.
- Monitor system health, application performance, availability, and reliability.
- Analyze logs and metrics to identify performance issues and operational risks.
- Establish proactive monitoring and alerting to reduce incidents and improve service reliability.
- Support root-cause analysis of production incidents using monitoring and observability data.
Incident Management & Troubleshooting
- Troubleshoot and resolve complexinfrastructure, application, deployment, and CI/CD issues.
- Work closely with development, infrastructure, security, and operations teams to resolve production and non-production issues.
- Participate in incident management, problem management, and root-cause analysis.
- Identify recurring issues and implement permanent solutions rather than relying on manual workarounds.
- Support production deployments and provide operational assistance when required.
- Contribute to continuous improvement initiatives based on incident trends and operational feedback.
SRE & Operational Excellence
- ApplySite Reliability Engineering (SRE) principles to improve system availability, scalability, performance, and resilience.
- Implement automation to reduce operational toil.
- Define and improve operational processes, reliability practices, and service standards.
- Support capacity planning, performance optimization, availability, and resilience initiatives.
- Contribute to reliability engineering practices such as monitoring, alerting, incident response, and continuous improvement.
- Balance project delivery responsibilities with operational support requirements.
DevOps & Application Lifecycle
- Work on projects frominception, design, implementation, testing, deployment, and transition to operations.
- Collaborate with application development teams to integrate DevOps practices into the software development lifecycle.
- Provide technical guidance on build, deployment, configuration, automation, and operational requirements.
- Promote automation-first approaches across development and operations teams.
- Support continuous improvement of engineering processes and delivery methodologies.
Required Technical Skills
Mandatory Skills
- Strong hands-on experience withDevOps methodologies and practices.
- Strong practical experience withGitHub Actions.
- Strong hands-on experience withJenkins and CI/CD pipelines.
- Good knowledge ofJenkins Shared Libraries, Groovy, Python, and/or Bash scripting.
- Experience managing and administering enterprise DevOps tooling.
- Strong understanding ofcontainer technologies and container-based deployments.
- Experience withGitOps and Argo CD.
- Strong understanding ofInfrastructure as Code (IaC) principles.
- Strong knowledge ofDevSecOps and security integration within CI/CD pipelines.
- Experience withELK, Prometheus, and Grafana or equivalent monitoring and logging technologies.
- Strong infrastructure and security fundamentals.
- Strong troubleshooting and problem-solving capabilities.
DevOps Tools
Hands-on experience with several of the following:
Jenkins | GitHub | GitHub Actions | Nexus | Nexus IQ | SonarQube | Checkmarx | Sysdig | Cosign | Argo CD | JIRA | Confluence
DevOps / SRE Knowledge
The candidate should have a comprehensive understanding of:
- DevOps principles and methodologies
- SRE principles and practices
- CI/CD and continuous delivery
- Infrastructure as Code
- GitOps
- DevSecOps
- Containerization
- Automation and configuration management
- 12-Factor Application principles
- Monitoring and observability
- Logging and incident management
- Application lifecycle management
- Infrastructure and application security
- Reliability, scalability, and availability engineering
Soft Skills & Behavioral Competencies
- Strong communication skills with the ability to communicate effectively withtechnical and non-technical stakeholders.
- Ability to establish transparent and professional relationships with internal teams, clients, vendors, and key stakeholders.
- Strong ownership and accountability for assigned projects and services.
- Ability to work independently while collaborating effectively within cross-functional teams.
- Strong analytical and pragmatic approach to problem-solving.
- Ability to work effectively under pressure and manage multiple priorities.
- Strong stakeholder and client management skills.
- Ability to manage client expectations while maintaining delivery commitments.
- Strong focus on operational excellence and continuous improvement.
- Willingness to learn and adopt emerging technologies and industry best practices.
- Ability to work across bothproject delivery and operational support functions.
- Ability to adapt to changing business and technology requirements.
- Demonstrate professional integrity, accountability, and a collaborative approach.
- Ability to contribute positively to team culture and promote organizational values.
Key Success Factors
The successful candidate will be expected to:
- Increase deployment automation and reduce manual intervention.
- Improve CI/CD pipeline reliability and efficiency.
- Strengthen security throughout the software delivery lifecycle.
- Improve infrastructure and application reliability.
- Reduce recurring production incidents through automation and root-cause analysis.
- Improve monitoring, logging, and observability.
- Establish scalable and reusable DevOps practices.
- Support faster and more reliable software delivery.
- Build strong relationships with development, infrastructure, security, operations, and business stakeholders.
- Continuously identify opportunities for process, tooling, and technology improvements.
Senior DevOps / SRE Engineer ยท GSSTech Group