AI
Lead GCP / DevOps Engineer
Ampcus Incorporated
๐บ๐ธ United States
On-site
Staff / Principal
3 days ago
- GCP
- Devops
- CI/CD
- IaC
- Terraform
- GKE
- Compute Engine
- BigQuery
- Cloud SQL
- GitHub Actions
- GitLab CI
- Jenkins
- Docker
- Python
- Bash
- Prometheus
- Grafana
- Datadog
- Incident Response
- Kubernetes
- AWS
- Azure
- Istio
3 days ago
Ampcus Inc. is a certified global provider of a broad range of Technology and Business consulting services. We are in search of a highly motivated candidate to join our talented Team.
CTH/FTE
Austin/Southlake, TX
Detailed Job Description
We are seeking a highly skilled and visionary Lead GCP SRE / DevOps Engineer to spearhead the reliability, scalability, and automation of our cloud infrastructure. In this role, you will bridge the gap between development and operations, leading a team to design and maintain robust CI/CD pipelines, optimize Google Cloud Platform (GCP) infrastructure, and enforce high availability across all production environments.
Key Responsibilities
Infrastructure & Architecture
- Design and maintain scalable, secure, and fault-tolerant infrastructure entirely within Google Cloud Platform (GCP).
- Architect Infrastructure as Code (IaC) templates using Terraform to manage multi-environment setups.
- Optimize cloud spend and maximize performance across GCP services (GKE, Compute Engine, BigQuery, Cloud SQL).
CI/CD & Automation
- Own the deployment lifecycle by building and optimizing automated CI/CD pipelines (using tools like GitHub Actions, GitLab CI, or Jenkins).
- Drive containerization strategies using Docker and orchestration via Google Kubernetes Engine (GKE).
- Automate repetitive operational tasks ("toil") using scripting languages like Python, Go, or Bash.
Observability & SRE Practices
- Define, measure, and report critical reliability metrics, including SLIs, SLOs, and Error Budgets.
- Implement comprehensive monitoring, logging, and alerting systems using GCP Cloud Monitoring/Logging, Prometheus, Grafana, or Datadog.
- Lead incident response rotations and facilitate constructive, blameless post-mortems to prevent recurrence.
Required Skills & Qualifications
Technical Essentials
- Experience: 8 years of experience in DevOps, Systems Engineering, or SRE roles, in GCP Platform.
- Cloud Platform: Deep production-level expertise with Google Cloud Platform (GCP).
- Orchestration: Advanced hands-on experience managing production workloads on Kubernetes (specifically GKE).
- Infrastructure as Code: Proficient with Terraform for state management and modular design.
- Programming: Strong coding skills in Python or Go, alongside excellent shell scripting capabilities.
- CI/CD: Proven track record building enterprise-grade pipelines.
- Experience migrating legacy workloads from on-premise or AWS/Azure into GCP.
Soft Skills & Leadership
- Strong communication skills with the ability to explain complex architectural concepts to non-technical stakeholders.
- Proven ability to manage high-pressure live production incidents calmly and methodically.
Preferred Qualifications (Nice to Have)
- Certifications: GCP Certified Professional Cloud Architect or GCP Certified Professional Cloud DevOps Engineer.
- Familiarity with service mesh technologies (e.g., Istio, Anthos Service Mesh).
Lead GCP / DevOps Engineer ยท Ampcus Incorporated