Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
PL

AWS L3 SME / Cloud Architect

PeopleNTech LLC
🇺🇸 United States
On-site
Manager or above
3 days ago
  • AWS
  • GCP
  • Kubernetes
  • Devops
  • Incident Management
  • EC2
  • VPC
  • IAM
  • RDS
  • CloudWatch
  • Compute Engine
  • Cloud SQL
  • Terraform
  • CloudFormation
  • Ansible
  • Python
  • Docker
  • EKS
  • GKE
  • Grafana
  • Prometheus
  • AWS Security Specialty
  • CKA
  • AWS Cloud
  • EBS
  • SNS
  • SQS
  • ACLS
  • DNS
  • KMS
  • SOC2
  • ISO 27001
  • HIPAA
  • NIST
  • IaC
  • Configuration Management
  • Change Management
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV

PSL304306-8

Warren NJ

Onsite

AWS L3 SME / Cloud Architect

80-85/hr

JD :

15-20 years exp.

Job Title

AWS L3 SME / Cloud Specialist (AWS Primary, GCP Secondary)

AWS | GCP | Cloud Operations | Infrastructure | Security | Automation | Kubernetes | DevOps

Role Summary

We are seeking an experiencedAWS L3 SME / Cloud Specialist to provide advanced operational support, engineering expertise, and platform optimization for enterprise cloud environments. The role will be primarily responsible for managing, troubleshooting, and enhancing AWS infrastructure while supporting workloads deployed on GCP.

The ideal candidate should possess strong hands-on expertise in AWS services, cloud operations, incident management, automation, security, networking, and cloud governance. The individual will act as the highest level technical escalation point for complex cloud incidents and contribute to continuous improvement initiatives.

Required Skills & Experience

Experience

·6-10 years of overall IT infrastructure and cloud experience.

·4+ years hands-on AWS experience.

  • Experience supporting production cloud environments.
  • Strong background in incident management and operational support.
  • Experience in enterprise or regulated environments preferred.

Technical Skills

AWS (Mandatory)

· EC2

  • VPC
  • S3
  • IAM
  • RDS
  • Route53
  • CloudWatch
  • Lambda
  • Auto Scaling
  • Load Balancers
  • Systems Manager (SSM)

GCP (Good to Have)

· Compute Engine

  • Cloud Storage
  • Cloud SQL
  • IAM
  • VPC

Automation

· Terraform

  • CloudFormation
  • Ansible
  • Python
  • Shell Scripting

Containers

· Kubernetes

  • Docker
  • EKS
  • GKE

Monitoring

· Grafana

  • Prometheus
  • CloudWatch
  • Logging & Alerting Platforms

Leadership & Soft Skills

· Strong troubleshooting and analytical capabilities

  • Excellent communication and stakeholder management skills
  • Ability to lead technical incident bridges
  • Documentation and knowledge-sharing mindset
  • Mentoring and coaching capability

Preferred Certifications

· AWS Certified Solutions Architect Associate / Professional

  • AWS Certified SysOps Administrator
  • AWS Advanced Networking Specialty (preferred)
  • AWS Security Specialty (preferred)
  • Google Associate Cloud Engineer (good to have)
  • Terraform Associate Certification
  • Kubernetes (CKA/CKAD) certification

Role Impact

This role is critical to ensuring:

High availability and reliability of cloud platforms

Fast resolution of business-critical incidents

  • Cloud security and compliance adherence
  • Infrastructure automation and operational efficiency
  • Optimized cloud performance and cost management

Key Responsibilities

1. AWS Platform Operations & Engineering

· Serve as the L3 technical SME for AWS cloud services.

  • Manage and support AWS environments including:

o EC2, EBS, S3, RDS, EFS

    • VPC, Load Balancers, Route53
    • Lambda, CloudWatch, SNS, SQS
  • Troubleshoot critical production incidents and perform root cause analysis.
  • Drive platform stability, reliability, and performance improvements.
  • Support cloud migrations and modernization activities.

2. GCP Knowledge & Support (Basic knowledge is enough)

· Provide operational support for GCP services.

  • Assist in administration of:

o Compute Engine

    • Cloud Storage
    • Cloud SQL
    • VPC Networking
    • IAM
  • Collaborate with GCP SMEs on issue resolution and optimization initiatives.
  • Support multi-cloud connectivity between AWS and GCP.

3. Cloud Networking

· Manage and troubleshoot:

o VPCs

    • Subnets
    • Route Tables
    • Security Groups
    • Network ACLs
    • VPN Connectivity
  • Resolve routing, DNS, latency, and connectivity issues.
  • Support hybrid connectivity with on-premises environments.
  • Implement secure network segmentation and access controls.

4. Security & Compliance

· Manage IAM users, roles, policies, and federation.

  • Implement and maintain security best practices.
  • Support vulnerability remediation and compliance initiatives.
  • Manage encryption services including KMS and Secrets Manager.
  • Assist with compliance requirements such as SOC2, ISO 27001, HIPAA, and NIST.

5. Automation & Infrastructure as Code

· Develop and maintain infrastructure automation solutions.

  • Work with:

o Terraform

    • CloudFormation
    • Ansible
    • Python
    • Shell Scripting
  • Automate provisioning, configuration management, patching, and operational tasks.
  • Contribute to self-healing and operational efficiency initiatives.

6. Cloud Monitoring

· Implement and support monitoring platforms.

  • Configure and maintain:

o CloudWatch

    • Grafana
    • Prometheus
    • Logging Solutions
  • Create operational dashboards and alerts.
  • Drive proactive monitoring and incident prevention.

7. Kubernetes & Container Platforms

· Support containerized workloads on:

o EKS

    • GKE (basic operational support)
  • Troubleshoot cluster, networking, ingress, and workload issues.
  • Support container security and platform upgrades.

8. Incident, Problem & Change Management

· Act as escalation point for Priority 1 and Priority 2 incidents.

  • Perform detailed root cause analysis and corrective action planning.
  • Participate in change reviews and implementation planning.
  • Drive service improvement initiatives.
  • Prepare post-incident and problem management reports.

9. Cost Optimization & Governance

· Identify opportunities for cloud cost reduction.

  • Support:

o Rightsizing

    • Reserved Instances
    • Savings Plans
    • Storage Optimization
  • Monitor cloud consumption and usage trends.
  • Ensure adherence to tagging and governance standards.

10. Collaboration & Knowledge Management

· Work closely with application, infrastructure, security, and DevOps teams.

  • Create and maintain operational runbooks and technical documentation.
  • Mentor L1 and L2 support engineers.
  • Participate in on-call and major incident support rotations.

AWS L3 SME / Cloud Architect · PeopleNTech LLC

Auto apply with Likeremote