Likeremote

Subscribe to the latest remote jobs:

  • Likeremote jobs on https://LinkedIn.com/
  • Likeremote jobs on https://telegram.org/
  • Likeremote jobs on Reddit.com
DL

AWS Cloud Operations

Diverse Lynx India
Location not stated
1 week ago
  • Azure
  • AWS
  • Dynatrace
  • Linux
  • Incident Management
  • Python
  • SQL
  • Ansible
  • OpenTelemetry
  • CI/CD
  • IaC
  • Incident Response
Not scoredNo CV on file. Upload one and this job gets a score out of 100.Upload CV
Primary Skills Azure, AWS, Dynatrace, Linux, Production Support, Incident Management & Troubleshooting
Secondary Skills Python/Shell Scripting, SQL, Automation, Infrastructure Engineering, Ansible
Preferred Skills SRE Practices, OpenTelemetry, CI/CD, Infrastructure-as-Code, Middleware Technologies
Key Responsibilities • Provide operational support for business-critical Quartz platform services across on-premises and cloud environments
• Monitor, troubleshoot, and resolve complex production issues using Dynatrace and cloud-native observability capabilities
• Support Azure and AWS-hosted workloads, including deployment validation, monitoring, incident response, and operational readiness
• Drive root cause analysis and remediation activities to improve platform stability and reduce recurring incidents
• Identify and implement opportunities for automation, self-healing, and operational efficiency improvements
• Support cloud migration initiatives, platform modernization efforts, and data center transformation programs
• Collaborate with engineering, infrastructure, and application teams to resolve cross-functional issues and improve service resilience
• Participate in patching, change, release, weekend maintenance, and recovery exercises as required
Required Experience • Hands-on experience supporting workloads inAzure and AWS environments
• Strong experience withDynatrace dashboards, alerting, troubleshooting, and observability concepts
• Strong Linux administration and troubleshooting skills
• Experience supporting large-scale production environments
• Strong incident management and root cause analysis skills
• Experience troubleshooting across infrastructure, middleware, application, and cloud layers
• Working knowledge of networking, compute, storage, and platform services
• Experience with scripting and automation using Python, Shell, or similar technologies
Desired Experience • Hybrid cloud operations and cloud migration experience
• SRE, Reliability Engineering, and Operational Excellence practices
• OpenTelemetry and modern observability platforms
• Automation frameworks and Infrastructure-as-Code concepts
• Financial Services / Trading Platform experience
Key Attributes • Strongsolution-oriented mindset focused on eliminating root causes rather than repeatedly addressing symptoms
• Demonstrated ability to identify

AWS Cloud Operations · Diverse Lynx India

Auto apply with Likeremote