M
Site Reliability Engineering (SRE)/Dev Ops - Site Reliability Engineering (SRE)/Dev Ops
Mindlance
🇨🇴 Colombia
On-site
4 weeks ago
- AWS
- CloudWatch
- DynamoDB
- Step Functions
- Aurora
- MySQL
- RDS
- Agile
- Change Management
- Incident Management
4 weeks ago
Location: Colombia – Remote
Duration: 12 months
Â
Key Responsibilities:
- Monitor and troubleshoot production applications and AWS infrastructure.
- Investigate and resolve production incidents using AWS CloudWatch, Lambda, DynamoDB, Step Functions, Aurora MySQL/RDS, and related tools.
- Support L1/L2 escalations and collaborate with Developers and Data Management teams.
- Assist with application releases through testing, validation, and phased deployments.
- Follow Agile SDLC and Change Management processes.
- Create and maintain production documentation, processes, and basic automations.
- Provide support for urgent production issues, including occasional after-hours coverage.
- Basic hands-on experience with AWS and cloud-based applications.
- Knowledge of CloudWatch, Lambda, DynamoDB, Step Functions, and RDS/Aurora MySQL.
- Strong troubleshooting and problem-solving skills.
- Understanding of production/application support and incident management.
- Good communication and ability to work with cross-functional teams.
- Bachelor's degree in Computer Science, IT, Engineering, or related field preferred.
“Mindlance is an Equal Opportunity Employer and does not discriminate in employment on the basis of – Minority/Gender/Disability/Religion/LGBTQI/Age/Veterans.”
Site Reliability Engineering (SRE)/Dev Ops - Site Reliability Engineering (SRE)/Dev Ops · Mindlance