VP, Site Reliability Engineer, Data Platform, Group Technology
- Core Data
- Data Visualization
- Trino
- Tableau
- OpenShift
- GCP
- AWS
- CI/CD
- Prometheus
- Grafana
- Devops
- Hadoop
- Kubernetes
- Elastic Stack
- IaC
Role Overview
We are seeking an experienced and strategic Vice President of Site Reliability Engineering (SRE) to join our Data & Visualization Platform team. The ideal candidate will be a seasoned technical leader with over 16 years of IT experience and deep specialization in Big Data, Monitoring, Visualization, and Observability platforms. You will be responsible for the architecture, operational excellence, and scalability of our critical Data and Visualization infrastructure, ensuring that our platforms—including sophisticated visualization suites—remain resilient, secure, and cost-efficient to support the bank’s strategic data initiatives.
Key Responsibilities
Foundational Platform Development & Engineering
Architect, build, and operationalize enterprise-grade Big Data platforms from the ground up, ensuring high scalability and reliability.
Manage the lifecycle and adoption of core data and Data Visualization tools (e.g., Trino, Superset, Tableau, Celerity) to ensure seamless usability for Lines of Business (LOBTs) and business stakeholders.
Maintain and enhance data lake and data processing infrastructures to support mission-critical banking applications.
Strategic Migration & Modernization
Lead large-scale application migration and modernization projects, transitioning legacy systems to containerized, cloud-native environments (OpenShift, GCP, AWS).
Streamline deployment pipelines and CI/CD workflows to improve engineering velocity and operational efficiency.
Operational Excellence & Observability
Implement and manage unified monitoring, logging, and visualization frameworks (e.g., ELK, Prometheus, Grafana) to provide end-to-end system visibility.
Proactively lead capacity planning and drive cost-optimization strategies, ensuring efficient resource utilization across the cloud ecosystem.
Leadership & Cross-Functional Collaboration
Partner with cross-functional stakeholders to bridge the gap between complex technical solutions and tangible business outcomes.
Oversee production support, ensuring strict adherence to compliance, security standards, and governance protocols.
Act as a senior technical mentor, driving best practices in SRE and DevOps across the platform.
Technical Qualifications
Big Data Ecosystem: Expert-level proficiency in Hadoop, Spark, and Kafka.
Data Visualization & Analytics: Expert-level proficiency in Trino, Superset, Tableau, Celerity, and associated data reporting tools.
Performance Engineering: Expert skills in analyzing, tuning, and optimizing system performance for high-throughput data platforms.
Cloud & Orchestration: Strong experience with Kubernetes, OpenShift, GCP, and AWS.
Observability: Hands-on experience with ELK Stack, Prometheus, and Grafana.
DevOps Practices: Proven track record in automation, pipeline development, and infrastructure-as-code.
Professional Experience
16+ years of IT experience, with at least 9+ years specifically in Big Data, Monitoring, and Observability.
Proven experience in leading large-scale data platform migrations and modernization initiatives in a highly regulated environment (Financial Services preferred).
Demonstrated success in managing complex production environments and driving significant cost-optimization initiatives.
Strong ability to align technical roadmaps with business strategic objectives.
Location:
DBS Asia HubJob:
TechnologySchedule:
RegularEmployee Status:
Full timeVP, Site Reliability Engineer, Data Platform, Group Technology · Dbs