Technology Operations Specialist, VP
Deutsche Bank
Job at a glance
Job Description: Job Title: Technology Operations Specialist Corporate Title: VP Location: Pune, India Role Description You will be operating within the Production Support Services team of the TST domain, spanning SES, TAS and Trade Finance Lending within Corporate Bank Production Services. The role is focused on providing hands-on production support, troubleshooting and operational ownership for applications hosted on Google Cloud Platform (GCP), with strong emphasis on Google Kubernetes Engine (GKE), containerised workloads, PostgreSQL/Postgres-backed services and related cloud-native technologies.
You will be expected to lead incident recovery, perform deep technical analysis across application, infrastructure, database and platform layers, improve observability, and drive automation to enhance service stability, resilience and operational efficiency. Lead technical and functional troubleshooting for production incidents, user requests and platform issues across GCP-hosted applications, including GKE, Cloud Run, Compute Engine, Cloud Storage, IAM, networking, PostgreSQL/Postgres databases and database connectivity.
Drive end-to-end incident recovery for P3H and above incidents by analysing logs, metrics, traces, deployment history, configuration changes, infrastructure events and application behaviour. Use GCP observability tools such as Cloud Logging, Cloud Monitoring, Error Reporting, dashboards, uptime checks and alert policies to identify root cause, reduce mean time to detect and improve service reliability.
Partner with application, infrastructure, SRE, security and engineering teams to troubleshoot cloud networking, container runtime, IAM, quota, performance, scaling and availability issues. Drive automation, toil reduction, platform hygiene, monitoring improvements and operational controls for cloud-native applications and supporting infrastructure. Prepare and distribute clear incident communications, technical updates, recovery timelines and service restoration summaries for senior stakeholders.
Participate in CAB, change validation and post-change support activities to assess operational risk, ensure rollback readiness and support safer change delivery. Collaborate with Global Incident Management, L2/L3 support, SRE and platform teams to orchestrate recovery of major incidents and implement preventive actions. What we’ll offer you As part of our flexible scheme, here are just some of the benefits that you’ll enjoy, Best in class leave policy.
Gender neutral parental leaves 100% reimbursement under childcare assistance benefit (gender neutral) Sponsorship for Industry relevant certifications and education Employee Assistance Program for you and your family members Comprehensive Hospitalization Insurance for you and your dependents Accident and Term life Insurance Complementary Health screening for 35 yrs. and above Your key responsibilities Provide hands-on technical and functional production support for applications deployed on GCP and integrated enterprise platforms within the TST domain.
Troubleshoot complex production issues across cloud runtime, containers, application services, middleware, databases, network connectivity, IAM permissions, certificates, secrets and deployment pipelines. Analyse GCP logs, metrics and alerts using Cloud Logging, Cloud Monitoring, dashboards and log-based metrics to identify root cause and restore service quickly. Support containerised workloads running on GKE and Cloud Run, including pod/container restarts, scaling behaviour, node pressure, cluster events, health checks, readiness/liveness failures, resource saturation, ingress/service issues and deployment rollbacks.
Troubleshoot PostgreSQL/Postgres production issues, including connection failures, query performance, locks, replication or failover symptoms, storage growth, backup/restore readiness and application-to-database connectivity. Build technical and functional subject matter expertise across supported applications, b