Vice President · SRE & Platform Engineering

Engineering Reliable, Scalable Enterprise Platforms for the AI Era.

Combining over two decades of enterprise engineering leadership in Site Reliability Engineering, Platform Engineering, Application Modernization, Enterprise Operations, and FinOps with hands-on AI to build resilient, scalable, and future-ready enterprise platforms.

📍 Hyderabad, India · JPMorgan Chase · Deutsche Bank · UBS · RBS · Infosys

Executive Summary

Two decades at the operational core of global banking systems.

With over two decades in enterprise technology, I've had the privilege of leading engineering teams and delivering resilient platforms for some of the world's leading financial institutions, including JPMorgan Chase, Deutsche Bank, UBS, and RBS.

My expertise spans Site Reliability Engineering, Platform Engineering, Observability, IT Application Operations, Cloud, and FinOps. I've led globally distributed teams, supported large-scale production environments, and driven initiatives focused on reliability, operational excellence, cloud modernization, and cost optimization.

Beyond my professional experience, I'm passionate about continuous learning. I regularly explore AI, cloud technologies, automation, and platform engineering through certifications, hands-on projects, technical writing, and reading. I believe the best engineering leaders remain curious, embrace new technologies, and never stop learning.

Core Domain SRE · Platform Engineering · Cloud Architecture · Application Modernization · AI · ITSM
Industry Sector BFSI / Enterprise Global Banking
Executive Education IIM Ahmedabad — Senior Management Programme · B.E. in Computer Science & Engineering

Active Technical Initiatives

Strategic Investments & Growth

A career transition is a deliberate investment in depth, not a gap. Since Jan 2026, I've treated this as structured, full-time work: certifications, hands-on projects, technical writing, and wide reading — all logged below as it happens.

20+ Certs / Courses
06+ Articles Published
12+ Executive Books Read
01+ Projects Shipped

Leadership Experience

Career Track Record

  1. VP Level

    Vice President — Site Reliability Engineering

    JPMorgan Chase • Hyderabad, India

    Directed Site Reliability Engineering (SRE) for enterprise observability platforms across APAC and EMEA, leading the operations, strategy, and engineering of large-scale enterprise level Splunk, ELK/EFK, and Kafka based streaming platforms processing petabytes of telemetry data.

    • Enterprise Platform Ownership: Directed end-to-end Site Reliability Engineering (SRE), and operations for enterprise-scale Splunk, ELK/EFK, and Kafka-based data streaming platforms processing petabytes of events / telemetry data. Delivered highly available, secure, scalable, and resilient services—including centralized logging, monitoring, alerting, dashboards, intelligent reporting and operational analytics—supporting thousands of mission-critical banking applications while driving operational excellence, platform reliability, and an exceptional user experience.
    • SRE Culture & Frameworks: Championed SRE methodologies by establishing SLIs, SLOs, and error budgets, systematically eliminating toil and maximizing platform uptime, stability and performance. Reduced incident ticket volume by ~30% while significantly accelerating Mean Time to Detect (MTTD) and Mean Time to Resolve (MTTR).
    • Cloud & Resilience Modernization: Modernizing Splunk on-premises platforms to AWS utilizing Infrastructure as Code (Terraform), Game Days, Chaos Engineering (Gremlin), and automated disaster recovery workflows.
    • ITSM & Governance: Drove operational excellence across incident, problem, change, blameless postmortems and capacity management while ensuring regulatory compliance, audit remediation, and FinOps cloud cost optimization.
    • AI Innovation & Engineering Leadership: Hands-on experience on AI-driven operational tools, including alert correlation algorithms and self-service support chatbots. Remained hands-on with Python automation, custom observability dashboards, and architectural reviews.
    • Forming Agile team : Spearheaded the formation and Agile transformation of a dedicated SRE team, navigating all stages of group development (Forming, Storming, Norming, Performing). Acted as de facto Scrum Master to establish sprint cadence, facilitate core Agile ceremonies, and foster a delivery-focused engineering culture.
  2. AVP Level

    Assistant Vice President — SRE & ITAO

    Deutsche Bank • Pune, India

    Spent over seven years leading Site Reliability Engineering (SRE), IT Application Ownership (ITAO), and Platform Engineering for mission-critical global banking platforms, spanning cloud-native SaaS applications, anti-financial crime systems, payments, sanctions screening, and enterprise integration services. Drove reliability engineering, cloud transformation, security, operational excellence, and large-scale modernization initiatives across both Google Cloud Platform (GCP) and on-premises environments.

    • Cloud-Native Platform Engineering: Led SRE and ITAO for a GCP-hosted, API-first, microservices-based SaaS platform, delivering enterprise-grade reliability, resilience, security, and scalability while integrating with APIGEE, NPCI (National Payments Corporation of India), Razorpay, Digio, and other strategic fintech ecosystems.
    • Reliability Engineering & Operational Excellence: Established SLI/SLO frameworks, error budgets, proactive observability, incident management, postmortems, and automation strategies to improve service reliability, eliminate operational toil, and accelerate engineering delivery.
    • Enterprise Architecture, Security & Governance: Architected resilient cloud environments using Terraform, Shared VPCs, IAM, network segmentation, and automated Disaster Recovery while leading ITAO governance across audit readiness, regulatory compliance, vulnerability remediation, release certification, and secure software delivery.
    • Production Engineering & Platform Modernization: Directed production management for high-volume sanctions screening and case management platforms, delivering performance optimization, CI/CD modernization, infrastructure upgrades, zero-downtime data center migrations, capacity planning, and regional disaster recovery programs.
    • Strategic Technology Leadership: Partnered with engineering, product, security, and business stakeholders to improve platform architecture, accelerate time-to-market, enhance operational resilience, resolve systemic production issues, and scale enterprise platforms supporting global Anti-Financial Crime (AFC) operations.
  3. Technical Leadership

    Technology Lead / Delivery Lead

    Infosys • Bengaluru, London, Zurich, Singapore

    Spent nearly a decade delivering production engineering, middleware integration, and enterprise platform support for global tier-1 financial institutions, including UBS, RBS, and Capital Group. Worked across Bengaluru, London, Zurich, and Singapore, leading mission-critical financial messaging, (payment, FX and security), financial gateways and payments routing / integration platforms while driving operational excellence, platform resilience, and engineering transformation.

    • Enterprise Financial Messaging & Platform Ownership: Led ITSM streams (Incient, Change, Transition and Relase management) for mission-critical SWIFT messaging, payment processing, securities settlement, and sanctions screening platforms, ensuring high availability, regulatory compliance, and uninterrupted global banking operations.
    • Middleware integration & Subject Matter Expertise: Served as the technical lead and Subject Matter Expert (SME) for enterprise middleware technologies including SWIFTNet (SAA, SAG, SNL), FircoSoft, IBM MQ, Apache Tomcat, and related integration platforms, providing architectural guidance, performance tuning, capacity planning, and lifecycle management.
    • Production Engineering & Operational Excellence: Directed incident management, root cause analysis, performance optimization, proactive monitoring, automation, and problem management, significantly improving platform stability, operational efficiency, and service availability across globally distributed production environments.
    • Resilience, Disaster Recovery & Service Governance: Led disaster recovery planning, infrastructure migrations, regulatory audits, change governance, release management, and ITIL-based service management practices, strengthening business continuity, operational resilience, and enterprise compliance.
    • Global Stakeholder Leadership & Technical Consulting: Collaborated with engineering teams, infrastructure groups, vendors, and business stakeholders across Europe and Asia to deliver platform modernization initiatives, mentor support teams, resolve complex production challenges, and continuously enhance enterprise service delivery.
  4. Software Consultant

    Software Developer

    Netsys Consulting • Bengaluru, India

    Engineered enterprise J2EE applications supporting e-commerce and HR platforms for over two years, building a strong core foundation in application architecture, Object-Oriented Design, and backend solution engineering.

  5. Faculty

    Teaching Staff / Computer Science Lecturer

    Scholar Academy of Computer Education • Kalaburagi, India

    Initiated career teaching fundamental computer science and programming concepts. Established early competencies in technical communication, mentoring, and leadership.

Technical Expertise

Technical Skills & Toolsets

Languages / Development

Python Bash / Shell Core Java REST / FastAPI

Core Capabilities

Hands-on experience building automation, operational tooling, REST APIs, and applying AI where appropriate to improve platform reliability, engineering productivity, and operational excellence.

Cloud & Infrastructure

AWS GCP Terraform

Core Capabilities

Hands-on experience architecting, modernizing, and securing enterprise cloud platforms, delivering cloud migrations, Infrastructure as Code (IaC), observability, capacity planning, and cloud cost optimization.

Observability & Monitoring

Splunk Dynatrace Geneos Prometheus Grafana ELK / EFK

Core Capabilities

Hands-on experience designing and implementing enterprise observability solutions, including end-to-end monitoring, centralized logging, dashboards, intelligent alerting, performance monitoring, troubleshooting, and operational analytics.

AI / Emerging Technologies

AWS Bedrock SageMaker LLMs / RAG GitHub Copilot Prompt Eng.

Core Capabilities

Hands-on experience applying Generative AI technologies, including AWS Bedrock, OpenAI, LLMs, RAG, LangChain, Prompt Engineering, and GitHub Copilot to improve developer productivity, automate knowledge discovery, and build practical enterprise solutions.

Messaging / Middleware

Kafka Apache Tomcat ActiveMQ IBM MQ

Core Capabilities

Hands-on experience administering and supporting enterprise messaging and middleware platforms, delivering reliable message processing, large-scale data and event streaming, performance optimization, monitoring, and automation.

Database & Reporting

Oracle PostgreSQL AWS Aurora Jasper Reports

Core Capabilities

Hands-on experience with Oracle, PostgreSQL, Amazon Aurora, and Jasper Reports, delivering database performance optimization through SQL tuning, Oracle AWR analysis, capacity planning, database migration, and enterprise reporting solutions.

DR & Chaos Engineering

DR Architect AWS FIS Gremlin Game Days

Core Capabilities

Hands-on experience architecting Disaster Recovery (DR) and Business Continuity (BCP) solutions across on-premises, hybrid, and cloud environments, including Game Days, RTO/RPO planning, resilience testing, and Chaos Engineering using AWS FIS and Gremlin.

Commercial Financial products

SWIFTNet HotScan FircoSoft Case Mgmt

Core Capabilities

Hands-on experience implementing and managing enterprise banking and observability platforms, including financial messaging and routing, payment/security gateways, sanctions and embargo screening, AML/AFC compliance, and enterprise observability solutions.

ITSM & Collaboration

ServiceNow Jira Agile / Scrum Kanban Confluence

Core Capabilities

Hands-on experience applying ITSM and Agile best practices using ServiceNow, Jira, Confluence, Scrum, and Kanban to improve engineering delivery, operational governance, cross-functional collaboration, and product owner role.

Thought Leadership

Published Articles & Insights

20-05-2026

Chaos Engineering

A Practitioner's Guide for SRE Teams

Article Highlight

Strategies for embedding proactive failure injection into modern CI/CD delivery pipelines to uncover hidden failure modes.

Read Full Article ↗
14-05-2026

DR Testing — Conducting the Test

Resilience Before Disaster Strikes

Article Highlight

How to organize high-stakes, multi-region disaster recovery Game Days with zero client impact in banking environments.

Read Full Article ↗
06-05-2026

Disaster Recovery Strategies

Which One Does Your Business Need?

Article Highlight

A comparative cost-benefit evaluation of Pilot Light, Warm Standby, and Active-Active multi-cloud DR patterns.

Read Full Article ↗
06-05-2026

A Leader's Guide to SLIs, SLOs & Error Budgets

Data-Driven Reliability Frameworks

Article Highlight

Translating technical SLOs into business value metrics to balance feature speed with platform resilience.

Read Full Article ↗
24-03-2026

Building an SRE Function in a GCC

From Legacy Support to Reliability Engineering

Article Highlight

A strategic blueprint for transforming Global Capability Centers from L1/L2 operational hubs into high-impact engineering centers.

Read Full Article ↗
13-02-2026

Gen-AI Real-World Use Cases

From Experimentation to Production

Article Highlight

Tactical architectural blueprints for deploying secure RAG systems and LLM workflows in strict enterprise IT settings.

Read Full Article ↗

Credentials

Certifications