whoami

Jay

Site Reliability Engineer II @ Mastercard

Your not-so-average Cloud & DevOps wizard.

jay@portfolio: ~ type help

watch -n1 kubectl top

Live systems dashboard

streaming
Request rate
req/s · last 60s
Latency p50 p99
ms · last 60s
Error budget
remaining
SLOs
  • checkout 99.9% Meeting ✓
  • auth 99.95% Meeting ✓
  • payments-api 99.9% Meeting ✓
  • ledger 99.5% Meeting ✓
0%
Production uptime
0M+
Daily transactions
Live events tail -f /var/log/cluster

while true; do ship; done

How I keep things running

Code
Build
Test
Deploy
Observe

cat about.md

About

I'm a Cloud & DevOps engineer who keeps large-scale systems fast, resilient, and observable. When I'm not moving mountains (or at least, data) to the cloud, I'm tinkering with on-prem setups — someone has to keep those ancient relics running, right? Today I work on reliability, automation, and observability as an SRE at Mastercard.

  • role Site Reliability Engineer II
  • company Mastercard
  • location Dublin, Ireland
  • email jai9176288@gmail.com

ls ~/skills

Tech & Tooling

Cloud

AzureAWSCloud ArchitectureAKSEC2

Infrastructure as Code

TerraformAnsibleInfrastructure as Code

Containers & Orchestration

KubernetesDockerArgo CDGitOpsPod Security

CI/CD & Automation

GitLab CIJenkinsPythonBashPowerShell

Observability & Reliability

DatadogSite24x7SLO/SLI designIncident responseOn-callITIL

git log --experience

Experience

Site Reliability Engineer II · Mastercard

Mar 2026 – Present

Dublin, Ireland

  • Ensure application scalability, performance, and resilience across production systems.
  • Manage incident response, conduct root cause analysis, and lead blameless post-mortems.
  • Implement DevOps automation and CI/CD pipelines to reduce manual intervention.
  • Establish and monitor Service Level Objectives (SLOs) to improve system reliability.
  • Partner with development teams on operational design, capacity planning, and monitoring.
  • Drive operational excellence and a developer-run ownership culture.

Technical Support Engineer — Azure Cloud · Zoho

May 2024 – Jul 2024

Chennai, India

  • Managed Azure production environments at 99.9% uptime with proactive monitoring, SLOs, and automated health checks — cutting mean time to detect by 35%.
  • Ran containerized workloads on Azure Kubernetes Service (AKS): troubleshooting pod failures, network policies, and resource optimization.
  • Automated infrastructure with PowerShell and Python, improving operational efficiency by 30% and reducing toil.
  • Participated in on-call rotation — incident triage, root cause analysis, and post-incident reviews.
  • Configured and troubleshot DNS, load balancers, and network security groups for secure, reliable connectivity.

Data Processing Analyst — Data Infrastructure · NielsenIQ

Oct 2022 – Apr 2024

Chennai, India

  • Led migration of on-premises Presto data systems to Azure Cloud, designing a highly available architecture supporting 5M+ daily transactions.
  • Built and managed Kubernetes clusters running 50+ containerized microservices with a focus on reliability and resource optimization.
  • Implemented Infrastructure as Code with Terraform and Ansible, cutting provisioning time by 40% and ensuring consistency across environments.
  • Managed PostgreSQL and Redis infrastructure at 99.9% uptime for business-critical processing serving Fortune 500 retail clients.
  • Built monitoring and alerting that proactively surfaced bottlenecks, reducing incident response time by 35%.
  • Led a team of six, coordinating delivery and fostering collaborative problem-solving.

Cyber Security Intern · Tevel Cyber Corps

Jun 2021 – Jul 2021

Chennai, India

  • Assisted in threat analysis, incident response, and vulnerability assessments.
  • Conducted security audits and contributed to compliance and security-awareness efforts.
  • Researched emerging threats and helped implement new security tooling.

cat ~/education

Education

Master of Science — Computing (DevOps) · Atlantic Technological University

Sep 2024 – Sep 2025

Coursework across Jenkins, Docker, Kubernetes, and cloud platforms (AWS, Azure) — automating infrastructure provisioning and managing DevOps pipelines end to end.

Bachelor of Science — Information Technology · Guru Nanak College

2019 – 2022

Foundations in computing, networks, and software — where my interest in cloud and DevOps first took shape.

ls ~/projects

Selected Work

Cloud Infrastructure Migration (IaC)

Led a Terraform + Ansible migration that codified infrastructure end to end, cutting deployment time by 40% and making environments reproducible.

TerraformAnsibleAzureIaC

Production Kubernetes Platform

Ran production Kubernetes / AKS at 99.9% uptime with pod security policies, container hardening, and resource optimization for demanding workloads.

KubernetesAKSSecurityArgo CD

Observability & SLO Automation

Stood up monitoring and alerting with Datadog and Site24x7 — SLOs plus automated health checks that cut mean time to detect by 35% and incident volume by 25%.

DatadogSite24x7SLOReliability

./contact.sh

All systems operational

Let's build reliable systems

Open to interesting problems in cloud, DevOps, and reliability.