Experience

Eight years, five roles

Platform engineering leader with 8 years of experience building Internal Developer Platforms, self-service infrastructure, and the reliability, security, and cost foundations that help product teams ship faster. I treat the platform as a product, designing golden paths and paved roads that cut developer cognitive load and remove ticket-based handoffs. Most of my work has been converting unstable, manually operated systems into declarative, GitOps-driven platforms, with hands-on depth across Infrastructure as Code, Kubernetes at scale, DevSecOps, and FinOps, alongside leading and mentoring SRE and platform teams.

Get in touchLinkedIn

Selected impact

Numbers across the last four years

~70%less config driftstandardised infrastructure provisioning on Terraform
~40%lower compute costSpot, On-Demand and Reservation blending, holding capacity for traffic spikes
~50%lower MTTRfull-stack observability paired with structured incident response
99.9%availability sustainedPod Disruption Budgets, topology spread, and mixed node pools
~40%faster rolloutsend-to-end ArgoCD GitOps with automated, auditable delivery

Timeline

Where the time went

M2P Fintech2025 to Present · 1y 7m

Senior SDE (Manager), Platform Engineering

TerraformCrossplaneArgoCDKyverno
Mad Street Den2021 to 2025 · 4y

SRE → Senior SRE → Technical Lead

KubernetesObservabilityDRFinOps
Qube Cinema Technologies2018 to 2021 · 3y

Software Engineer (Associate → SE)

AutomationAWSBilling

8y 7m across 3 teams

Roles

What I did, in detail

M2P Fintech

Senior SDE (Manager), Platform Engineering

2025 to Present
  • Standardised infrastructure provisioning across all services with Terraform, cutting manual configuration drift by ~70% and speeding up environment onboarding.
  • Introduced Crossplane to expose cloud infrastructure as self-service, Kubernetes-native APIs, giving product teams golden paths to provision resources without ticket-based handoffs.
  • Authored reusable Helm charts for all platform services, making Kubernetes deployments consistent, repeatable, and version-controlled.
  • Established end-to-end GitOps with ArgoCD, reducing rollout time by ~40% and delivering auditable, self-service continuous delivery across environments.
  • Built GitOps automation agents that auto-raise pull requests for platform updates to each product team's DevOps repositories, propagating changes consistently while product teams keep control through standard review and approval.
  • Developed a Claude-powered self-service skill that walks product teams through platform upgrades and new releases step by step, lowering onboarding friction and reducing support load on the platform team.
  • Drove a cultural shift from manual deployments to declarative, Kubernetes-native workflows, helping engineering teams adopt GitOps practices.
  • Led the migration from Ingress to Kubernetes Gateway API, improving traffic management, security control, and extensibility across services.
  • Built CI pipelines with integrated security scanning and policy gates, improving deployment safety and reducing rollback frequency.
  • Deployed Kyverno for admission-time policy enforcement and set up firewall monitoring and alerting to catch network-level threats early.
TerraformCrossplaneArgoCDHelmGateway APIKyverno

Mad Street Den

Technical Lead, Site Reliability Engineering

2024 to 2025
  • Led the migration of legacy workloads to Kubernetes, improving deployment consistency, resource utilisation, and operational scalability across the organisation.
  • Right-sized workloads from load-test and production metrics, reducing compute costs by ~35%.
  • Engineered pod-stability strategies using Pod Disruption Budgets, topology spread constraints, and Spot plus On-Demand node pools, sustaining 99.9% availability at lower instance cost.
  • Deployed a full DR setup covering circuit breakers, rate limiting, and automated failover, which reduced blast radius during incidents and hardened system resilience.
  • Built end-to-end observability with OpenTelemetry, Prometheus, Grafana, and Kibana / ELK, cutting MTTR by ~50% and improving SLA compliance.
  • Hardened security posture with WAF rules and Kubernetes network policies across all production services, and added caching layers that cut backend load and improved response times by ~40%.
  • Led a zero-data-loss AWS account migration covering DNS migration, service re-routing and cutover with minimal downtime, and performed cluster version upgrades with no production impact.
KubernetesOpenTelemetryPrometheusGrafanaWAF

Senior Site Reliability Engineer

2022 to 2024
  • Owned cloud cost engineering: built cost dashboards for engineering and leadership that surfaced spend leaks from NAT gateways, unused volumes, and inefficient data retention.
  • Optimised the EC2 fleet with a Spot, On-Demand, and Reservation blend, achieving ~40% compute-cost reduction while holding capacity for high-volume traffic spikes.
  • Implemented data backup and retention policies across storage systems, reducing storage costs by ~30% without compromising availability.
  • Ran capacity-planning exercises combining load-test results with production metrics to right-size compute and define scaling policies.
  • Configured intelligent auto-scaling and alarm rules for traffic surges, reducing over-provisioning by ~25% while protecting against performance degradation.
  • Managed Elasticsearch clusters at scale, tuning indexing, shard strategy, and retention to improve query performance and reduce storage overhead.
FinOpsAWSElasticsearchCapacity Planning

Site Reliability Engineer

2021 to 2022
  • Hired directly by the Director of Tech Support and SRE Director to stabilise a critical AI product suite that was struggling to meet SLA commitments.
  • Audited all organisation-wide products to map instabilities and failure patterns, then built the observability foundation from scratch: business and system metrics, dashboards, and alerting across all services.
  • Authored incident-response playbooks and runbooks for known failure patterns, then formed and led a 24/7 support team with on-call rotations, escalation paths, triage, and automation that resolved recurring issues and cut operational toil.
  • Established an SLO and error-budget framework, moving the team from reactive firefighting to proactive reliability management and cutting MTTR by ~40%.
SLOsIncident ResponseObservability

Qube Cinema Technologies

Software Engineer (Associate to SE)

2018 to 2021
  • Built an automated testing tool that reduced manual QA effort by ~50% and improved release velocity.
  • Cut AWS S3 storage costs by ~35% by implementing intelligent multi-tier storage and lifecycle policies.
  • Built the credit user-management feature in Qube Billing and the distributor-theatre report generation feature, improving financial workflow efficiency and cutting manual reporting effort by ~40%.
AutomationAWSBilling

Skills

Tools and practices

Languages
GoPythonNode.jsJavaScriptShell
Cloud
AWSGCPAzure
Containers & Orchestration
Kubernetes (EKS, GKE, AKS)DockerHelmAirflow
Kubernetes Platform
Gateway APIHPA / VPAPod Disruption BudgetsTopology SpreadKarpenterNetwork PoliciesKyverno
IaC & Control Planes
TerraformCrossplanePulumiCloudFormationAnsible
GitOps & CI/CD
ArgoCDGitHub ActionsJenkins
Observability
PrometheusGrafanaVictoriaMetricsOpenTelemetryJaegerLokiELKCloudWatchOpsgenie
DevSecOps & Policy
KyvernoTrivySonarQubeCheckovWAF
Databases
PostgreSQLElasticsearchRedisDynamoDBCosmosDBetcd
Reliability
Circuit BreakersRate LimitingCaching LayersAutomated FailoverDR Runbooks
FinOps
Cost DashboardsSpend-Leak DetectionCapacity PlanningSpot Strategy
MLOps (foundational)
MLflowKubeflow

Also

Focus areas and education

Expertise
Platform EngineeringInternal Developer PlatformPlatform as a ProductDeveloper ExperienceGolden Paths / Paved RoadsSelf-Service InfrastructureGitOpsInfrastructure as CodeDevSecOpsSite Reliability EngineeringFinOpsKubernetesObservabilityDisaster RecoveryIncident Response
Education

Bachelor of Engineering, Electronics & Communications

Sri Venkateswara College of Engineering, Sriperumbudur

2014 to 2018