Ajay Matharu

DevOps Guy

About Me

As a Staff DevOps and Site Reliability Engineer, building and running resilient cloud infrastructure for financial and technology companies is the primary mission. Core expertise centers around AWS, Kubernetes, and the surrounding ecosystem—including CI/CD pipelines, Terraform, Ansible, and robust observability stacks.

Solving complex infrastructure challenges requires more than just technical know-how; it involves translating architectural decisions into strategies that non-technical stakeholders can easily understand and act upon. Key career milestones include migrating full production stacks across global regions to drastically cut latency with zero unplanned downtime. Additionally, a major focus has been modernizing legacy Kubernetes environments into highly available, multi-AZ EKS clusters to support high-traffic healthcare applications and production trading workloads without cluster-level incidents.

In high-stakes environments, directing incident response for production systems ensures rapid resolution and thorough post-mortems when downtime carries a real business cost. Consistent value is delivered by bridging the gap between development, security, and business operations. This involves leading cross-functional engineering teams across Hong Kong and India, and partnering directly with trading desks and developers to ensure infrastructure changes are production-ready long before they go live.

Core Expertise

  • Cloud & Infrastructure: AWS (VPC, EC2, EKS, RDS, MSK), GCP, Terraform, and Ansible.

  • Operations & Security: Zero-trust networking (Zscaler, ScaleFT), IAM hardening, Privileged Access Management, and CI/CD automation.

  • Containerization & Tooling: Kubernetes, Docker, GitHub Actions, Jenkins, and Buildkite.

  • Databases & Observability: PostgreSQL, MongoDB, Redis, Kafka (AutoMQ, MSK), Grafana, Loki, and Mimir.

Education

  • B.Sc Computer Science From Mumbai University