Sandeep Sidhu

Sandeep Sidhu

Professional summary

Site reliability engineer with 20+ years across production infrastructure: public cloud operations at scale, Kubernetes and AWS estates, and infrastructure as code with Go, Python, Terraform, and Ansible.

Now focused on security and observability for software teams that do not have a dedicated security function. I'm the founder of AlertKick, a security monitoring and compliance platform with an eBPF-based agent, and I operate its multi-region production fleet - where AI agents carry out real operational work (deploys, diagnostics, change management) under audited change control. That combination - long SRE experience, security engineering, and hands-on AI operations - is what I bring to every engagement.

I'm process driven with a knack for finding problem areas, and I only rest when things are streamlined and under control. Proven leadership of diverse teams of up to 8 engineers.

Skills

  • Software development: Go throughout my recent work, plus years of Python tooling; frontend with React and web technologies.
  • Systems engineering: Linux/Unix, virtualization, high availability, load balancing, MySQL, clustering, and all things cloud.
  • Security engineering: eBPF-based detection, file integrity monitoring, SSH hardening and access control, MITRE ATT&CK mapping, compliance frameworks (SOC 2, PCI, SOX).
  • Networking: Linux-based systems and firewalls; Kubernetes networking design and management.
  • Leadership: Building motivated teams, workflow and process improvement, automation; led teams of up to 8 Linux administrators.

Work experience

AlertKick (January 2025 - present) - Founder

Building a security monitoring and compliance platform for teams without a dedicated security function: eBPF-based security agent (Linux and Windows), multi-region SaaS, SOC 2/PCI/SOX-style compliance reporting, and closed-loop change control for AI agents operating production systems. Started as a side project in January 2025; incorporated May 2026.
Go · eBPF · Ansible-managed fleet · Kafka · MongoDB · React

GoDaddy EMEA (2021 - present) - Senior Site Reliability Engineer

SRE on GoDaddy EMEA's Kubernetes-based infrastructure.
Go · Kubernetes · AWS · Terraform · Linux

Sky (May 2020 - 2021)

GCP, Kubernetes, Terraform, Puppet, Ansible, Python, Groovy, and lots and lots of Jenkins pipelines - improving deployment processes through automated pipelines while supporting new feature requests and handling failures.

YNAP (YOOX NET-A-PORTER GROUP) (July 2017 - March 2020)

Designed, built, and maintained infrastructure as code via Terraform supporting 10+ environments across multiple AWS accounts. Deployments via Jenkins declarative pipelines executed in Docker, mainly using Ansible.

Education

MCA - Master of Computer Applications, Punjab Technical University, Jalandhar

Certifications

Previously held (now expired):

  • Red Hat Certified Engineer
  • Certified Kubernetes Security Specialist
  • Certified Kubernetes Administrator
  • SCSA - Sun Certified Systems Administrator

Detailed history on LinkedIn.