sanjay.hona
Local time in Ottawa, Canada:--:--ET

all systems operational· open to new roles

Sanjay Hona

Site Reliability Engineer @ EdgeSignal

  • DevOps
  • DevSecOps
  • SRE
  • AIOps · learning

I build and secure cloud platforms that ship fast and stay up. Seven years across Kubernetes, Terraform, and hardened CI/CD, plus the observability and automation that keep production reliable. Reliability and security, I treat as the same job.

Page me LinkedIn
❯ kubectl get engineerNAME          READY   STATUS    RESTARTS   AGEsanjay-hona   1/1     Running   0          7y❯ kubectl describe engineer sanjay-honaLocation:    Ottawa, CanadaRole:        SRE · EdgeSignalFocus:       AWS · K8s · TerraformLearning:    CKA · in progressReading:     Designing Data-Intensive AppsPracticing:  Vim · mastery goal 2027Events:        Normal  Available  open to new roles❯ 

Numbers from production

The results I can point to, from real systems and real on-call rotations.

Time in production

7+years

DevOps → DevSecOps → SRE, since 2019

Edge fleet operated

40+devices

Running in the field, watched remotely

Incident resolution

−30%MTTR

ELK + Prometheus + Grafana · BeyondID

Release frequency

3×more releases

Standard CI/CD over 10+ services · Innovate Tech

Environment setup

<2hwas 3+ days

Company-wide Terraform migration · BeyondID

Critical security audit

0findings

Passed clean · BeyondID

Service status

Every discipline I run, tracked like a service. One cell per quarter since 2016.

  • Software engineeringsince 2016 · 10y
    operational
  • DevOpssince 2019 · 7y
    operational
  • DevSecOpssince 2021 · 5y
    operational
  • SREsince 2025 · 1y
    operational
  • AIOpssince 2026 · new
    canary
  • in production
  • canary · learning
  • no data

Experience

Roles, newest first. Every one of them shipped to production.

  1. v4.02025 — nowrunning

    Site Reliability Engineer @ EdgeSignal

    Own production reliability for a distributed system — deployment automation, fleet-wide observability, and incident response on Datadog and CloudWatch. Keep the rollout pipeline boring and the pager quiet.

    • AWS
    • Docker
    • Datadog
    • CloudWatch
    • Python
  2. v3.02021 — 2024exit 0

    DevOps / DevSecOps Engineer @ BeyondID

    Built AWS EKS platforms on GitOps (ArgoCD, Helm) and cut deploy failures 15%. Led the company-wide Terraform migration — env setup from 3+ days to under 2 hours — and stood up ELK/Prometheus/Grafana that took 30% off MTTR. Passed a critical security audit with zero findings.

    • −15% deploy failures
    • 3+ days → <2 h env setup
    • −30% MTTR
    • 0 audit findings
    • AWS EKS
    • Terraform
    • ArgoCD
    • Prometheus
    • Security Audit
  3. v2.02019 — 2021exit 0

    DevOps Engineer @ Innovate Tech

    Migrated legacy CloudFormation to a modular Terraform framework; standardized CI/CD across 10+ microservices — deployment errors down 35%, release frequency tripled. Built proactive security monitoring with Prometheus and AWS CloudTrail.

    • −35% deployment errors
    • 3× release frequency
    • 10+ services on one CI/CD
    • Terraform
    • GitLab CI
    • CloudTrail
  4. v1.02016 — 2019exit 0

    Software Engineer @ Sastra Creations · Three Monks

    Containerized a monolithic GPS platform with Docker to absorb a 200% IoT traffic spike; built a real-time layer with Spring Boot and WebSockets delivering sub-second fleet telemetry. Earlier, shipped data-driven Java/JSP web apps on Nginx and Tomcat.

    • 200% IoT traffic spike absorbed
    • sub-second telemetry
    • Docker
    • Spring Boot
    • Node.js
    • Java

Platform stack

What I run, drawn the way it runs: delivery on top of runtime on top of cloud, with security and observability wired through every layer.

Security

cross-cutting
  • CloudTrail
  • audits
  • least-privilege
  • supply-chain
  1. Delivery

    IaC · CI/CD
    • Terraform
    • GitLab CI
    • GitHub Actions
    • ArgoCD
  2. Runtime

    Orchestration
    • Kubernetes
    • Docker
    • Helm
  3. Infrastructure

    Cloud
    • AWS
    • Azure

Observability

cross-cutting
  • Datadog
  • Prometheus
  • Grafana
  • ELK

About

My path started in Kathmandu and runs through seven years of production systems: Java services first, then containers, then whole platforms on AWS and Azure. The thread: the best infrastructure is the kind nobody notices, and the most secure is the kind an attacker never gets a foothold in.

Off the clock I run a three-node Kubernetes homelab, practice Vim, and I'm working through AIOps: anomaly detection and automated remediation on top of the observability I already run.

Contact

on call · accepting pages

Open to platform engineering, reliability, and security work.

The fastest way to reach me is email.

Page me
  1. L1Email[email protected]primary
  2. L2LinkedIn/in/sanjayhona fallback
  3. TZLocationOttawa, Canada · --:-- ETlocal time