Arash Haghighat

Site Reliability Engineer and platform consultant. Fifteen years keeping critical infrastructure reliable: cloud platforms, Kubernetes, and the practices that keep systems resilient. Now as a tech lead and trainer.
- 15+years in infrastructure
- Kubestronautall five CNCF Kubernetes certifications
- −40%cloud cost reduction, monolith to Kubernetes
- Head of Infraat a fast-growing logistics scaleup
Skills
Infrastructure
Cloud or on-premises, I design, implement, and maintain robust infrastructures. AWS, GCP, and Azure, with solid Proxmox experience and IaC in Terraform and Ansible.
Cloud native
Building cloud-native applications, managing orchestration with Kubernetes, Helm, and Harbor, and driving GitOps workflows with Flux across the container lifecycle.
Development and CI/CD
From GitLab CI and CircleCI to end-to-end testing and supply chain security, I architect and maintain build pipelines and platform services.
Observability
Prometheus, Grafana, Datadog, ELK Stack. I design monitoring and alerting strategies that give teams real-time insight into system health at scale.
Platform engineering
Self-service tooling, golden paths, and developer onboarding. I build internal platforms that improve DevEx and help teams ship faster.
Reliability
SLOs, incident response, chaos engineering, disaster recovery, and compliance. I build the practices and culture that keep systems resilient.
Professional milestones
- to
Consultant, Tech Lead, and Trainer
I lead platform and reliability engagements at enterprise scale, serving as Tech Lead on mission-critical multi-region landing zones and container platforms.
I build compliance automation, self-service infrastructure, and help organizations adopt modern engineering practices.
I train teams on SRE, platform engineering, and developer experience, combining hands-on IC delivery with team leadership.
- to
Head of Infrastructure and Operations
I led cross-functional engineering teams at a fast-growing logistics scaleup, owning the full stack from CI/CD to observability and high availability.
I drove a 40% cost reduction through cloud migration and Kubernetes adoption, moving from monolith to microservices.
I also led the expansion into new countries on GCP, scaling infrastructure and performance to match the pace of the business.
- to
Site Reliability Engineer and Tech Lead
I pioneered SRE and DevOps practices at enterprise companies in telco and finance, often as the first hire in these roles.
I built teams from the ground up, shaped engineering culture, and developed deep expertise in Linux, automation, and production operations.
It was an exciting time of rapid growth, turning new disciplines into core capabilities.
- to
Software Engineer
I kicked off my career as a Software Engineer, focusing on mobile and embedded systems, all while finishing up university.
I joined the university's ACM team and took part in programming contests, which sharpened my problem-solving skills.

