Hai Hoang

Sydney NSW 2000 Australian Citizen in

Summary

Systems Engineer with 5+ years building, automating, and operating large-scale distributed systems on AWS. Develops production services and tooling in Go and Python, designs infrastructure-as-code with TypeScript CDK, and debugs complex issues across networking, IAM, compute, and security layers.

Core Skills

Software Development

  • Production service development in Go and Python
  • Code review, testing, and CI/CD pipeline design
  • AWS CDK (TypeScript) infrastructure-as-code
  • Data structures and algorithms in distributed systems contexts

Distributed Systems & Cloud

  • Large-scale system design, deployment, and operation
  • AWS: EC2, Lambda, VPC, IAM, Systems Manager, Config, CloudWatch
  • Networking, security architecture, and configuration management
  • Observability, monitoring, and incident resolution at scale

Systems Engineering & Automation

  • Production debugging across distributed systems (networking, IAM, compute, storage)
  • Incident response, root-cause analysis, and corrective actions
  • Automation for configuration enforcement and proactive remediation
  • Operational excellence: toil reduction and observability improvements

Homelab & Personal Projects

  • Kubernetes/Docker homelab running distributed services
  • Prometheus/Grafana monitoring, GitOps automation, CI/CD workflows
  • Prototyping system designs and evaluating emerging tooling

Certifications

Professional Experience

Amazon Web Services — Amazon Managed Services Sydney

Systems Dev Engineer IIMar 2025 – Present
System Engineer IINov 2022 – Mar 2025
Operations EngineerNov 2020 – Nov 2022
Operations AssociateNov 2019 – Nov 2020

Key Achievements

  • Developed and maintained a distributed AWS-native service in Go and Python, delivering features and performance improvements to production.
  • Reviewed code from peers, enforcing best practices around style, testability, accuracy, and efficiency.
  • Designed system architecture with stakeholders, selecting technologies for scalability and security requirements.
  • Triaged and debugged production issues across IAM, networking, compute, and storage — analysing root causes and coordinating fixes with upstream service teams.
  • Built automation tooling (SSM Automation Documents) for configuration enforcement and proactive remediation across the organisation.
  • Authored and maintained documentation, runbooks, and educational content, adapting materials based on operational feedback.
  • Modernised CI/CD pipelines using TypeScript AWS CDK, improving deployment reliability and developer velocity.

White Clarke Asia Pacific — Systems Administrator

Systems AdministratorNov 2014 – Nov 2019
  • Designed and deployed multi-site infrastructure (SAN, vCenter, SD-WAN) serving operations across multiple countries.
  • Built automated backup, replication, and secure remote access systems supporting 50+ staff.

Technical Support / Service Desk Roles

Various Roles2009 – 2014
  • L1/L2 support, incident triage, escalation management, and operational troubleshooting foundations.

Education

Bachelor of Software Engineering (Computer Science Specialty)

University of South Australia — 2006–2010