Staff DevOps Engineer · Based in Singapore

Building reliable
cloud platforms
that scale.

DevOps · Site Reliability · Platform Engineering

Platform and reliability engineer with 15+ years in infrastructure and over a decade operating containerised AWS platforms in high-availability production. I set the standards and build them hands on, from EKS and secure CI/CD to AI-assisted operations.

Alamsyah Ho
10+ yrs
AWS & Kubernetes
5×
Certifications
300+
Pods at peak
25–40%
Infra cost reduction
30 min
Cloud cutover downtime
1,000+
Security findings closed
About

Experience you can rely on.

I operate containerised AWS platforms in high-availability production, with depth in Kubernetes, SLO-driven operations, incident response, secure CI/CD and IAM least-privilege. AWS DevOps Professional and CKS certified.

My track record spans cloud migration, toil reduction and mentoring engineers, including work in MAS-regulated payments and insurance environments. Lately I have been building Claude-based operational agents and MCP gateways that take real toil off the team.

My core values center on passion and lifelong learning, with a particular enthusiasm for automation, containerization and cloud-native engineering.

🎯

Experience

A proven track record across production platforms, regulated environments and high-traffic systems built over more than 15 years.

⚡

Autonomy

Self-reliant problem solving. I take ownership end to end, from architecture decisions to the hands-on implementation and the on-call that follows.

🤝

Involvement

Collaborative by default. I set platform standards, mentor engineers and run enablement so teams ship securely and confidently.

Core skills

The toolkit

A decade of cloud-native engineering across platform, reliability, security and AI-assisted operations.

☁️ Cloud & Platform

AWS EKSEC2LambdaIAMVPCCloudWatchKubernetesHelmDockerLinux (RHEL)

🤖 AI & Agentic Ops

MCP server designFastMCPboto3 / SigV4Bedrock AgentCoreClaude ops agentsCost anomaly detectionIAM reviewRelease monitoring

🛡️ Reliability & SRE

SLI / SLOError budgetsIncident commandPostmortemsCapacity planningOn-callToil reduction

📊 Observability

DatadogPrometheusGrafanaELKCloudWatch

🚀 CI/CD & IaC

JenkinsGitLab CIGitHub ActionsArgoCDTerraformAnsiblePuppetPythonBash

🔐 Security

Image hardeningCrowdStrikeSecrets ManagerSSM Parameter StoreIAM least-privilegeGuardDutyControl TowerSecurity Hub
Career

Where I have built things

Incube8 Pte Ltd

Jun 2015 – Present
Singapore
Staff DevOps Engineer (2021–Present) · Senior DevOps Engineer (2018–2021) · Senior Linux Systems Administrator (2015–2017)
  • Architected and led the migration of 5+ production applications from legacy EC2 to Kubernetes (EKS), re-engineering Ansible pipelines into Helm releases for consistent, auditable deployments.
  • Own production on-call and incident response for 10+ services; introduced blameless postmortems, runbook standards and severity-based escalation, improving MTTR by 40%.
  • Tuned Kubernetes autoscaling and capacity planning to sustain 300+ pods at peak, cutting monthly spend 25–40% with no SLO regression.
  • Built Claude-based operational tooling and a remote MCP gateway exposing Amazon Managed Prometheus metrics to an AI agent, removing 5 to 10 engineer-hours of toil per week.
  • Led the on-premises (Switch, Las Vegas) to AWS migration end to end, executing cutover with 30 minutes of total downtime.
  • Mentor junior and mid-level engineers, set platform standards, and run internal enablement on CI/CD, Kubernetes and secure pipeline practice.

GovTech Singapore

Mar 2021 – Jun 2021
Lead DevOps Engineer
  • DevOps lead on the Ministry of Manpower (MOM) engagement, directing a team of 6 across platform operations and release management within a government security baseline.

Network for Electronic Transfers (NETS)

Mar 2017 – Apr 2018
Senior Consultant, DevOps
  • Core member of the DevOps team driving Infrastructure as Code adoption in a MAS-regulated payments environment.
  • Designed automated CI/CD workflows that removed configuration drift and standardised release practice across 3 application teams.
  • Worked within change-control and audit requirements aligned to MAS TRM Guidelines and PCI-DSS.

PT Chubb Life Assurance Indonesia

Mar 2014 – Jun 2015
IT Assistant Manager, IT Infrastructure · Jakarta, Indonesia
  • Operated and optimised the enterprise Java application server estate for a life insurer, owning deployment, troubleshooting and the availability of business-critical systems.
  • Led infrastructure design and optimisation projects, deepening the production reliability and enterprise platform experience that underpins my SRE work today.

PT NTT Data Indonesia

Feb 2013 – Feb 2014
Systems Engineer · Jakarta, Indonesia
  • Designed, implemented and supported centralised infrastructure solutions, keeping business systems highly available against defined SLAs.
  • Owned technical solutions end to end, an early grounding in the SLA-driven, availability-first mindset that is central to how I approach reliability engineering now.

PT Andaman Lestari Multikreasi

Oct 2008 – Dec 2012
IT Assistant Manager · Jakarta, Indonesia
  • Managed the IT department day to day and supervised full-time systems and developer staff, an early step into the technical leadership and mentoring I do today.
  • Advised client companies on solutions and support, building the cross-functional and stakeholder-facing habits I still rely on as a platform engineer.

PT Sinarmas Multifinance

Apr 2007 – Oct 2008
Systems Engineer · Jakarta, Indonesia
  • Began my career supporting Linux and Windows production systems, resolving application issues and backing the helpdesk.
  • This hands-on systems support built the operational discipline and troubleshooting instinct that my cloud and DevOps career grew from.
Selected engineering detail

Problem → Approach → Result

Incube8 · 2021–2022

EC2 to Kubernetes platform migration

Problem 5+ revenue-critical applications on hand-managed EC2 with inconsistent, unauditable releases.
Approach Re-engineered Ansible deployment pipelines into Helm releases on EKS, with hardened base images and policy gates in CI.
Result 10 to 20 deploys per week, rollback in 15 minutes, and a single audit trail across environments.
Incube8 · 2022

Observability & reliability programme

Problem Incidents surfaced by users before monitoring, with no agreed service objectives.
Approach Unified Datadog, Prometheus/Grafana, ELK and CloudWatch; built endpoint-level monitoring; defined SLIs/SLOs and alert routing per tier.
Result MTTD down 50%, MTTR down 40%, and 2 repeat incidents eliminated.
Incube8 · 2024

Pipeline & runtime security hardening

Problem Inconsistent base images and long-lived IAM credentials across 15 services, with findings surfacing only in production.
Approach Set golden base images, moved scanning left into CI, extended CrowdStrike runtime coverage, and replaced static IAM users with scoped service roles.
Result Critical image findings down 50%, static credentials retired across 12 IAM accounts, least-privilege across 15 services.
Incube8 · 2025–2026

AWS Security Hub remediation programme

Problem A large backlog of high and critical Security Hub findings across workload accounts, with no owner or triage process.
Approach Triaged by severity and control, grouped into repeatable remediation patterns, drove fixes with owning teams, and folded controls into account baselines.
Result 1,000+ high and critical findings closed across multiple accounts, with a sustained remediation cadence.
Credentials

Certified & grounded

AWS
AWS Certified DevOps Engineer — Professional
CKS
Certified Kubernetes Security Specialist
CKA
Certified Kubernetes Administrator
CKAD
Certified Kubernetes Application Developer
RHCE
Red Hat Certified Engineer
Education

Bachelor of Electrical Engineering

Tarumanagara University
2001 – 2007
Languages
English · Bahasa Indonesia

Status
Singapore Citizen

Let's build something reliable.

Open to conversations about platform engineering, SRE and cloud architecture roles. The fastest way to reach me is email.