aia
aia3d ago
New↻ Repost

Senior Manager, Cloud Solution Architect (Observability)

MalaysiaMalaysia·Kuala Lumpursenior
OtherCloud Solution Architect
1 views0 saves0 applied

Quick Summary

Key Responsibilities

Observability Architecture & Design• Define and maintain enterprise observability reference architecture, standards, patterns, and governance across AIA Group and Business Units.

Technical Tools
OtherCloud Solution Architect

AIA Digital+ is a Technology, Digital and Analytics innovation hub dedicated to powering AIA to be more efficient, connected and innovative as it fulfils its Purpose to help millions of people across Asia-Pacific live Healthier, Longer, Better Lives.

If you are hungry and driven to play an active role in shaping a better tomorrow, we want to hear from you. Because the work we do at AIA Digital+ makes a difference in the lives of millions of people, every day. We will equip you with the critical skills, tools and technology, and endless opportunities to learn, contribute and thrive in a dynamic and exciting environment.

If you want to shape a brighter future at AIA Digital+, please read on.

About the Role

~1 min read
The Observability Architect is responsible for defining, designing, implementing, and governing enterprise observability capabilities across AIA multi-cloud, hybrid, and containerized environments. The role will establish consistent monitoring, logging, tracing, event correlation, service health, and operational intelligence practices across Azure, Alibaba Cloud, and related enterprise platforms.

Dynatrace will be the primary observability platform, integrated with ServiceNow ITOM to support event management, incident enrichment, service mapping, root cause analysis, and operational automation. The role will also define complementary patterns for logging and open-source monitoring platforms such as Elastic, OpenSearch, Prometheus, Grafana, and OpenTelemetry.

A key objective of this role is to drive AIOps-enabled operations and self-healing automation to reduce alert noise, improve Mean Time To Detect (MTTD), reduce Mean Time To Resolve (MTTR), and improve overall platform and application reliability.

Responsibilities

~1 min read

Requirements

~2 min read

Skills:
•    10+ years relevant experience in Cloud Architecture, Infrastructure, Operations, Observability, or Application Performance Management.
•    5+ years practical experience designing and governing enterprise observability platforms.
•    Strong hands-on experience with Dynatrace, including APM, infrastructure monitoring, Kubernetes monitoring, dashboards, management zones, service mapping, and Davis AI capabilities.
•    Experience integrating observability platforms with ServiceNow ITOM / Event Management / ITSM processes.
•    Strong experience with logging platforms such as Elastic Stack, OpenSearch, or equivalent enterprise log analytics solutions.
•    Experience with open-source monitoring and visualization platforms such as Prometheus, Grafana, and OpenTelemetry.
•    Hands-on experience with Azure native monitoring services including Azure Monitor, Log Analytics, Application Insights, Network Watcher, Azure Managed Prometheus, and Azure Managed Grafana.
•    Hands-on experience with Alibaba Cloud native monitoring services including CloudMonitor, Log Service (SLS), ActionTrail, ARMS, Managed Service for Prometheus, and related observability services.
•    Strong understanding of Kubernetes, containers, microservices, APIs, network monitoring, distributed tracing, cloud networking, IAM, and security principles.
•    Experience implementing AIOps, event correlation, anomaly detection, automated remediation, and self-healing operations.
•    Experience working in enterprise or regulated environments is highly desirable.
•    Relevant professional certifications such as Dynatrace Associate/Professional, Microsoft Azure Solutions Architect Expert, Alibaba Cloud Professional Architect, CKA, ITIL, SRE Foundation, or TOGAF will be an advantage.
•    Sound understanding of IT partner ecosystem and partner collaboration in a multinational corporation.
•    Experience in top-tier multinational corporation will be an advantage.
•    Strong problem-solving and analytical skills.
•    Excellent communication, stakeholder management, and collaboration skills.
•    Ability to translate complex operational telemetry into actionable service reliability insights.
•    Ability to operationalize disruptive technology services, including building implementation roadmaps for observability, AIOps, and self-healing automation.
 

Build a career with us as we help our customers and the community live healthier, longer, better lives.

You must provide all requested information, including Personal Data, to be considered for this career opportunity. Failure to provide such information may influence the processing and outcome of your application. You are responsible for ensuring that the information you submit is accurate and up-to-date.

Location & Eligibility

Where is the job
Kuala Lumpur, Malaysia
On-site at the office
Who can apply
MY

Listing Details

Posted
August 3, 2026
First seen
August 3, 2026
Last seen
August 6, 2026

Posting Health

Days active
0
Repost count
1
Trust Level
44%
Scored at
August 3, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

aiaSenior Manager, Cloud Solution Architect (Observability)