System Reliability Engineer, Consultant
Quick Summary
At AIA we’ve started an exciting movement to create a healthier, more sustainable future for everyone. As pioneering innovators for over 100 years,
As pioneering innovators for over 100 years, we’re now transforming our organisation to be faster, simpler and more connected. Because we want to be even better equipped to develop digital solutions and experiences that help more people live Healthier, Longer, Better Lives.
To get there, we need people with tech/digital/analytics expertise and passion to help develop positive, sustainable change through digitally enhanced experiences that will impact the lives of millions of people and create a healthier future for everyone.
About the Role
~1 min readResponsibilities
~1 min readMonitor and report on application performance, and highlight any deviations or issues.
Collaborate with application engineers and developers to identify root causes and implement durable fixes.
Participate as a Subject Matter Advisor during production incidents and outages.
Provide insights backed by system monitoring, code review, and database analysis.
Support post‑mortem reviews and drive follow‑up actions.
Automate operational tasks such as monitoring, alerts, and recovery processes.
Build scripts and internal tools to eliminate manual toil and improve operational efficiency.
Implement telemetry and observability practices to track system health, latency, and error rates.
Manage the Dynatrace platform and its integrations with application services.
Support teams in designing dashboards and visualization setups.
Work with Security teams to ensure systems comply with regulatory and industry standards (e.g., PCI‑DSS, GDPR).
Implement necessary access controls, encryption, and audit capabilities within SRE scope.
Analyze usage trends to forecast demand and support scaling decisions.
Contribute to cost‑performance optimization efforts across infrastructure and applications.
Collaborate closely with development, QA, and infrastructure teams to embed reliability into the SDLC.
Maintain clear and up‑to‑date operational documentation, runbooks, and architecture diagrams.
Champion SRE principles across the organization to foster resilience and accountability.
Requirements
~1 min readBachelor’s degree in Computer Science, Software Engineering, IT, or related fields.
3–5 years of experience in SRE, DevOps, or Software Engineering roles.
Experience supporting front‑end applications in production environments, ideally within financial services or other regulated industries.
Strong understanding of front‑end performance monitoring and instrumentation.
Hands‑on experience with Real User Monitoring (RUM), Synthetic Monitoring, and APM tools (e.g., Dynatrace, New Relic, Datadog).
Proficiency in building dashboards and alerts using Dynatrace, Grafana, Prometheus, Elastic Stack, or Splunk.
Familiarity with OpenTelemetry for distributed tracing.
Scripting skills in Python, Bash, or JavaScript.
Experience with CI/CD pipelines (e.g., GitHub Flow).
Practical experience with cloud technologies (AWS or Azure).
Knowledge of Docker and Kubernetes.
Understanding of secure coding practices for front‑end applications.
Awareness of financial compliance standards such as PCI‑DSS.
What We Offer
~1 min readLocation & Eligibility
Listing Details
- First seen
- September 27, 2026
- Last seen
- September 27, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 56%
- Scored at
- September 27, 2026
Signal breakdown
Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.