2mo ago
New

Site Reliability Engineer | Hybrid - Centris/Makati

PhilippinesPhilippines·Makati City
EngineeringDevops Engineer
0 views0 saves0 applied

Quick Summary

Key Responsibilities

You will design and define standards, patterns, and automations opportunities that elevate monitoring and reliability across platforms and applications, with a strong focus on Azure Monitor,

Requirements Summary

Bachelor’s degree in IT /Computer science/Engineering, or related field.

Technical Tools
EngineeringDevops Engineer
  • Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipeline:
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Knowledge of Grafana, Prometheus, App Dynamics, ThousandEyes
  • Uses SLI/SLOs, postmortems, and CMDB and other context to reduce noise, drive self-healing, and measurably improve MTTR and KPIs.

Ensure the reliability of a client's critical monitoring and event management systems and services, provide standards and governance around monitoring and observability which enable Chevron business-critical processes to operate safely, efficiently, and reliably.

Responsibilities

~1 min read
  • →You will design and define standards, patterns, and automations opportunities that elevate monitoring and reliability across platforms and applications, with a strong focus on Azure Monitor, ServiceNow ITOM Event Management, Grafana, and APM/Synthetics tooling
  • →You’ll partner with product teams to implement SLO/SLI-driven operations, reduce alert noise, accelerate incident response, and embed self-healing where it matters most.
  • →Engineer enterprise monitoring & event patterns by authoring and maintaining reference architectures, runbooks, and event management models (alert → event → incident) with actionable alerts and incidents routing.
  • →Contribute to Monitoring and Observability & Event Management Strategy and tooling intake/governance checkpoints and coach product teamsInternal - General Use
  • →Excellent communication skills to drive continuous improvement by reducing alert noise, shorten MTTR, and improve change success by embedding postmortem learnings into patterns, rules, and pipelines.

  • SRE Practices: Observability and Monitoring
  • Cloud Observability: Azure Monitor/App Insights/Log Analytics (KQL)
  • Grafana/Prometheus for metrics visualization where applicable
  • ServiceNow ITOM Event Management
  • Azure Fundamentals, Azure Monitor
  • DevOps and Automation Tools
  • Grafana, Prometheus, App Dynamics, ThousandEyes
  • Application Performance Monitoring and Digital User Experience tools

Requirements

~1 min read
  • Bachelor’s degree in IT /Computer science/Engineering, or related field.
  • 3+ years in monitoring/observability/SRE roles with hands-on experience in Azure Monitor/App Insights (KQL) and ServiceNow Event Management.
  • Strong knowledge in Azure Log Analytics, KQL, Telemetry, APM implementations
  • Demonstrated ability to collaborate across IT Operations team, platform, cyber, network, and product teams, strong written verbal communication for standards and enablement.

  • 5+ years of experience with SRE role and deep understanding of monitoring and application performance management
  • Knowledge of SLO platforms (e.g., Nobl9) and experience contributing to standards/governance artifacts.
  • Knowledge of proactive monitoring using Azure monitor services, telemetry, and synthetic transactions.
  • Understanding of network architecture and security: WAN/LAN, TCP/IP, PKI.Internal - General Use
  • Familiarity with ITSM processes and tools (e.g., ServiceNow), and compliance processes
  • Have AIOps vision and awareness

  1. Communication & Teaming – Able to translate complex reliability patterns into consumable standards and coach IT operations team via office hours/CoP sessions.
  2. Technical Depth in Monitoring and Observability Stack – Hands-on in ServiceNow Event Management, Azure Monitor/KQL, and automation.
  3. Analytical & Systems Thinking – Uses SLI/SLOs, postmortems, and CMDB context to reduce noise, drive self-healing, and measurably improve MTTR and KPIs.

  • Work set-up: Hybrid 3x / RTO 2x per week | Eton, Centris
  • Work shift: Nightshift

Location & Eligibility

Where is the job
Makati City, Philippines
On-site at the office

Listing Details

Posted
July 1, 2026
First seen
September 28, 2026
Last seen
September 28, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
16%
Scored at
September 28, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

Site Reliability Engineer | Hybrid - Centris/Makati