2mo ago
New

SRE

IndiaIndia·Airolimid
EngineeringDevops Engineer
2 views0 saves0 applied

Quick Summary

Key Responsibilities

Actively monitor production environments, enterprise dashboards, and telemetry feeds using toolsets like Datadog, Dynatrace, and Grafana to spot anomalies before they impact end-users. [1, 2,

Requirements Summary

1 to 3 years of hands-on experience in an L1 Support, Infrastructure Monitoring, or Junior SRE role. Observability Tools: Hands-on experience navigating and tracking alerts within Datadog, Dynatrace,

Technical Tools
EngineeringDevops Engineer
Job Title: Site Reliability Engineer (SRE) / L1 Monitoring Engineer Job Summary We are seeking a proactive and technically driven SRE / L1 Monitoring Engineer with 1 to 3 years of experience to join our core digital infrastructure operations team. In this role, you will serve as the first line of defense ensuring the high availability, security, and performance of critical financial services and digital banking applications. You will be responsible for real-time system monitoring, tracking alerts across modern observability stacks, performing initial triage on infrastructure bottlenecks, and managing API traffic performance. This is an excellent opportunity for an early-career engineer looking to scale their skills in a high-volume, secure cloud infrastructure environment. [1] Key Responsibilities L1 Infrastructure Monitoring & Alerts Real-time Surveillance: Actively monitor production environments, enterprise dashboards, and telemetry feeds using toolsets like Datadog, Dynatrace, and Grafana to spot anomalies before they impact end-users. [1, 2, 3] Alert Triage: Acknowledge, validate, and categorize incoming infrastructure, database, and application alerts generated by Prometheus and application performance monitoring (APM) agents using predefined Standard Operating Procedures (SOPs). [1, 2, 3, 4, 5] Incident Escalation: Document incident details clearly in the ticketing system and swiftly escalate unresolved P1/P2 issues to L2 engineers or specialized DevOps teams with complete log snippets and context. Application Delivery & API Traffic Management Nginx Operations: Monitor web server logs, verify reverse proxy configurations, and troubleshoot basic traffic routing or SSL/TLS certificate errors. [1, 2, 3] API Gateways: Use Apigee to monitor API proxy performance, track error rates (5xx/4xx codes), track latency spikes, and check developer portal connectivity. [1, 2, 3, 4] Kubernetes Support: Monitor cluster health, inspect pod statuses, view application logs using kubectl, and track resource usage (CPU/Memory limits). [1, 2] Cloud Operations & Reliability GCP Monitoring: Utilize Google Cloud logging, native monitoring tools, and integrated observability dashboards to check the health of virtual machines, storage, and networking layers. [1, 2, 3, 4] Health Checks: Perform routine daily morning sanity checks and post-deployment validation steps for critical banking services. Runbook Execution: Execute automated or manual scripts to restart failed services, clear disk space, or cycle pods safely in staging and production environments. Required Qualifications & Technical Skills Experience: 1 to 3 years of hands-on experience in an L1 Support, Infrastructure Monitoring, or Junior SRE role. Observability Tools: Hands-on experience navigating and tracking alerts within Datadog, Dynatrace, Prometheus, and Grafana. Web Servers: Practical understanding of Nginx (reverse proxy, load balancing, log analysis). Containerization: Foundational knowledge of Kubernetes (K8s) (understanding pods, deployments, services, and basic troubleshooting commands like kubectl logs and kubectl get pods). Cloud Platform: Familiarity with Google Cloud Platform (GCP) core services and cloud monitoring concepts. API Management: Exposure to Apigee or equivalent API gateways for monitoring traffic flow and checking endpoint health. Operating Systems: Strong command-line comfort in Linux/Unix environments for navigating directories and tailing logs. Shift Flexibility: Readiness to work in a 24/7 rotating shift model (including night shifts and weekends) to maintain uninterrupted banking infrastructure support

Location & Eligibility

Where is the job
Airoli, India
On-site at the office

Listing Details

Posted
July 10, 2026
First seen
September 25, 2026
Last seen
September 29, 2026

Posting Health

Days active
3
Repost count
0
Trust Level
20%
Scored at
September 29, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.