1d ago
New

Linux SME / SRE Engineer

IndiaIndiaยทHyderabadFull-timemid
OtherSre Engineer
0 views0 saves0 applied

Quick Summary

Technical Tools
OtherSre Engineer

๐—ง๐—ต๐—ถ๐˜€ ๐—ฟ๐—ผ๐—น๐—ฒ ๐—ถ๐˜€ ๐—ณ๐—ผ๐—ฟ ๐—ผ๐—ป๐—ฒ ๐—ผ๐—ณ ๐˜๐—ต๐—ฒ ๐—ช๐—ฒ๐—ฒ๐—ธ๐—ฑ๐—ฎ๐˜†'๐˜€ ๐—ฐ๐—น๐—ถ๐—ฒ๐—ป๐˜๐˜€

๐—ฆ๐—ฎ๐—น๐—ฎ๐—ฟ๐˜† ๐—ฟ๐—ฎ๐—ป๐—ด๐—ฒ: ๐—ฅ๐˜€ ๐Ÿฎ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ - ๐—ฅ๐˜€ ๐Ÿฏ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ๐Ÿฌ (๐—ถ๐—ฒ ๐—œ๐—ก๐—ฅ ๐Ÿฎ๐Ÿฌ-๐Ÿฏ๐Ÿฌ ๐—Ÿ๐—ฃ๐—”)

Experience: 5+ yrs

Location: Bengaluru, Karnataka, India, Hyderabad, Telangana, India

Job Type: Full-time

We are looking for an experienced Linux SME / SRE Engineer with strong expertise in Core Linux Administration, RHEL, and PCS/Pacemaker clustering to support business-critical production environments.

The role focuses on maintaining highly available Linux infrastructure, resolving complex production issues, ensuring system reliability, and supporting clustered environments. The ideal candidate will have strong hands-on troubleshooting capabilities, a solid understanding of high-availability architectures, and the ability to work effectively with clients and technical stakeholders.

Requirements

~2 min read

Key Responsibilities

  • Administer and support Linux-based production environments across critical infrastructure.
  • Perform day-to-day Core Linux administration, configuration, monitoring, maintenance, and troubleshooting.
  • Manage, monitor, configure, and troubleshoot PCS/Pacemaker high-availability clusters.
  • Ensure availability, reliability, stability, and performance of Linux infrastructure and clustered services.
  • Troubleshoot complex and critical production incidents and drive issues through to resolution.
  • Perform root-cause analysis and implement sustainable solutions for recurring infrastructure problems.
  • Monitor system and cluster health and proactively identify potential availability or performance issues.
  • Support failover, recovery, maintenance, and operational activities across high-availability environments.
  • Collaborate with clients, infrastructure teams, application teams, and other technical stakeholders on incidents and enhancements.
  • Participate in incident management, problem management, change management, and production maintenance activities.
  • Follow SRE practices for monitoring, reliability improvement, incident response, and operational efficiency.
  • Maintain technical documentation, operational procedures, troubleshooting guides, and support records.
  • Participate in rotational shifts to provide continuous production support.
  • Identify opportunities to automate repetitive infrastructure tasks and improve operational efficiency.
  • Support infrastructure changes, upgrades, patching, and maintenance activities in accordance with established processes.
  • Contribute to service reliability, availability, and continuous improvement initiatives.

What Makes You a Great Fit

  • 5โ€“9 years of overall experience in Linux administration, infrastructure engineering, SRE, or production support, with a maximum of 10 years preferred.
  • Minimum 4 years of hands-on experience with PCS/Pacemaker cluster administration.
  • Strong expertise in Core Linux Administration and production infrastructure support.
  • Strong hands-on experience with RHEL (Red Hat Enterprise Linux).
  • Solid understanding of High Availability, clustering, failover, resource management, and cluster troubleshooting.
  • Proven experience supporting critical production environments with strict availability and reliability requirements.
  • Strong troubleshooting, debugging, root-cause analysis, and incident-resolution capabilities.
  • Experience working with production monitoring, incident management, and infrastructure maintenance processes.
  • Strong understanding of SRE and ITIL practices is desirable.
  • Excellent communication and client-facing skills with the ability to explain technical issues clearly to stakeholders.
  • Strong stakeholder-management and collaboration skills.
  • Ability to work effectively under pressure during critical production incidents.
  • Willingness to work in rotational shifts, including scheduled production-support coverage.
  • Experience with VMware administration is an advantage.
  • Exposure to AWS or other cloud platforms is desirable.
  • Knowledge of Oracle Database and its infrastructure dependencies is an advantage.
  • Strong ownership mindset with a focus on system reliability, operational excellence, and continuous improvement.

Location & Eligibility

Where is the job
Hyderabad, India
On-site at the office

Listing Details

Posted
September 28, 2026
First seen
September 29, 2026
Last seen
September 29, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
54%
Scored at
September 29, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust
Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

Linux SME / SRE Engineer