Sr. Lead Software Engineer - Performance and Resiliency Engineering -
Quick Summary
Lead performance engineering efforts: load/stress/soak testing, capacity modeling, performance tuning, and regression prevention (KPIs, guardrails, and acceptance criteria).
Performance and Resiliency Engineering Lead,
ATTENTION MILITARY AFFILIATED JOB SEEKERS - Our organization works with partner companies to source qualified talent for their open roles. The following position is available to Veterans, Transitioning Military, National Guard and Reserve Members, Military Spouses, Wounded Warriors, and their Caregivers. If you have the required skill set, education requirements, and experience, please click the submit button and follow the next steps. All positions are onsite, unless otherwise stated.
Job Description:
Performance and Resiliency Engineering Lead, Merchant Services (Senior Vice President or Executive Director)
Merchant Services is hiring an SVP/ED Performance & Resiliency Engineer to improve the performance and stability of critical, high-volume platforms. The role focuses on latency reduction, throughput/TPS improvements, capacity planning, and peak-event readiness, while also strengthening resiliency and release safety.
You will partner closely with application development teams and infrastructure/platform engineering to identify bottlenecks across the stack (application/runtime, database, network, compute, and platform), implement durable fixes, and raise engineering standards through technical leadership, mentorship, and strong cross-team collaboration.
Key Responsibilities:
- Lead performance engineering efforts: load/stress/soak testing, capacity modeling, performance tuning, and regression prevention (KPIs, guardrails, and acceptance criteria).
- Improve production readiness: observability (metrics/logs/traces/APM), actionable alerting, incident triage, and root-cause analysis leading to durable remediation.
- Strengthen resiliency patterns: timeouts/retries, circuit breakers, backpressure/rate limiting, graceful degradation, and failover readiness.
- Drive safer releases via canary/progressive delivery and automated rollback patterns.
- Optimize containerized workloads across EKS (primary) and ECS/other compute where applicable; drive autoscaling strategy and right-sizing.
- Senior experience in performance engineering for distributed systems and/or SRE-style reliability engineering in production.
- Strong cloud/container background (AWS + Kubernetes/EKS; ECS exposure beneficial).
- Experience with modern observability tooling (e.g., Datadog, Dynatrace, Grafana, OpenTelemetry, CloudWatch or equivalent).
- KEDA and/or Karpenter: large plus.
- Akamai: strongly preferred.
- Demonstrated ability to lead through influence, mentor engineers, and work effectively across teams.
Location & Eligibility
Listing Details
- Posted
- October 6, 2026
- First seen
- October 6, 2026
- Last seen
- October 6, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 57%
- Scored at
- October 6, 2026
Signal breakdown
Similar Lead Software Engineer jobs
View all →Browse Similar Jobs
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.