Quick Summary
Description Builds and runs large scale batch data pipelines, keeping jobs fast and affordable as data volumes grow. Works with Apache Spark to build and manage large scale data processing jobs,
Builds and runs large scale batch data pipelines, keeping jobs fast and affordable as data volumes grow. Works with Apache Spark to build and manage large scale data processing jobs, including performance tuning to maintain efficient and reliable pipeline execution. Supports data modelling for analytics and organises data in a lakehouse so it is usable downstream. Handles automated scheduling, monitoring, and data quality checks, while working with platform and product teams on end to end data flows.
Requirements
~1 min read- Strong hands on Apache Spark including performance tuning, not Spark usage through a managed notebook only.
- Data modelling for analytics and organising data in a lakehouse so it is usable downstream.
- Automated scheduling, monitoring, and data quality checks.
- Works with platform and product teams on end to end data flows.
- Apache Iceberg or other open table formats.
- Trino or similar query engines.
- On premises or self managed cluster experience.
Location & Eligibility
Listing Details
- Posted
- September 23, 2026
- First seen
- September 28, 2026
- Last seen
- September 28, 2026
Posting Health
- Days active
- 0
- Repost count
- 0
- Trust Level
- 38%
- Scored at
- September 28, 2026
Signal breakdown
Stay ahead of the market
Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.
No spam. Unsubscribe at any time.