Senior Data Engineer

Remotesenior
Data EngineerData
0 views0 saves0 applied

Quick Summary

Key Responsibilities

Architect, build, and manage a cloud-based lakehouse environment using Apache Iceberg, AWS S3, and AWS Glue Catalog.

Requirements Summary

Experience operating and maintaining Apache Iceberg tables at enterprise scale. Experie

Technical Tools
Data EngineerData
Job Title: Senior Data Engineer
Location: Remote
Position Type: Full-Time 
Clearance Required: Secret Clearance

About the company:  
At VivSoft, we aim to solve complex federal problems using emerging and open technologies in a collaborative and rewarding environment. VivSoft is a diverse team of strategists, engineers, designers, and creators experienced in building high-performance, effective software, with a focus on impactful organisational design and software delivery dynamics. We build secure Software Factories based on DoD reference designs and NIST Frameworks for Cloud and DevSecOps. These factories deliver AI/ML Applications, Data Science Platforms, Blockchain and Microservices for DoD, Healthcare and Civilian Agencies


Job Summary:
We are seeking a Senior Data Engineer to support a United States Air Force (USAF) program responsible for building and operating a modern, scalable, and secure data platform. This role will lead the design and optimization of an enterprise lakehouse architecture on AWS, leveraging Apache Iceberg, Apache Spark, and cloud-native technologies to enable advanced analytics, AI/ML initiatives, and operational reporting. The ideal candidate will possess deep expertise in large-scale data platforms, distributed processing, data governance, and cloud infrastructure while providing technical leadership and mentoring across engineering teams.

Key Responsibilities:
  • Architect, build, and manage a cloud-based lakehouse environment using Apache Iceberg, AWS S3, and AWS Glue Catalog.
  • Develop and optimize Apache Spark data pipelines on EMR and Kubernetes environments.
  • Design and implement event-driven data ingestion solutions using S3 events and Amazon SQS.
  • Optimize query performance across Athena, Trino, and Spark SQL environments.
  • Orchestrate data workflows and ETL pipelines using Apache Airflow.
  • Manage infrastructure deployment and automation using Terraform.
  • Develop, publish, monitor, and maintain data products and analytical dashboards.
  • Implement data quality, governance, lineage, access controls, and cost management best practices.
  • Create technical documentation, architectural designs, and operational runbooks.
  • Provide technical leadership and mentorship to junior engineering team members.

Required Skills:
  • Must possess an active Secret Clearance
  • 8+ years of professional experience in Data Engineering or large-scale data platform development.
  • Expertise in Apache Spark, including performance tuning, partitioning, memory optimization, and handling data skew.
  • Strong proficiency in Python or Scala and advanced SQL development.
  • Experience with open table formats such as Apache Iceberg (preferred) or Delta Lake.
  • Strong understanding of distributed query engines, including Athena, Trino, and Spark SQL.
  • Hands-on experience with AWS services, including S3, Glue, EMR, Athena, EC2, SQS, and event-driven architectures.
  • Experience with Apache Airflow for workflow orchestration.
  • Proficiency with Terraform and Infrastructure as Code (IaC).
  • Experience working with Kubernetes environments.
  • Experience developing dashboards and data products using Grafana or similar visualization platforms.
  • Strong understanding of data quality, data governance, lineage, and access control frameworks.
  • Excellent technical leadership, mentoring, and stakeholder communication skills.
  • Ability to translate business, operational, and analytics requirements into scalable data platform solutions.
  • Strong collaboration skills with Data Scientists, AI Engineers, Cloud Engineers, and business stakeholders.
  • Proven ability to lead technical discussions and mentor junior engineers.
  • Excellent written and verbal communication skills.
  • Strong attention to detail regarding data quality, governance, lineage, security, and operational reliability.

Preferred Skills:
  • Experience operating and maintaining Apache Iceberg tables at enterprise scale.
  • Experience with Trino administration and performance optimization.
  • Knowledge of dbt or similar data transformation frameworks.
  • Experience with Helm and GitOps deployment methodologies.
  • Experience evaluating and implementing modern data visualization platforms beyond Grafana.
  • Familiarity with data catalog, metadata management, lineage, and governance tools.
  • AWS Data Analytics Certification and/or Certified Kubernetes Administrator (CKA).
  • Prior DoD, USAF, or Federal Government data platform experience.
  • Experience supporting AI/ML, data science, or advanced analytics workloads in cloud environments.

Benefits:  
  • Comprehensive Medical, Dental, and Vision Plans (Healthcare benefits are 100% employer-paid for employees only)  
  • Life Insurance  
  • Paid Time Off (Flexible/Combined PTO, Bereavement Leave, 11 Company Paid Holidays)  
  • 401K Retirement Plan with employer match  
  • Professional Development Training Reimbursement

Salary Range: $160K to $180K per Annually

Location & Eligibility

Where is the job
Worldwide
Fully remote, anywhere in the world
Who can apply
Same as job location

Listing Details

Posted
October 8, 2026
First seen
October 8, 2026
Last seen
October 8, 2026

Posting Health

Days active
0
Repost count
0
Trust Level
63%
Scored at
October 9, 2026

Signal breakdown

freshnesssource trustcontent trustemployer trust

4 other jobs at

View all →

Explore open roles at .

Newsletter

Stay ahead of the market

Get the latest job openings, salary trends, and hiring insights delivered to your inbox every week.

A
B
C
D
Join 12,000+ marketers

No spam. Unsubscribe at any time.

Senior Data Engineer