We are looking for a Senior Data Software Engineer to help deliver and optimize an Apache Iceberg-based data lakehouse on AWS while strengthening platform automation and deployment workflows. You will collaborate closely with senior engineers to improve reliability and governance—apply now to build scalable data foundations.
Responsibilities
- Build and optimize an Apache Iceberg lakehouse on Amazon Web Services
- Execute Iceberg optimization techniques including partitioning, compaction, snapshot retention, orphan-file removal, and schema evolution
- Improve data quality and reliability across Bronze-layer landing zones
- Enhance the configuration-management framework used across the lakehouse platform
- Build self-service tenant-onboarding automation for landing zones, IAM, and monitoring
- Maintain the automated GitHub-to-S3 deployment pipeline
- Contribute to PR-based validation workflows and uphold review standards
- Collaborate with senior engineers to align implementations with platform architecture
- Evaluate Iceberg-to-Snowflake integration options and document trade-offs
Requirements
- 3+ years of experience in data engineering with Amazon Web Services
- Experience building data lakehouse solutions in production
- Experience implementing Apache Iceberg features and optimization techniques
- Ability to collaborate under senior guidance and align with defined platform architecture
- Strong project execution skills in a staff augmentation engagement model
- Advanced knowledge of partitioning, compaction, snapshot retention, and schema evolution in Iceberg
- Strong automation skills for configuration management, onboarding workflows, and CI/CD-style deployments
- Strong data quality and reliability skills across landing zones and Bronze-layer patterns
- Effective communication skills for documentation, trade-off analysis, and code review participation
- Upper-Intermediate English proficiency (B2)
Nice to have
- Snowflake integration experience
- AWS Glue development experience