Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Guidehouse

AWS Lakehouse Data Engineer

Guidehouse

AWS Lakehouse Data Engineer building Guidehouse’s S3 and Iceberg data platform for AI, analytics, reporting, and visualization. Automating ingestion, governance, CI/CD, and cloud operations.

Posted 8/31/2026full-timeRemote • 🇺🇸 United StatesMid-LevelSenior💰 $113,000 - $188,000 per yearWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and implementing cloud-native data platforms, with a strong focus on AWS-native services, ETL/ELT pipeline development, and data governance. Proficient in optimizing performance, reliability, and security of data workflows while ensuring compliance with best practices.

Highest-signal resume keywords
AWS-Native Data Lake ArchitecturePython ETL/ELT Pipeline DevelopmentApache Iceberg ImplementationMetadata Management and GovernanceCI/CD for Data Workflows

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
PythonPySparkSQLETLELTData ModelingData TransformationAWS GlueAmazon S3Apache Parquet
Soft Skills
CollaborationCommunicationProblem-Solving
Tools & Technologies
Amazon AthenaAmazon EMRAWS Lake FormationAWS RedshiftInfrastructure as Code
Industry Keywords
Data GovernanceData QualityCloud-NativeData VisualizationOperational Runbooks

Tech Stack

Tools & technologies
Amazon RedshiftApacheAWSCloudETLPySparkPythonSDLCSQL

About the role

Key responsibilities & impact
  • Design, implement, and operate a cloud-native data platform powering AI/ML, analytics, reporting, and data visualization
  • Design and implement batch and streaming ingestion from APIs, relational databases, file drops, event streams, and external partners
  • Implement, test, and optimize Python and PySpark ETL/ELT pipelines for curated analytics-ready datasets
  • Implement incremental processing, change data capture, data contracts, schema validation, and reusable transformation frameworks
  • Improve pipeline reliability through automated testing, orchestration, monitoring, retry handling, and operational runbooks
  • Design and implement a Delta Lakehouse-style platform using AWS-native services
  • Build and manage a scalable lakehouse on Amazon S3 using Apache Iceberg and Apache Parquet
  • Implement ACID transactions, schema evolution, partition evolution, snapshot isolation, time travel, and rollback capabilities
  • Enable fast interactive querying using Amazon Athena, Amazon EMR, AWS Glue, and Amazon Redshift
  • Optimize performance and cost through partitioning, compaction, file sizing, statistics, caching, lifecycle policies, and compute-storage separation
  • Establish standardized development, test, and production environments with controlled promotion
  • Implement governance, metadata management, lineage, fine-grained access control, classification, retention, encryption, and secure data handling
  • Build operational data quality checks and publish measurable SLAs/SLOs
  • Implement AWS provisioning with Infrastructure as Code and secure-by-default baselines
  • Build and enhance CI/CD for data pipelines and lakehouse components, including testing, security checks, deployment, promotion, and rollback
  • Implement observability with metrics, logs, traces, alerts, dashboards, runbooks, and incident-response procedures
  • Evaluate and improve platform performance, scalability, reliability, security, and cost
  • Collaborate with data, application, analytics, AI/ML, security, networking, and cloud platform teams
  • Maintain architecture diagrams, data models, SOPs, interface specifications, runbooks, and secure configuration baselines
  • Present technical findings, trade-offs, risks, and recommendations to technical and non-technical stakeholders

Requirements

What you’ll need
  • Bachelor's degree in Engineering, Information Technology, Computer Science, Data Engineering, or a related field, or FOUR (4) years equivalent practical experience in lieu of degree
  • SIX (6) years of relevant experience
  • Hands-on experience implementing AWS-native data lake or lakehouse architectures using Amazon S3 and services such as AWS Glue, Amazon Athena, Amazon EMR, AWS Lake Formation, and Amazon Redshift
  • Strong experience developing production ETL/ELT pipelines using Python and PySpark, including data modeling, transformation, testing, performance tuning, and error handling
  • Hands-on experience with Apache Iceberg, including ACID transactions, snapshots, schema and partition evolution, time travel, table maintenance, and query optimization
  • Advanced SQL skills and experience supporting analytical queries, semantic layers, reporting tools, and data visualization workloads
  • Experience implementing metadata management and governance capabilities, including cataloging, lineage, ownership, classification, policy enforcement, and fine-grained access controls
  • Experience with AWS security fundamentals, including IAM and least privilege, KMS encryption, secrets management, network security, logging, and secure SDLC practices
  • Experience provisioning AWS resources using IaC and operating data platforms across multiple environments
  • Experience building or operating CI/CD pipelines for data workflows, including testing, packaging, deployment automation, environment promotion, and rollback
  • Ability to troubleshoot distributed data-processing workloads and optimize performance, reliability, and cost
  • Ability to obtain Public Trust clearance

Benefits

Comp & perks
  • Medical, Rx, Dental & Vision Insurance
  • Personal and Family Sick Time & Company Paid Holidays
  • Parental Leave
  • 401(k) Retirement Plan
  • Group Term Life and Travel Assistance
  • Voluntary Life and AD&D Insurance
  • Health Savings Account, Health Care & Dependent Care Flexible Spending Accounts
  • Transit and Parking Commuter Benefits
  • Short-Term & Long-Term Disability
  • Tuition Reimbursement, Personal Development, Certifications & Learning Opportunities
  • Employee Referral Program
  • Corporate Sponsored Events & Community Outreach
  • Care.com annual membership
  • Employee Assistance Program
  • Supplemental Benefits via Corestream (Critical Care, Hospital Indemnity, Accident Insurance, Legal Assistance and ID theft protection, etc.)
  • Position may be eligible for a discretionary variable incentive bonus
  • Flexible benefits package