Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
Proton.ai

Senior Data Engineer – AI-Native, Data Layer

Proton.ai

Senior Data Engineer for Proton, building and operating the Data Layer for AI-driven products. Ingesting, modeling, and ensuring data reliability and correctness across various systems.

Posted 7/22/2026full-timeRemote • 🇺🇸 United StatesSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in building and operating data ingestion and transformation pipelines, ensuring data integrity and reliability across various sources. Proficient in SQL and orchestration frameworks, with a strong focus on scalable data architecture and effective communication with cross-functional teams.

Highest-signal resume keywords
Data EngineeringSQL SkillsCloud Data Warehouse ExperiencePipeline OrchestrationData Consistency Management

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Data IngestionData ModelingELT PipelinesQuery OptimizationData ValidationSchema DesignData LineageIncremental LoadsIdempotencyDebugging
Soft Skills
Strong CommunicationOwnershipJudgmentStartup Mindset
Tools & Technologies
Claude CodeCursor Agent ModeCodexCloud PlatformsOrchestration Frameworks
Industry Keywords
Medallion ModelData ContractsBatch FilesStreaming EventsAPI-based SourcesData Consistency Failure Modes

Tech Stack

Tools & technologies
CloudSQL

About the role

Key responsibilities & impact
  • Own the Data Layer end to end: ingestion from file-, event-, and API-based sources; the medallion-style model (raw → refined → curated); and the serving layer that powers the product and the AI brain.
  • Build and operate the ingestion and transformation pipelines that power the Data Layer, using a modern orchestration framework and cloud data warehouse.
  • Ingest and reconcile large, messy, real-world data across many source types and shapes — batch files, streaming events, and APIs.
  • Model data across medallion layers so it's trustworthy, queryable, and stable for downstream teams and the AI.
  • Help take the Data Layer to the next level — better architecture, better tooling, more scale, more sources — and have a real say in what that looks like.
  • Operate AI coding agents (Claude Code and similar) at a high level: scope work, structure context, run agents in parallel where it makes sense, and ship reviewed, production-quality output.
  • Build the systems that make data trustworthy — validation, reconciliation, lineage, backfills, idempotent and incremental loads — so downstream teams and the AI don't inherit silent errors.
  • Partner with backend, AI, and product engineers (and occasionally customers' IT teams) to define the data contracts they build on.

Requirements

What you’ll need
  • 7+ years hands-on as a data engineer with real, demonstrable production ownership — pipelines and data models serving real users at scale.
  • Strong fundamentals. You understand what your code and your queries are doing and why. You can read a query plan, reason about a slow or expensive pipeline, and debug a data-correctness bug to its root.
  • Strong programming and SQL skills. You build efficient pipelines, schemas, and queries, and can model data for both transactional and analytical access patterns.
  • Hands-on orchestration experience, building reliable ingestion/ELT pipelines against messy upstream sources.
  • Experience with a cloud data warehouse and a major cloud platform.
  • Experience ingesting from multiple source types: file-based, event/streaming, and API-based.
  • Solid grasp of data-consistency failure modes — partial loads, late or out-of-order data, idempotency, backfills, schema drift.
  • Daily, hands-on use of agentic dev tools (Claude Code, Cursor agent mode, Codex, or equivalent) to ship real work. You can talk concretely about how you structure prompts, manage context, parallelize agents, and verify their output.
  • Ownership and judgment. You take data systems from idea to production and exercise good taste on what to build and what to cut.
  • Startup mindset and strong communication — pragmatic, fast, biased to ship, and able to explain data decisions to engineers, PMs, and customers in writing.
  • English at C1 or above.

Benefits

Comp & perks
  • Professional development opportunities