FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in data-platform engineering, including architecting streaming and batch pipelines, and managing identity data models. Proficient in leveraging AI-assisted engineering tools and collaborating with cross-functional teams to drive technical direction and improvements.
Highest-signal resume keywords
Data-Platform EngineeringJava ProficiencyAWS Data PrimitivesSpark/EMR ExperienceAI-Assisted Engineering Tools
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data Identity ArchitectureStreaming PipelinesBatch PipelinesCross-Team Data ContractsSchema VersioningDeduplication TechniquesIncident ResponseCost ManagementTechnical DirectionDistributed Storage
Soft Skills
CollaborationComfort with AmbiguityInfluencing Peers
Tools & Technologies
HBaseDatabricksAWS EMRAWS S3AWS SNSAWS SQSClaudeCycloneDXVEXSPDX
Industry Keywords
Software EngineeringAgile EnvironmentTechnical DiscussionsVendor EngineeringData Quality Signals
Tech Stack
Tools & technologiesAWSHBaseJavaSpark
About the role
Key responsibilities & impact- Set technical direction for the Data Identity team’s next generation
- Partner with the team lead and product on multi-quarter architectural decisions
- Own the identity data model, including coordinate schemas, cross-ecosystem mapping, deduplication, and re-discovery
- Architect streaming and batch pipelines using Java, Spark/EMR, HBase, Databricks, and AWS data primitives
- Drive migrations, cost improvements, reliability improvements, and infrastructure choices
- Design stable, versioned data contracts for downstream teams
- Lead ingestion of third-party SBOM feeds, including CycloneDX, VEX, and SPDX
- Represent Sonatype in technical discussions with vendor engineering teams
- Explore AI/ML applications for identity re-discovery, package deduplication, and data-quality signals
- Use AI-assisted engineering tools for coding, code review, investigation, and data exploration
- Lead alarm triage, SLO discipline, cost management, and incident response
- Influence peer teams and collaborate with cross-functional stakeholders
Requirements
What you’ll need- 10+ years of professional software engineering experience
- 3+ years at Staff level or equivalent scope
- Deep experience with data-platform engineering at scale, including streaming and batch pipelines and distributed storage
- Experience with Spark, HBase, Databricks, or comparable systems
- Strong Java proficiency or a systems-programming background enabling rapid productivity
- Experience designing and coordinating cross-team data contracts, schema versioning, and backward compatibility
- AWS-native operating experience with EMR, S3, SNS, and SQS
- Comfort with ambiguity and vendor-facing technical work
- Working fluency with AI-assisted engineering tools such as Claude or equivalent
- Ability to judge when to apply LLMs and classical ML to data-pipeline problems
- Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience
- Experience in an agile environment and cross-functional collaboration
Benefits
Comp & perks- Parental leave
- Diversity and inclusion working groups
- Flexible working practices
- Paid Volunteer Time Off (VTO)
- Equal-opportunity employer
- Disability or special-needs accommodation
