FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Data Engineer – AI, Java, Python, Spark
Infotree Global SolutionsSenior Data & GenAI Engineer building scalable data, cloud and LLM platforms for Infotree Global Solutions. Designing pipelines, distributed systems and agentic AI applications in Warsaw.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in developing and maintaining high-quality software, with a strong focus on building scalable cloud-native services and optimizing distributed data processing. Proficient in Python, Java, and modern data platforms, with a solid understanding of GenAI applications and engineering best practices.
Highest-signal resume keywords
Python DevelopmentJava DevelopmentApache SparkKubernetesGenAI Applications
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Software EngineeringData EngineeringDistributed SystemsData Pipeline DesignPerformance OptimizationTesting and Code QualitySystem DesignML/AI DevelopmentLLM OrchestrationCloud-Native Technologies
Soft Skills
Technical LeadershipCollaboration
Tools & Technologies
DatabricksSnowflakeLangChainLangGraph
Industry Keywords
Cloud Data WarehousesLakehouse PlatformsDistributed Processing SystemsEngineering Best Practices
Tech Stack
Tools & technologiesApacheCloudDistributed SystemsJavaKubernetesPythonSpark
About the role
Key responsibilities & impact- Develop, test and maintain high-quality, production-ready software
- Design and implement large-scale data pipelines and distributed processing systems
- Build scalable cloud-native services and platforms
- Provide technical leadership for cross-team initiatives and complex engineering projects
- Design and develop reusable libraries, frameworks and platform components
- Optimize distributed data processing workloads for performance, scalability and reliability
- Work with Databricks, Apache Spark and Snowflake data platforms
- Develop and deploy applications using Python and/or Java
- Build and operate containerized workloads using Kubernetes and cloud-native technologies
- Design and implement GenAI/LLM-based applications and services
- Use LangChain and LangGraph for LLM orchestration and agentic workflows
- Collaborate with data scientists, software engineers, architects and product teams
- Establish engineering best practices around testing, observability, reliability and deployment
Requirements
What you’ll need- 5+ years of professional software/data engineering experience
- Strong hands-on experience with Python and/or Java
- Strong experience with Apache Spark and distributed data processing
- Experience with Databricks and/or modern lakehouse platforms
- Experience with Snowflake or comparable cloud data warehouses
- Practical experience with Kubernetes and cloud-native technologies
- Experience designing and maintaining large-scale data pipelines
- Strong understanding of distributed systems, scalability and production engineering
- Experience developing ML/AI or GenAI applications
- Experience with LLM-based applications, RAG, AI agents or LLM orchestration
- Familiarity with LangChain, LangGraph or similar GenAI frameworks
- Strong software engineering fundamentals including testing, code quality and system design
Benefits
Comp & perks- Supportive team focused on employee growth and well-being
- Opportunities to grow and challenge yourself in the role
- Dedicated recruiter and consultant care representative support