FREE ACCESS
5,000–10,000 jobs/day
See all jobs on JobTailor
Search thousands of fresh jobs every day.
Discover
- Fresh listings
- Fast filters
- No subscription required
Create a free account and start exploring right away.

Research Engineer, Curator
TransPerfectResearch Engineer (RE) Curator at DataForce by TransPerfect, creating evaluation benchmarks for AI systems. Involves Python proficiency and experimental research collaboration.
Core Competencies
Role fitCore Competencies
Use this summary to align your resume positioning with the role.
Demonstrates expertise in designing and implementing AI evaluation benchmarks and dataset frameworks, with strong proficiency in Python programming and experience in machine learning and model testing. Capable of conducting systematic research and collaborating effectively with cross-functional teams to translate complex concepts into actionable insights.
Highest-signal resume keywords
Python ProgrammingMachine LearningAI EvaluationBenchmark CurationResearch Mindset
ATS Keywords
Tailor your resumeApplicant Tracking System Keywords
Tip: use these terms in your resume and cover letter to boost ATS matches.
Hard Skills
Data AnalysisExperimental ResearchModel Testing FrameworksLLMsScalable Code Writing
Soft Skills
Analytical CapabilitiesProblem-Solving
Tools & Technologies
JupyterColabGitModern IDEs
Certifications & Qualifications
MScPhD
Industry Keywords
AIMLComputer ScienceStatisticsMathematicsPhysicsComputational Sciences
Tech Stack
Tools & technologiesPython
About the role
Key responsibilities & impact- Design, curate, and implement complex AI evaluation benchmarks and dataset frameworks to test advanced LLM capabilities and limitations
- Conduct experimental research, red teaming, and systematic testing to identify model failure modes, edge cases, and reasoning gaps
- Utilize Python, Jupyter/Colab environments, and modern IDEs to write scalable code, process datasets, and automate evaluation pipelines
- Collaborate with AI research and engineering teams to translate complex subject-matter concepts into actionable model benchmarks
- Document research findings, track codebase and dataset changes via Git, and deliver precise sourcing and evaluation reports.
Requirements
What you’ll need- Successful completion of a role-specific coding/research assessment and background check
- MSc or PhD in a STEM discipline (Computer Science, AI/ML, Statistics, Mathematics, Physics, Computational Sciences, or related field)
- Proven background as a Research Engineer, Applied Scientist, ML Engineer, Research Scientist, or AI Researcher
- Strong hands-on Python programming skills, with daily fluency in Git, modern IDEs, and Jupyter/Colab environments
- Core experience in Machine Learning, LLMs, data analysis, and structured experimental research
- Direct experience in AI evaluation, benchmark curation/development, red teaming, or model testing frameworks (preferred)
- Strong research mindset paired with exceptional analytical and problem-solving capabilities.
Benefits
Comp & perks- Flexible work arrangements