Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
TransPerfect

Research Engineer, Curator

TransPerfect

Research Engineer (RE) Curator at DataForce by TransPerfect, creating evaluation benchmarks for AI systems. Involves Python proficiency and experimental research collaboration.

Posted 7/25/2026contractRemote • 🇺🇸 United StatesMid-LevelSeniorWebsite

Core Competencies

Role fit
Core Competencies

Use this summary to align your resume positioning with the role.

Demonstrates expertise in designing and implementing AI evaluation benchmarks and dataset frameworks, with strong proficiency in Python programming and experience in machine learning and model testing. Capable of conducting systematic research and collaborating effectively with cross-functional teams to translate complex concepts into actionable insights.

Highest-signal resume keywords
Python ProgrammingMachine LearningAI EvaluationBenchmark CurationResearch Mindset

ATS Keywords

Tailor your resume
Applicant Tracking System Keywords

Tip: use these terms in your resume and cover letter to boost ATS matches.

Hard Skills
Data AnalysisExperimental ResearchModel Testing FrameworksLLMsScalable Code Writing
Soft Skills
Analytical CapabilitiesProblem-Solving
Tools & Technologies
JupyterColabGitModern IDEs
Certifications & Qualifications
MScPhD
Industry Keywords
AIMLComputer ScienceStatisticsMathematicsPhysicsComputational Sciences

Tech Stack

Tools & technologies
Python

About the role

Key responsibilities & impact
  • Design, curate, and implement complex AI evaluation benchmarks and dataset frameworks to test advanced LLM capabilities and limitations
  • Conduct experimental research, red teaming, and systematic testing to identify model failure modes, edge cases, and reasoning gaps
  • Utilize Python, Jupyter/Colab environments, and modern IDEs to write scalable code, process datasets, and automate evaluation pipelines
  • Collaborate with AI research and engineering teams to translate complex subject-matter concepts into actionable model benchmarks
  • Document research findings, track codebase and dataset changes via Git, and deliver precise sourcing and evaluation reports.

Requirements

What you’ll need
  • Successful completion of a role-specific coding/research assessment and background check
  • MSc or PhD in a STEM discipline (Computer Science, AI/ML, Statistics, Mathematics, Physics, Computational Sciences, or related field)
  • Proven background as a Research Engineer, Applied Scientist, ML Engineer, Research Scientist, or AI Researcher
  • Strong hands-on Python programming skills, with daily fluency in Git, modern IDEs, and Jupyter/Colab environments
  • Core experience in Machine Learning, LLMs, data analysis, and structured experimental research
  • Direct experience in AI evaluation, benchmark curation/development, red teaming, or model testing frameworks (preferred)
  • Strong research mindset paired with exceptional analytical and problem-solving capabilities.

Benefits

Comp & perks
  • Flexible work arrangements