Apply

Ready to go for it?

AI Apply speeds things up—apply directly if you prefer.

FREE ACCESS
5,000–10,000 jobs/day
JobTailor Logo

See all jobs on JobTailor

Search thousands of fresh jobs every day.

Discover
  • Fresh listings
  • Fast filters
  • No subscription required
Create a free account and start exploring right away.
InnoData

Applied Research Scientist, LLM Evaluation – Post-Training

InnoData

Applied Research Scientist at Innodata advancing LLM evaluation and post-training methods. Collaborating on evaluation design, measurement strategies, and feedback signals for model improvement.

Posted 7/27/2026full-timeRemote • 🇺🇸 United StatesMid-LevelSenior💰 $175,000 - $225,000 per yearWebsite

Tech Stack

Tools & technologies
PythonPyTorchTensorflow

About the role

Key responsibilities & impact
  • Lead research and experimentation on how evaluation design, measurement strategies, and feedback signals influence model improvement.
  • Help define the next generation of evaluation-driven model improvement workflows.
  • Study how different evaluation approaches (human, automated, hybrid) shape model selection and post-training outcomes.
  • Design experiments that produce credible, actionable conclusions.
  • Support customer engagements by bringing scientific rigor to evaluation strategy, methodology review, and technical recommendations.
  • Define and execute a research agenda focused on LLM evaluation and post-training, especially evaluation-driven model improvement.
  • Design rigorous experiments to study how evaluation methodologies impact fine-tuning and post-training outcomes.
  • Develop and validate evaluation frameworks for LLM and multimodal systems.
  • Analyze model behavior and failure patterns; generate actionable recommendations for model improvement and evaluation redesign.
  • Collaborate with AI/ML Research Engineers to translate research methods into scalable evaluation and post-training pipelines.

Requirements

What you’ll need
  • MS/PhD in Computer Science, Machine Learning, Statistics, Applied Mathematics, AI, or a related quantitative scientific field (PhD strongly preferred)
  • 5+ years of relevant experience in applied research / research science in ML/AI, with substantial work in LLMs or foundation models
  • Demonstrated experience with LLM evaluation, benchmarking, alignment, post-training, or model quality research
  • Strong foundation in experimental design, statistical analysis, and scientific reasoning for ML systems
  • Strong coding skills in Python for research experimentation and analysis (e.g., data processing, evaluation pipelines, statistical analysis, visualization)
  • Experience working with modern ML tooling/frameworks (e.g., PyTorch, Hugging Face, JAX/TensorFlow as applicable) sufficient to design and execute model/evaluation experiments
  • Ability to evaluate and compare human and automated evaluation methods, including tradeoffs in cost, reliability, validity, and scalability
  • Experience designing evaluation studies and protocols that are reproducible across datasets, model versions, and evaluation runs
  • Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data scientists, and customer technical counterparts
  • Strong communication skills and ability to present nuanced technical conclusions, assumptions, and limitations clearly.

Benefits

Comp & perks
  • 🌐 Worldwide ❌ Jobs You've Hidden ⭐️ Saved Jobs ✅ Applied Jobs ✉️ Email Alerts 👤 Account InnoData Website LinkedIn All Job Openings 2 - 10 employees Founded 2019 🤝 B2B 💼 Consulting 🌍 Social Impact B2B
  • Consulting
  • Social Impact InnoData is INNOvation DATA SCS, an Italian social cooperative based in Foggia that identifies itself as a provider of technological solutions. The company website (currently under maintenance) highlights "Soluzioni tecnologiche" (technological solutions) and emphasizes social impact ("Impatto sociale"). Contact details listed include Via Francesco Crispi 65, 71121 Foggia, Italy. Based on the available information, InnoData appears to operate at the intersection of technology and social impact, likely offering tech-focused services to other organizations. Applied Research Scientist, LLM Evaluation – Post-Training Job not on LinkedIn 🔥 25 minutes ago 🇺🇸 United States – Remote 💵 $175k - $225k / year ⏰ Full Time 🟡 Mid-level 🟠 Senior 🧬 Research Scientist Apply Now Find Hiring Managers Customize resume + cover letter Report problem ☆ Save ☑️ Mark as applied ❌ Hide 📋 Description
  • Lead research and experimentation on how evaluation design, measurement strategies, and feedback signals influence model improvement.
  • Help define the next generation of evaluation-driven model improvement workflows.
  • Study how different evaluation approaches (human, automated, hybrid) shape model selection and post-training outcomes.
  • Design experiments that produce credible, actionable conclusions.
  • Support customer engagements by bringing scientific rigor to evaluation strategy, methodology review, and technical recommendations.
  • Define and execute a research agenda focused on LLM evaluation and post-training, especially evaluation-driven model improvement.
  • Design rigorous experiments to study how evaluation methodologies impact fine-tuning and post-training outcomes.
  • Develop and validate evaluation frameworks for LLM and multimodal systems.
  • Analyze model behavior and failure patterns; generate actionable recommendations for model improvement and evaluation redesign.
  • Collaborate with AI/ML Research Engineers to translate research methods into scalable evaluation and post-training pipelines. 🎯 Requirements
  • MS/PhD in Computer Science, Machine Learning, Statistics, Applied Mathematics, AI, or a related quantitative scientific field (PhD strongly preferred)
  • 5+ years of relevant experience in applied research / research science in ML/AI, with substantial work in LLMs or foundation models
  • Demonstrated experience with LLM evaluation, benchmarking, alignment, post-training, or model quality research
  • Strong foundation in experimental design, statistical analysis, and scientific reasoning for ML systems
  • Strong coding skills in Python for research experimentation and analysis (e.g., data processing, evaluation pipelines, statistical analysis, visualization)
  • Experience working with modern ML tooling/frameworks (e.g., PyTorch, Hugging Face, JAX/TensorFlow as applicable) sufficient to design and execute model/evaluation experiments
  • Ability to evaluate and compare human and automated evaluation methods, including tradeoffs in cost, reliability, validity, and scalability
  • Experience designing evaluation studies and protocols that are reproducible across datasets, model versions, and evaluation runs
  • Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data scientists, and customer technical counterparts
  • Strong communication skills and ability to present nuanced technical conclusions, assumptions, and limitations clearly. Apply Now 📊 Check your resume score for this job Improve your chances of getting an interview by checking your resume score before you apply. Check Resume Score Similar Jobs Research Scientist IV, Information Science 🔥 6 hours ago Arizona 201 - 500 📣 Marketing 📱 Media Website LinkedIn All Job Openings Research Scientist IV developing and evaluating AI/NLP systems for scientific feasibility assessment. Leading design and implementation in a collaborative research environment with Python expertise. 🇺🇸 United States – Remote ⏰ Full Time 🟡 Mid-level 🟠 Senior 🧬 Research Scientist Senior RTI Researcher 🔥 7 hours ago hermeneutic Investments 11 - 50 💸 Finance Website LinkedIn All Job Openings Senior RTI Researcher focusing on real-time information in crypto markets. Collaborating with traders to identify news-driven trading opportunities and building alert systems. 🇺🇸 United States – Remote ⏰ Full Time 🟠 Senior 🧬 Research Scientist Research Assistant – Floater, Cross-Border Travel Required 🔥 7 hours ago Centricity Research 201 - 500 🏥 Healthcare 💊 Pharmaceuticals 🤝 B2B Website LinkedIn All Job Openings Research Assistant Floater managing data entry and administrative tasks for clinical trials across U.S. and Canadian sites. In-person support required at various research locations along with remote work capabilities. 🇺🇸 United States – Remote 💵 $21 - $24 / hour ⏰ Full Time 🟡 Mid-level 🟠 Senior 🧬 Research Scientist Research Assistant, Technology in Education 🔥 11 hours ago JHU EEHPC 11 - 50 📚 Education 🤖 Artificial Intelligence 🏥 Healthcare Website LinkedIn All Job Openings Research Assistant overseeing data collection and management for research studies in educational technology. Collaborating with faculty and research teams to support data accuracy and reporting. 🇺🇸 United States – Remote 💵 $17 - $30 / hour ⏰ Full Time 🟡 Mid-level 🟠 Senior 🧬 Research Scientist Senior Market Researcher 🕒 2 days ago HubSpot 1001 - 5000 🤝 B2B ☁️ SaaS 📣 Marketing Website LinkedIn All Job Openings Senior Market Researcher leading quantitative and mixed-method research at HubSpot. Collaborating with teams to derive insights for product strategies and market opportunities. 🇺🇸 United States – Remote 💵 $112k - $168k / year ⏰ Full Time 🟠 Senior 🧬 Research Scientist 🦅 H1B Visa Sponsor View More Research Scientist Jobs 🌐 Worldwide Built by Lior Neu-ner. I'd love to hear your feedback — Get in touch via DM or support@remoterocketship.com Search Search Jobs by country Search jobs by city Search jobs by job title Search entry-level jobs Search junior-level jobs Search senior-level jobs Search jobs by tech stack Search jobs by contract type Search remote internships Search remote part-time jobs Remote jobs Anywhere in the World Companies Hiring Anywhere in the World Companies Hiring Sales People Anywhere in the World Companies Hiring Software Engineers Anywhere in the World Resources Advice Tips for finding remote jobs Interview questions and answers Resume examples Cover letter examples Post a job Affiliates Is Remote Rocketship legit? Privacy policy Terms of service Job board SEO course OpenClaw job finder Find jobs using your resume Jobs by Country Remote jobs anywhere in the world (Worldwide remote jobs) Remote jobs United States Remote jobs Australia Remote jobs Brazil Remote jobs Canada Remote jobs France Remote jobs Ireland Remote jobs Germany Remote jobs Netherlands Remote jobs Spain Remote jobs UK Popular Jobs Remote data analyst jobs Remote customer support jobs Remote executive assistant jobs Remote marketing jobs Remote product designer jobs Remote product manager jobs Remote project manager jobs Remote recruiter jobs Remote sales jobs Remote software engineer jobs Jobs by Type Remote full-time jobs Remote part-time jobs Remote contract jobs Remote internship jobs Remote entry-level jobs Remote jobs with no experience required Remote junior jobs (1-3 years of experience) Digital nomad jobs Remote jobs with no degree required Freelance remote jobs Temporary remote jobs Remote jobs hiring now Stay at home mom jobs