Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI/ML Research Engineer, LLM Post-Training & Evaluation

$110k - $240k per year
Full-time

Innodata Inc.

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.

Scope of the Role: 

Innodata is expanding its team of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation, you will build and optimize the technical foundations that power model improvement for foundation model builders and leading labs.

This role is ideal for someone who has hands-on experience fine-tuning and evaluating large language models (and ideally multimodal models), and who can bridge research and engineering in real-world customer environments. You will work closely with Language Data Scientists, Applied Research Scientists, data engineers, and client technical stakeholders to design and implement robust training/evaluation pipelines using both human-in-the-loop and AI-augmented methods.

The ideal candidate brings a strong computer science / machine learning engineering background, experience with modern LLM post-training workflows, and the ability to engage credibly with technical counterparts at leading AI organizations.

What You’ll Own:

As an AI/ML Research Engineer, LLM Training & Evaluation, you will design and implement the pipelines and tooling that connect data, evaluation, and post-training. You will help customers and internal teams move from evaluation findings to measurable model improvements.

Your work may include building fine-tuning workflows (e.g., supervised fine-tuning and preference-based optimization), integrating evaluation harnesses into model development loops, improving experiment reliability and throughput, and supporting advanced evaluation scenarios such as long-context, cross-modal, and dynamic multi-turn interactions.

You will also contribute to Innodata’s internal R&D efforts, including benchmark datasets, evaluation frameworks, and reusable infrastructure for model assessment and post-training experimentation. Additional responsibilities include (but are not limited to):


  • Lead or co-lead technically complex ML engineering projects from initial customer discussions through implementation and delivery

  • Design, build, and improve LLM training and post-training pipelines, including data ingestion, preprocessing, fine-tuning, evaluation, and experiment tracking

  • Implement and optimize evaluation systems for LLMs and multimodal models, including offline benchmarks and task-specific test harnesses

  • Integrate human-in-the-loop and AI-augmented evaluation signals into model development workflows

  • Build robust infrastructure and tooling for reproducible experimentation, metrics logging, and regression monitoring

  • Diagnose model behavior and pipeline failures, including data issues, training instability, metric inconsistencies, and evaluation drift

  • Collaborate with Language Data Scientists and Applied Research Scientists to translate evaluation frameworks into executable systems

  • Work closely with customer technical stakeholders to understand goals, constraints, and success criteria; propose and implement technically sound solutions

  • Contribute to internal research and platform development, including benchmark frameworks, evaluation tooling, and post-training workflow improvements

  • Contribute to best practices and standards for LLM training, evaluation, and quality assurance across projects

  • Mentor junior engineers and contribute to technical design reviews, documentation, and engineering rigor across the team

You’ll Thrive in This Role If You Have:


  • BS/MS/PhD in Computer Science, Machine Learning, AI, Applied Mathematics, or a related quantitative technical field (MS/PhD preferred)

  • 2-3 years of relevant industry or research engineering experience in ML/AI systems

  • Hands-on experience with LLM training / fine-tuning / post-training, including at least one of:


    • supervised fine-tuning (SFT)

    • preference optimization (e.g., DPO or related methods)

    • RLHF / RLAIF-style workflows

    • task- or domain-adaptation of foundation models

  • Strong programming skills in Python and experience building production-quality ML code

  • Experience with modern ML frameworks (e.g., PyTorch, JAX, TensorFlow) and model libraries/tooling (e.g., Hugging Face ecosystem, vLLM, distributed training stacks)

  • Experience designing and implementing evaluation pipelines for LLM/ML systems, including metrics computation, dataset handling, and experiment comparisons

  • Strong understanding of data pipelines and ML systems engineering, including reproducibility, observability, and debugging

  • Experience with large-scale distributed ML systems and performance optimization for training/evaluation workloads (GPU/accelerator environments preferred)

  • Experience with large-scale data processing and workflow orchestration in support of model training/evaluation

  • Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data engineers, and customer technical leads

  • Strong written and verbal communication skills, including the ability to explain complex technical tradeoffs to both technical and non-technical audiences

Technical Skills

ML / LLM Engineering


  • Experience training, fine-tuning, and evaluating transformer-based models

  • Understanding of post-training workflows and model iteration loops

  • Familiarity with inference-time considerations (latency, throughput, memory/performance tradeoffs) where relevant to evaluation or deployment

Evaluation & Experimentation


  • Experience implementing automated evaluation pipelines and test harnesses

  • Experience with experiment tracking, versioning, and reproducibility practices

  • Ability to assess metric quality and ensure consistency across model comparisons

Software / Data Engineering


  • Proficiency in Python and strong software engineering fundamentals

  • Experience with data processing pipelines, storage formats, and scalable dataset workflows

  • Familiarity with CI/CD, testing, and engineering quality practices for ML systems

The expected salary range for this position is $110,000 – $240,000 CAD per year, based on experience, skills, and qualifications.

 

 

Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at  

If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at  View email address on jobs.jobcopilot.com and consider reporting it to the FTC at  ReportFraud.ftc.gov .

Vacancy posted 8 hours ago
Similar jobs that could be interesting for youBased on the AI/ML Research Engineer, LLM Post-Training & Evaluation in Remote vacancy
  • $130k - $220k per year

     ...and the age of 5G and IoT, with world class engineering, best-in-class user experience, and...  ...to join our team. We’re looking for an AI/ML Engineer who will develop, optimize, and...  ...Design & Deploy Conversational / Multi-Agent LLM Solutions Craft multi-agent conversational... 
    Suggested
    Full time
    Work from home
    Flexible hours

    Us Mobile

    Remote
    15 hours ago
  • $150k - $180k per year

     ...rise to the occasion  About the role:   We are seeking an AI/ML Engineer to join BPM’s Enterprise Technology Solutions team. This role...  ...Key Responsibilities:    Model Development : Design and evaluate AI/ML models tailored to BPM use cases (e.g., tax AI bot, Copilot... 
    Training
    Remote job
    Full time
    Local area
    Immediate start
    Flexible hours
    Shift work

    Bpm Llp

    Remote
    15 hours ago
  •  ...Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and LLMs . The SOC-TH team builds...  ..., and secure Continuously evaluate and improve model performance through...  ...and designing regression testing for LLM pipelines ~ Ability to rapidly prototype... 
    Suggested
    Full time
    Remote work
    Worldwide

    Censys

    Remote
    15 hours ago
  •  ...people, encourage their ideas and reward their results. As an  AI/ML Research Intern , you will be an integral member of a team of...  ...responsibility, mentored by industry-leading experts, and attend a robust training program to ensure your success at DRW. How will you make an... 
    Training
    Summer work
    Internship
    Work at office
    Immediate start
    Day shift

    Drw Inc

    Remote
    15 hours ago
  • $159k - $176k per year

     ...With the largest long-term and post-acute care dataset and a Marketplace...  ...of our revenue back into research and development, ensuring our employees...  ...human-first and accelerated by AI to create meaningful and...  ..., and we continue to invest in training and development to nurture innovation... 
    Training
    Remote job
    Long term contract
    Full time
    Manual labor
    Work at office
    Flexible hours

    Pointclickcare

    Remote
    15 hours ago
  •  ...combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions...  ...avec les cadres d'évaluation de LLM et l'analyse statistique. ~ Capacité à...  ...experiences and client partnerships. As a QA / AI Evaluation Engineer, you will join a highly... 
    Full time
    Apprenticeship
    Immediate start

    Appnovation Technologies

    Remote
    8 hours ago
  •  ...Architect / Principal, AI ML Accomplished Tech Visionary:  Embark on an exciting journey...  ...across multiple domains. You will evaluate emerging technologies, define architectural...  ...systems including data pipelines, model training/inference, and MLOps. Serve as a strategic... 
    Training
    Full time
    Work at office
    Remote work
    Flexible hours

    3pillar

    Remote
    15 hours ago
  •  ...We are seeking a machine learning (ML) research developer to join our team working on a novel AI safety agenda. In this role, you will work closely with ML research scientists to solve difficult training and inference problems using very large models. Key responsibilities... 
    Training
    Full time
    Flexible hours

    LawZero

    Remote
    15 hours ago
  •  ...analytics, data warehousing, observability, and AI workloads. The company’s sustained,...  ...journey! About the team The AI/ML Engineering team builds and operates ClickHouse's AI...  ...monitoring, versioning, observability, and evaluation. Oncall : Participate in the daytime... 
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Clickhouse

    Remote
    15 hours ago
  • $161.5k - $191.5k per year

     ...customer connections using real-time, AI-driven insights. We’re now...  .... Your Role As an AI Engineer: Voice Designer, you’ll own the...  ...TTS vendor APIs while leading research and prototyping for open-source...  ...Management: Design and manage LLM and TTS prompts and parameters... 
    Training
    Full time
    Manual labor
    Work at office
    Shift work

    Dialpad

    Remote
    15 hours ago
  •  ...Senior ML Engineer About Invoca   Invoca is an AI-powered revenue execution platform that brings together marketing, commerce, and contact center teams...  ...team owns the full ML lifecycle at Invoca, from model training and fine-tuning through inference optimization and production... 
    Training
    Long term contract
    Full time
    Remote work
    Flexible hours

    Invoca

    Remote
    15 hours ago
  • $129.39k - $217.13k per year

     ...Grafana Cloud's actually useful AI, organizations can see,...  ...only at this time. Senior AI Engineer  The Opportunity:  At Grafana...  ...including shipping and evolving LLM- or agent-powered workflows for...  ...information provided in CVs to job postings. The recruitment team will... 
    Full time
    Local area
    Remote work
    Flexible hours

    Grafana Labs

    Remote
    15 hours ago
  • $207k - $242.5k per year

     ...specified location above.  We are AI Native We are building an...  ...— firmware, app, and cloud engineering working as one team to ship the...  ...up through mass production and post-launch refinement. We're moving...  ...is the scale this on-device ML platform should be built to handle... 
    Training
    Remote job
    Full time
    Contract work
    Summer work
    Flexible hours

    Life360

    Remote
    15 hours ago
  • $50 per hour

     ...specialists with project-based AI opportunities for leading...  ...companies, focused on testing, evaluating, and improving AI systems.  Participation...  ...create challenging tasks and evaluation criteria within realistic...  ...data labeling Not prompt engineering Not writing code from... 
    Permanent employment
    Full time
    Temporary work

    Mindrift

    Remote
    15 hours ago
  • $160k - $180k per year

     ...building What we're working on AI agents are showing up in...  ...tackle this space. We're a focused engineering team working on the hard parts:...  ...systems, applied AI, and security research. We're early enough that the...  ...enforceable policy.  Be the AI/ML voice in our broader... 
    Full time
    Flexible hours

    Tigera

    Remote
    15 hours ago
  • $128.8k - $193.2k per year

     ...Opportunity When people talk about generative AI and other ML-powered solutions in today's conversation, they often refer to generative pre-trained transformers like ChatGPT that can respond...  ...for strategic product areas including LLM Inference and the AI Gateway. We are at... 
    Full time
    Internship
    Work at office
    Remote work
    Relocation package
    3 days per week

    Nutanix Inc.

    Remote
    15 hours ago
  • $110k - $130k per year

     ...KUBRA HQ is KUBRA’s unified, AI-powered, cloud-native platform...  ...journey seamlessly. The AI Engineer is responsible for designing, building...  ...(including generative AI and LLM-based systems), production-...  ...feature/prompt engineering, model training, orchestration, deployment, and... 
    Training
    Long term contract
    Permanent employment
    Full time
    Casual work
    Lmia
    Work at office
    Work visa

    Kubra

    Remote
    15 hours ago
  •  ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our...  ...intelligence by providing the data, evaluation frameworks, and human...  ...validity of datasets used to train, fine-tune, and evaluate health... 
    Training
    Full time
    Fixed term contract
    Shift work

    Innodata Inc.

    Remote
    15 hours ago
  • $210k - $245k per year

     ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our...  ...intelligence by providing the data, evaluation frameworks, and human...  ...validity of the datasets used to train, fine-tune, evaluate, and monitor... 
    Training
    Full time
    Fixed term contract
    Shift work

    Innodata Inc.

    Remote
    15 hours ago
  • $127k - $225k per year

     ...Huawei Canada has an immediate 12-month contract opening for a Research Engineer. About the team: The Huawei Digital Trust Lab is on a mission...  ...& authorization, kernel & hardware security, data security, ai agent security, advancement of agentic ai for product security.... 
    Full time
    Contract work
    Internship
    Immediate start
    Worldwide
    Shift work

    Huawei Technologies Canada Co.,

    Remote
    7 hours ago
  • $85k - $120k per year

     .... ABOUT YOU You are an AI practitioner who builds things,...  ...product works. You have shipped LLM-powered systems into production...  ...sits between product strategy and engineering architecture — with strong influence on both. This is not an ML research role. It is not a... 
    Training
    Long term contract
    Full time
    Flexible hours

    Xsolla

    Remote
    15 hours ago
  •  ...transforming SecOps with Morpheus, our AI-driven Autonomous Security Operations Center...  ...is looking for a hands-on Agentic AI Engineer — a software engineer who designs,...  ...is a software engineering role, not a research or model-training position — you will build production agents... 
    Training
    Permanent employment
    Full time
    Work at office
    Monday to friday

    D3 Security Management Systems

    Remote
    15 hours ago
  •  ...product company integrating cutting-edge AI capabilities into our core offering...  .... Our AI work spans task-specific ML models, large language model (LLM) integration, and agentic systems that...  ...We’re looking for a senior-level AI Engineer who is equally strong in backend engineering... 
    Training
    Full time

    Novoed

    Remote
    15 hours ago
  •  ...At Pinpoint Global, we build training and compliance platforms that organizations...  ...We're looking for true builders: engineers who orchestrate systems, ship production-grade AI, and modernize massive legacy...  ...~ Integrate AI capabilities (LLM APIs, intelligent automation, personalization... 
    Training
    Full time
    Internship
    Shift work

    Valsoft

    Remote
    15 hours ago
  • $209.3k - $313.8k per year

     ...Team:   As a Staff Machine Learning Engineer focused on End-to-End (E2E) Model Development...  ...leadership role focused on core model research and large-scale ML development, not feature-layer logic...  ...cost functions. Drive large-scale training and evaluation for E2E learning,... 
    Training
    Remote job
    Full time
    Work at office
    Relocation

    Torc Robotics

    Remote
    15 hours ago
  • $50k - $62k per year

     ...- $62,000 (Depending on prior ML experience, degree(s) obtained)...  ...clients.   We're hiring an ML Engineering Intern to join our ML Products...  ..., design, and extend internal AI-assisted skills that automate...  ...; you’ll have no bad habits to train you out of. PySpark or distributed... 
    Training
    Full time
    Temporary work
    Internship
    Work at office
    Immediate start
    Work from home
    Relocation
    Shift work

    Geocomply

    Remote
    15 hours ago
  •  ...We are seeking a senior distributed machine learning (ML) research developer to join our team working on a novel AI safety agenda. In this role, you will work closely with ML research scientists to solve difficult training and inference problems using very large models.... 
    Training
    Full time
    Work at office

    LawZero

    Remote
    15 hours ago
  •  ...Machine Learning Engineer (NLP/Multi Agent Systems) About us: At Spin...  ...at the intersection of AI and modern marketing, building...  ...work with new technologies and research in a forward-thinking organization...  ...Do Model Development: Train, fine-tune, and evaluate transformer... 
    Training
    Remote job
    Long term contract
    Full time
    Direct hire
    Immediate start
    Work from home
    Monday to friday

    TLNT Group

    Remote
    15 hours ago
  • $171.5k - $201k per year

     ...specified location above.  We are AI Native We are building an...  ...a product-minded, AI-Native engineering team. That means AI isn’t just...  ...builds its experimentation and ML infrastructure, develops its services...  ...agentic workflows, understands research-plan-implement cycle but doesn’... 
    Full time
    Summer work
    Internship
    Remote work
    Flexible hours

    Life360

    Remote
    15 hours ago
  •  ...solutions in Strategy, Analytics, Digital Engineering, Cloud, Data & AI, Experience Design, and Marketing....  ...is a production-focused role — not research or prototyping. We are looking for senior...  ..., and who understand failure modes, evaluation practices, and governance for mission... 
    Full time
    Internship
    Local area
    Worldwide
    Work visa

    Bounteous

    Remote
    15 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI/ML Research Engineer, LLM Post-Training & Evaluation. Be the first to apply!