Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

AI/ML Research Engineer, LLM Post-Training & Evaluation

$110k - $240k per year
Full-time

Innodata Inc.

Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.

Scope of the Role: 

Innodata is expanding its team of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation, you will build and optimize the technical foundations that power model improvement for foundation model builders and leading labs.

This role is ideal for someone who has hands-on experience fine-tuning and evaluating large language models (and ideally multimodal models), and who can bridge research and engineering in real-world customer environments. You will work closely with Language Data Scientists, Applied Research Scientists, data engineers, and client technical stakeholders to design and implement robust training/evaluation pipelines using both human-in-the-loop and AI-augmented methods.

The ideal candidate brings a strong computer science / machine learning engineering background, experience with modern LLM post-training workflows, and the ability to engage credibly with technical counterparts at leading AI organizations.

What You’ll Own:

As an AI/ML Research Engineer, LLM Training & Evaluation, you will design and implement the pipelines and tooling that connect data, evaluation, and post-training. You will help customers and internal teams move from evaluation findings to measurable model improvements.

Your work may include building fine-tuning workflows (e.g., supervised fine-tuning and preference-based optimization), integrating evaluation harnesses into model development loops, improving experiment reliability and throughput, and supporting advanced evaluation scenarios such as long-context, cross-modal, and dynamic multi-turn interactions.

You will also contribute to Innodata’s internal R&D efforts, including benchmark datasets, evaluation frameworks, and reusable infrastructure for model assessment and post-training experimentation. Additional responsibilities include (but are not limited to):


  • Lead or co-lead technically complex ML engineering projects from initial customer discussions through implementation and delivery

  • Design, build, and improve LLM training and post-training pipelines, including data ingestion, preprocessing, fine-tuning, evaluation, and experiment tracking

  • Implement and optimize evaluation systems for LLMs and multimodal models, including offline benchmarks and task-specific test harnesses

  • Integrate human-in-the-loop and AI-augmented evaluation signals into model development workflows

  • Build robust infrastructure and tooling for reproducible experimentation, metrics logging, and regression monitoring

  • Diagnose model behavior and pipeline failures, including data issues, training instability, metric inconsistencies, and evaluation drift

  • Collaborate with Language Data Scientists and Applied Research Scientists to translate evaluation frameworks into executable systems

  • Work closely with customer technical stakeholders to understand goals, constraints, and success criteria; propose and implement technically sound solutions

  • Contribute to internal research and platform development, including benchmark frameworks, evaluation tooling, and post-training workflow improvements

  • Contribute to best practices and standards for LLM training, evaluation, and quality assurance across projects

  • Mentor junior engineers and contribute to technical design reviews, documentation, and engineering rigor across the team

You’ll Thrive in This Role If You Have:


  • BS/MS/PhD in Computer Science, Machine Learning, AI, Applied Mathematics, or a related quantitative technical field (MS/PhD preferred)

  • 2-3 years of relevant industry or research engineering experience in ML/AI systems

  • Hands-on experience with LLM training / fine-tuning / post-training, including at least one of:


    • supervised fine-tuning (SFT)

    • preference optimization (e.g., DPO or related methods)

    • RLHF / RLAIF-style workflows

    • task- or domain-adaptation of foundation models

  • Strong programming skills in Python and experience building production-quality ML code

  • Experience with modern ML frameworks (e.g., PyTorch, JAX, TensorFlow) and model libraries/tooling (e.g., Hugging Face ecosystem, vLLM, distributed training stacks)

  • Experience designing and implementing evaluation pipelines for LLM/ML systems, including metrics computation, dataset handling, and experiment comparisons

  • Strong understanding of data pipelines and ML systems engineering, including reproducibility, observability, and debugging

  • Experience with large-scale distributed ML systems and performance optimization for training/evaluation workloads (GPU/accelerator environments preferred)

  • Experience with large-scale data processing and workflow orchestration in support of model training/evaluation

  • Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data engineers, and customer technical leads

  • Strong written and verbal communication skills, including the ability to explain complex technical tradeoffs to both technical and non-technical audiences

Technical Skills

ML / LLM Engineering


  • Experience training, fine-tuning, and evaluating transformer-based models

  • Understanding of post-training workflows and model iteration loops

  • Familiarity with inference-time considerations (latency, throughput, memory/performance tradeoffs) where relevant to evaluation or deployment

Evaluation & Experimentation


  • Experience implementing automated evaluation pipelines and test harnesses

  • Experience with experiment tracking, versioning, and reproducibility practices

  • Ability to assess metric quality and ensure consistency across model comparisons

Software / Data Engineering


  • Proficiency in Python and strong software engineering fundamentals

  • Experience with data processing pipelines, storage formats, and scalable dataset workflows

  • Familiarity with CI/CD, testing, and engineering quality practices for ML systems

The expected salary range for this position is $110,000 – $240,000 CAD per year, based on experience, skills, and qualifications.

 

 

Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at  

If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at  View email address on jobs.jobcopilot.com and consider reporting it to the FTC at  ReportFraud.ftc.gov .

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the AI/ML Research Engineer, LLM Post-Training & Evaluation in Canada vacancy
  • $150k - $180k per year

     ...rise to the occasion  About the role:   We are seeking an AI/ML Engineer to join BPM’s Enterprise Technology Solutions team. This role...  ...Key Responsibilities:    Model Development : Design and evaluate AI/ML models tailored to BPM use cases (e.g., tax AI bot, Copilot... 
    Training
    Remote job
    Full time
    Local area
    Immediate start
    Flexible hours
    Shift work

    Bpm Llp

    Canada
    10 hours ago
  •  ...Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and LLMs . The SOC-TH team builds...  ..., and secure Continuously evaluate and improve model performance through...  ...and designing regression testing for LLM pipelines ~ Ability to rapidly prototype... 
    Suggested
    Full time
    Remote work
    Worldwide

    Censys

    Canada
    10 hours ago
  •  ...Architect / Principal, AI ML Accomplished Tech Visionary:  Embark on an exciting journey...  ...across multiple domains. You will evaluate emerging technologies, define architectural...  ...systems including data pipelines, model training/inference, and MLOps. Serve as a strategic... 
    Training
    Full time
    Work at office
    Remote work
    Flexible hours

    3pillar

    Canada
    10 hours ago
  •  ...high-growth company delivering AI solutions that address some of...  ...physics, mathematics, medicine, engineering, and other specialties. The company...  ...R&D team builds leading-edge ML and physics-based models ("LQMs...  ..., Tahoe-100M, and DILImap), train and evaluate expression-perturbation... 
    Training
    Permanent employment
    Full time
    Seasonal work
    Flexible hours

    Sandbox Aq

    Canada
    10 hours ago
  •  ...analytics, data warehousing, observability, and AI workloads. The company’s sustained,...  ...journey! About the team The AI/ML Engineering team builds and operates ClickHouse's AI...  ...monitoring, versioning, observability, and evaluation. Oncall : Participate in the daytime... 
    Suggested
    Full time
    Local area
    Remote work
    Home office
    Flexible hours

    Clickhouse

    Canada
    10 hours ago
  • $186.37k - $230k per year

     ...Grafana Cloud's actually useful AI, organizations can see,...  ...only at this time. Staff AI Engineer  The Opportunity:  At Grafana...  ...including shipping and evolving LLM- or agent-powered workflows for...  ...information provided in CVs to job postings. The recruitment team will... 
    Remote job
    Full time
    Local area
    Flexible hours

    Grafana Labs

    Canada
    10 hours ago
  • $207k - $242.5k per year

     ...specified location above.  We are AI Native We are building an...  ...— firmware, app, and cloud engineering working as one team to ship the...  ...up through mass production and post-launch refinement. We're moving...  ...is the scale this on-device ML platform should be built to handle... 
    Training
    Remote job
    Full time
    Contract work
    Summer work
    Flexible hours

    Life360

    Canada
    10 hours ago
  • $50 per hour

     ...specialists with project-based AI opportunities for leading...  ...companies, focused on testing, evaluating, and improving AI systems.  Participation...  ...create challenging tasks and evaluation criteria within realistic...  ...data labeling Not prompt engineering Not writing code from... 
    Permanent employment
    Full time
    Temporary work

    Mindrift

    Canada
    10 hours ago
  •  ...At Pinpoint Global, we build training and compliance platforms that organizations...  ...We're looking for true builders: engineers who orchestrate systems, ship production-grade AI, and modernize massive legacy...  ...~ Integrate AI capabilities (LLM APIs, intelligent automation, personalization... 
    Training
    Full time
    Internship
    Shift work

    Valsoft

    Canada
    10 hours ago
  • $210k - $245k per year

     ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our...  ...intelligence by providing the data, evaluation frameworks, and human...  ...validity of the datasets used to train, fine-tune, evaluate, and monitor... 
    Training
    Full time
    Fixed term contract
    Shift work

    Innodata Inc.

    Canada
    10 hours ago
  •  ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our...  ...intelligence by providing the data, evaluation frameworks, and human...  ...validity of datasets used to train, fine-tune, and evaluate health... 
    Training
    Full time
    Fixed term contract
    Shift work

    Innodata Inc.

    Canada
    10 hours ago
  • $171.5k - $201k per year

     ...specified location above.  We are AI Native We are building an...  ...a product-minded, AI-Native engineering team. That means AI isn’t just...  ...builds its experimentation and ML infrastructure, develops its services...  ...agentic workflows, understands research-plan-implement cycle but doesn’... 
    Full time
    Summer work
    Internship
    Remote work
    Flexible hours

    Life360

    Canada
    10 hours ago
  •  ...agencies, product teams, and independent engineers that own original codebases and want to explore a new revenue stream through AI training licenses . It is a licensing...  ...to connect and submit the repository for evaluation. Code must not include third-party material... 
    Training
    Full time

    Gramian Consulting Group

    Canada
    10 hours ago
  •  ...like China to destinations worldwide. As an AI Engineer, you will own the design, development,...  ...intelligent automation and agentic workflows to LLM integrations embedded across our product...  ...to full agentic workflows Build, evaluate, and iterate on AI systems using a... 
    Full time
    Remote work
    Worldwide

    Portless

    Canada
    10 hours ago
  •  ...cloud infrastructure for the global AI economy. We are building a full-stack...  ...and enterprises from data and model training through to production deployment, without...  ...of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU... 
    Training
    Long term contract
    Full time
    Temporary work
    Immediate start
    Remote work

    Nebius

    Canada
    10 hours ago
  •  ...FirstPrinciples FirstPrinciples is a research organization building AI infrastructure for discovery in fundamental...  ...-first team of builders, researchers, engineers, and thinkers working across Canada,...  ...-as-code practices Partner with ML engineers, researchers, and software engineers... 
    Training
    Long term contract
    Full time
    Remote work

    First Principles, Llc

    Canada
    10 hours ago
  • $16 per hour

     ...passion for quality? We are looking for AI Response Evaluators to work on projects aimed at advancing...  ...detail Ability to perform sufficient research within allocated time, working within...  ...protect client confidentiality Must pass training and a required quality test before... 
    Training
    Long term contract
    Full time
    Contract work
    Part time
    Freelance
    Second job
    Remote work
    Weekend work
    3 days per week

    Welocalize

    Canada
    10 hours ago
  • $85k - $120k per year

     .... ABOUT YOU You are an AI practitioner who builds things,...  ...product works. You have shipped LLM-powered systems into production...  ...sits between product strategy and engineering architecture — with strong influence on both. This is not an ML research role. It is not a... 
    Training
    Long term contract
    Full time
    Flexible hours

    Xsolla

    Canada
    10 hours ago
  •  ...Analytics is a global leader in AI and advanced analytics...  ...continents, building cutting-edge ML and data solutions at scale. Join...  ...enterprise AI.  As an Innovation engineer, you will play a pivotal role...  ...platform engineers, product, and researchers. This is a hands-on, highly... 
    Full time
    Internship
    Local area

    Tiger Analytics

    Canada
    10 hours ago
  • $95k per year

     ...edge cloud platform, coupled with AI-driven analytics tools, unlocks...  ...experienced Senior Software Engineer, AI to perform a key role in our...  ..., with a strong focus on LLM applications, agentic AI systems...  ...backend engineering, data pipelines, evaluation, and production operations. You... 
    Training
    Full time
    Flexible hours

    Calabrio, Inc.

    Canada
    10 hours ago
  •  ...Analytics is a global leader in AI and analytics, helping Fortune 1000...  ...for a highly skilled  AI Engineer with 7+ years of experience in software...  ...integrating Generative AI/LLM APIs, AWS Bedrock, and other model...  ...with DevOps, CI/CD pipelines, and ML pipelines within the AWS... 
    Long term contract
    Full time
    Local area

    Tiger Analytics

    Canada
    10 hours ago
  • $146.25k - $195k per year

     ...in these provinces. This is Engineering at Lattice Lattice's...  ...and organizations to thrive. As AI becomes fundamental to every product...  ...how AI quality is measured, evaluated, and continuously improved...  ...experimentation platforms, or ML infrastructure. Knowledge of... 
    Long term contract
    Full time
    Work at office
    Remote work
    Work from home

    Lattice

    Canada
    10 hours ago
  • $27 - $33 per hour

     ...Description Poste chez Dumas Position : Engineer in Training – Operations Reports To : Operation Manager / Area Manager Direct Reports : None Company : Dumas Contracting Ltd Business Address : 865 Mountjoy Street South, PO Box 1600, Timmins, ON... 
    Training
    Long term contract
    Permanent employment
    Full time
    Contract work
    Temporary work
    For contractors
    Internship
    Work at office
    Shift work

    Carrières Dumas

    Canada
    10 hours ago
  •  ...Ericsson Ventures. We are seeking a Senior AI Security Engineer to focus on the emerging security...  ...autonomous AI agents. In this role, you will research, design, and implement novel techniques...  ...signals, and adversarial behaviors in LLM-powered agents. Agent Security... 
    Full time

    Menlo Security

    Canada
    10 hours ago
  •  ...Description Job Role: AI-First Lead Engineer Location: Ontario - Remote Who we are...  ...patterns, and fallback mechanisms Own AI evaluation frameworks, establishing quality metrics...  ...experience, including 5+ years delivering ML/AI solutions into production... 
    Full time
    Remote work

    Ezra Company

    Canada
    10 hours ago
  •  ...looking for experienced  Forward Deployment Engineer (Generative AI)  with Gen AI experience to join our...  ...has been recognized by various market research firms, including Forrester and Gartner...  ...● Utilize Vertex AI for model training, tuning, and deployment, ensuring seamless... 
    Training
    Full time
    Local area

    Tiger Analytics

    Canada
    10 hours ago
  •  ...Question: Great Question is the all-in-one AI customer research platform for understanding your...  ...redefining user research Product Engineer (Full-Stack) — Canada, Remote   I’m...  ...that: breaking work into sensible pieces, evaluating the output carefully, and most... 
    Training
    Full time
    Live In
    Remote work
    Flexible hours
    Shift work
    Day shift

    Great Question

    Canada
    10 hours ago
  •  ...About The Role:   We're looking for a rare kind of engineer: someone who thinks AI-first, ships end-to-end, and moves at a pace that makes a small...  ...What You Bring: Hands-on experience building with LLM APIs and SDKs (OpenAI, Anthropic, or open models) - writing... 
    Full time
    For contractors
    Local area
    Remote work
    Weekend work

    Ad Sauce

    Canada
    10 hours ago
  • $120k - $155k per year

     ...connection. Do you love solving tough engineering challenges and want to shape how AI is used in real products? At...  ...LlamaIndex, LangChain) Optimize LLM-powered features for performance, scalability...  ...-augmented generation (RAG), and evaluation of system quality—not necessarily... 
    Remote job
    Long term contract
    Full time
    Flexible hours

    Mediafly

    Canada
    10 hours ago
  •  ...hedge fund professionals to support a high-impact research initiative with a leading AI lab focused on training advanced foundational models. In this role, you will...  ...apply your expertise in long/short equity investing to evaluate and enhance financial reasoning content, helping... 
    Training
    Remote job
    Hourly pay
    Weekly pay
    Full time
    For contractors
    10 hours per week
    Flexible hours

    Weekday Ai

    Canada
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to AI/ML Research Engineer, LLM Post-Training & Evaluation. Be the first to apply!