AI/ML Research Engineer, LLM Post-Training & Evaluation
$110k - $240k per yearInnodata Inc.
Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.
Scope of the Role:
Innodata is expanding its team of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation, you will build and optimize the technical foundations that power model improvement for foundation model builders and leading labs.
This role is ideal for someone who has hands-on experience fine-tuning and evaluating large language models (and ideally multimodal models), and who can bridge research and engineering in real-world customer environments. You will work closely with Language Data Scientists, Applied Research Scientists, data engineers, and client technical stakeholders to design and implement robust training/evaluation pipelines using both human-in-the-loop and AI-augmented methods.
The ideal candidate brings a strong computer science / machine learning engineering background, experience with modern LLM post-training workflows, and the ability to engage credibly with technical counterparts at leading AI organizations.
What You’ll Own:
As an AI/ML Research Engineer, LLM Training & Evaluation, you will design and implement the pipelines and tooling that connect data, evaluation, and post-training. You will help customers and internal teams move from evaluation findings to measurable model improvements.
Your work may include building fine-tuning workflows (e.g., supervised fine-tuning and preference-based optimization), integrating evaluation harnesses into model development loops, improving experiment reliability and throughput, and supporting advanced evaluation scenarios such as long-context, cross-modal, and dynamic multi-turn interactions.
You will also contribute to Innodata’s internal R&D efforts, including benchmark datasets, evaluation frameworks, and reusable infrastructure for model assessment and post-training experimentation. Additional responsibilities include (but are not limited to):
- Lead or co-lead technically complex ML engineering projects from initial customer discussions through implementation and delivery
- Design, build, and improve LLM training and post-training pipelines, including data ingestion, preprocessing, fine-tuning, evaluation, and experiment tracking
- Implement and optimize evaluation systems for LLMs and multimodal models, including offline benchmarks and task-specific test harnesses
- Integrate human-in-the-loop and AI-augmented evaluation signals into model development workflows
- Build robust infrastructure and tooling for reproducible experimentation, metrics logging, and regression monitoring
- Diagnose model behavior and pipeline failures, including data issues, training instability, metric inconsistencies, and evaluation drift
- Collaborate with Language Data Scientists and Applied Research Scientists to translate evaluation frameworks into executable systems
- Work closely with customer technical stakeholders to understand goals, constraints, and success criteria; propose and implement technically sound solutions
- Contribute to internal research and platform development, including benchmark frameworks, evaluation tooling, and post-training workflow improvements
- Contribute to best practices and standards for LLM training, evaluation, and quality assurance across projects
- Mentor junior engineers and contribute to technical design reviews, documentation, and engineering rigor across the team
You’ll Thrive in This Role If You Have:
- BS/MS/PhD in Computer Science, Machine Learning, AI, Applied Mathematics, or a related quantitative technical field (MS/PhD preferred)
- 2-3 years of relevant industry or research engineering experience in ML/AI systems
- Hands-on experience with LLM training / fine-tuning / post-training, including at least one of:
- supervised fine-tuning (SFT)
- preference optimization (e.g., DPO or related methods)
- RLHF / RLAIF-style workflows
- task- or domain-adaptation of foundation models
- Strong programming skills in Python and experience building production-quality ML code
- Experience with modern ML frameworks (e.g., PyTorch, JAX, TensorFlow) and model libraries/tooling (e.g., Hugging Face ecosystem, vLLM, distributed training stacks)
- Experience designing and implementing evaluation pipelines for LLM/ML systems, including metrics computation, dataset handling, and experiment comparisons
- Strong understanding of data pipelines and ML systems engineering, including reproducibility, observability, and debugging
- Experience with large-scale distributed ML systems and performance optimization for training/evaluation workloads (GPU/accelerator environments preferred)
- Experience with large-scale data processing and workflow orchestration in support of model training/evaluation
- Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data engineers, and customer technical leads
- Strong written and verbal communication skills, including the ability to explain complex technical tradeoffs to both technical and non-technical audiences
Technical Skills
ML / LLM Engineering
- Experience training, fine-tuning, and evaluating transformer-based models
- Understanding of post-training workflows and model iteration loops
- Familiarity with inference-time considerations (latency, throughput, memory/performance tradeoffs) where relevant to evaluation or deployment
Evaluation & Experimentation
- Experience implementing automated evaluation pipelines and test harnesses
- Experience with experiment tracking, versioning, and reproducibility practices
- Ability to assess metric quality and ensure consistency across model comparisons
Software / Data Engineering
- Proficiency in Python and strong software engineering fundamentals
- Experience with data processing pipelines, storage formats, and scalable dataset workflows
- Familiarity with CI/CD, testing, and engineering quality practices for ML systems
The expected salary range for this position is $110,000 – $240,000 CAD per year, based on experience, skills, and qualifications.
Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at
If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at View email address on jobs.jobcopilot.com and consider reporting it to the FTC at ReportFraud.ftc.gov .
- ...Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and LLMs . The SOC-TH team builds... ..., and secure Continuously evaluate and improve model performance through... ...and designing regression testing for LLM pipelines ~ Ability to rapidly prototype...SuggestedFull timeRemote workWorldwide
- ...Architect / Principal, AI ML Accomplished Tech Visionary: Embark on an exciting journey... ...across multiple domains. You will evaluate emerging technologies, define architectural... ...systems including data pipelines, model training/inference, and MLOps. Serve as a strategic...TrainingFull timeWork at officeRemote workFlexible hours
- ...high-growth company delivering AI solutions that address some of... ...physics, mathematics, medicine, engineering, and other specialties. The company... ...R&D team builds leading-edge ML and physics-based models ("LQMs... ..., Tahoe-100M, and DILImap), train and evaluate expression-perturbation...TrainingPermanent employmentFull timeSeasonal workFlexible hours
- ...analytics, data warehousing, observability, and AI workloads. The company’s sustained,... ...journey! About the team The AI/ML Engineering team builds and operates ClickHouse's AI... ...monitoring, versioning, observability, and evaluation. Oncall : Participate in the daytime...SuggestedFull timeLocal areaRemote workHome officeFlexible hours
$207k - $242.5k per year
...specified location above. We are AI Native We are building an... ...— firmware, app, and cloud engineering working as one team to ship the... ...up through mass production and post-launch refinement. We're moving... ...is the scale this on-device ML platform should be built to handle...TrainingRemote jobFull timeContract workSummer workFlexible hours$186.37k - $230k per year
...Canada time zones only at this time. Staff AI Engineer The Opportunity: At Grafana, we... ...users, including shipping and evolving LLM- or agent-powered workflows for incident lifecycle... ...Opportunity Employer: We will recruit, train, compensate and promote regardless of race...Full timeLocal areaRemote work$50 per hour
...specialists with project-based AI opportunities for leading... ...companies, focused on testing, evaluating, and improving AI systems. Participation... ...create challenging tasks and evaluation criteria within realistic... ...data labeling Not prompt engineering Not writing code from...Permanent employmentFull timeTemporary work- ...At Pinpoint Global, we build training and compliance platforms that organizations... ...We're looking for true builders: engineers who orchestrate systems, ship production-grade AI, and modernize massive legacy... ...~ Integrate AI capabilities (LLM APIs, intelligent automation, personalization...TrainingFull timeInternshipShift work
$171.5k - $201k per year
...specified location above. We are AI Native We are building an... ...a product-minded, AI-Native engineering team. That means AI isn’t just... ...builds its experimentation and ML infrastructure, develops its services... ...agentic workflows, understands research-plan-implement cycle but doesn’...Full timeSummer workInternshipRemote workFlexible hours$210k - $245k per year
...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our... ...intelligence by providing the data, evaluation frameworks, and human... ...validity of the datasets used to train, fine-tune, evaluate, and monitor...TrainingFull timeFixed term contractShift work- ...like China to destinations worldwide. As an AI Engineer, you will own the design, development,... ...intelligent automation and agentic workflows to LLM integrations embedded across our product... ...to full agentic workflows Build, evaluate, and iterate on AI systems using a...Full timeRemote workWorldwide
- ...cloud infrastructure for the global AI economy. We are building a full-stack... ...and enterprises from data and model training through to production deployment, without... ...of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU...TrainingLong term contractFull timeTemporary workImmediate startRemote work
$85k - $120k per year
.... ABOUT YOU You are an AI practitioner who builds things,... ...product works. You have shipped LLM-powered systems into production... ...sits between product strategy and engineering architecture — with strong influence on both. This is not an ML research role. It is not a...TrainingLong term contractFull timeFlexible hours- ...FirstPrinciples FirstPrinciples is a research organization building AI infrastructure for discovery in fundamental... ...-first team of builders, researchers, engineers, and thinkers working across Canada,... ...-as-code practices Partner with ML engineers, researchers, and software engineers...TrainingLong term contractFull timeRemote work
- ...Analytics is a global leader in AI and analytics, helping Fortune 1000... ...for a highly skilled AI Engineer with 7+ years of experience in software... ...integrating Generative AI/LLM APIs, AWS Bedrock, and other model... ...with DevOps, CI/CD pipelines, and ML pipelines within the AWS...Long term contractFull timeLocal area
- ...We’re looking for a Senior Software Engineer to take ownership of the production... ...platform that power our Machine Learning (ML) and Operations Research (OR) work. This is a hands-on,... ...need for experimentation, deployment, evaluation, and reproducibility. Turn...Full timeFor contractors
- ...Analytics is a global leader in AI and advanced analytics... ...continents, building cutting-edge ML and data solutions at scale. Join... ...enterprise AI. As an Innovation engineer, you will play a pivotal role... ...platform engineers, product, and researchers. This is a hands-on, highly...Full timeInternshipLocal area
$95k per year
...edge cloud platform, coupled with AI-driven analytics tools, unlocks... ...experienced Senior Software Engineer, AI to perform a key role in our... ..., with a strong focus on LLM applications, agentic AI systems... ...backend engineering, data pipelines, evaluation, and production operations. You...TrainingFull timeFlexible hours- ...About Haven Safety: Haven Safety AI is building the enterprise learning... ...examines evidence, controls, engineering factors, procedures, regulations, training, and organizational history, then produces... ...: You will work on the AI and LLM engineering layer that connects Haven...TrainingFull timeInternship
$146.25k - $195k per year
...in these provinces. This is Engineering at Lattice Lattice's... ...and organizations to thrive. As AI becomes fundamental to every product... ...how AI quality is measured, evaluated, and continuously improved... ...experimentation platforms, or ML infrastructure. Knowledge of...Long term contractFull timeWork at officeRemote workWork from home$100k - $130k per year
...participation. Novara’s combination of software, training, and AI-powered tools put people and safety first... ...products. You will partner closely with product, engineering, and the broader QA team to bring rigor to how we test LLM-driven and agentic workflows, while also...TrainingFull timeLocal areaImmediate startFlexible hoursShift work- ...supply usage. We’re building an AI-driven software solution that... ...Role: As a Computer Vision/AI Engineer, you will be responsible for... ...world challenges into practical CV/ML solutions that create a... ...efficiency and robustness Establish evaluation frameworks for CV and VLM...Remote jobFull timeFlexible hours
- ...Question: Great Question is the all-in-one AI customer research platform for understanding your... ...redefining user research Product Engineer (Full-Stack) — Canada, Remote I’m... ...that: breaking work into sensible pieces, evaluating the output carefully, and most...TrainingFull timeLive InRemote workFlexible hoursShift workDay shift
- ...Description Job Role: AI-First Lead Engineer Location: Ontario - Remote Who we are... ...patterns, and fallback mechanisms Own AI evaluation frameworks, establishing quality metrics... ...experience, including 5+ years delivering ML/AI solutions into production...Full timeRemote work
$120k - $155k per year
...connection. Do you love solving tough engineering challenges and want to shape how AI is used in real products? At... ...LlamaIndex, LangChain) Optimize LLM-powered features for performance, scalability... ...-augmented generation (RAG), and evaluation of system quality—not necessarily...Remote jobLong term contractFull timeFlexible hours$27 - $33 per hour
...Description Position at Dumas Position : Engineer in Training – Operations Reports To : Operation Manager / Area Manager Direct Reports : None Company : Dumas Contracting Ltd Business Address : 865 Mountjoy Street South, PO Box 1600, Timmins,...TrainingLong term contractPermanent employmentFull timeContract workTemporary workFor contractorsInternshipWork at officeShift work- ...looking for experienced Forward Deployment Engineer (Generative AI) with Gen AI experience to join our... ...has been recognized by various market research firms, including Forrester and Gartner... ...● Utilize Vertex AI for model training, tuning, and deployment, ensuring seamless...TrainingFull timeLocal area
- ...About The Role: We're looking for a rare kind of engineer: someone who thinks AI-first, ships end-to-end, and moves at a pace that makes a small... ...What You Bring: Hands-on experience building with LLM APIs and SDKs (OpenAI, Anthropic, or open models) - writing...Full timeFor contractorsLocal areaRemote workWeekend work
- ...things data! The team will be composed of data engineers, analytics managers and data scientists.... ...and implement data infrastructure for AI/ML initiatives, including building pipelines... ...integrations and data-to-embedding pipelines for LLM applications. ~ Deep experience with...Long term contractFull timeManual labor
- ...seeking a forward-thinking Senior Security Engineer, AI & DevSecOps to help shape the future of... ...coding, AI agents, and AI workflows. Evaluate new AI technologies and recommend secure... ...minimizing friction. Continuously research emerging AI technologies, threats, and best...Full timeInternshipLocal areaFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI/ML Research Engineer, LLM Post-Training & Evaluation. Be the first to apply!
- machine learning engineer Canada
- junior machine learning developer Canada
- deep learning research engineer Canada
- mechanical research engineer Canada
- ingénieur de recherche Canada
- research engineer Canada
- machine learning researcher Canada
- software engineer - ai machine learning Canada
- machine learning Canada
- machine learning part time Canada

