AI/ML Research Engineer, LLM Post-Training & Evaluation
$110k - $240k per yearInnodata Inc.
Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.
Scope of the Role:
Innodata is expanding its team of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation, you will build and optimize the technical foundations that power model improvement for foundation model builders and leading labs.
This role is ideal for someone who has hands-on experience fine-tuning and evaluating large language models (and ideally multimodal models), and who can bridge research and engineering in real-world customer environments. You will work closely with Language Data Scientists, Applied Research Scientists, data engineers, and client technical stakeholders to design and implement robust training/evaluation pipelines using both human-in-the-loop and AI-augmented methods.
The ideal candidate brings a strong computer science / machine learning engineering background, experience with modern LLM post-training workflows, and the ability to engage credibly with technical counterparts at leading AI organizations.
What You’ll Own:
As an AI/ML Research Engineer, LLM Training & Evaluation, you will design and implement the pipelines and tooling that connect data, evaluation, and post-training. You will help customers and internal teams move from evaluation findings to measurable model improvements.
Your work may include building fine-tuning workflows (e.g., supervised fine-tuning and preference-based optimization), integrating evaluation harnesses into model development loops, improving experiment reliability and throughput, and supporting advanced evaluation scenarios such as long-context, cross-modal, and dynamic multi-turn interactions.
You will also contribute to Innodata’s internal R&D efforts, including benchmark datasets, evaluation frameworks, and reusable infrastructure for model assessment and post-training experimentation. Additional responsibilities include (but are not limited to):
- Lead or co-lead technically complex ML engineering projects from initial customer discussions through implementation and delivery
- Design, build, and improve LLM training and post-training pipelines, including data ingestion, preprocessing, fine-tuning, evaluation, and experiment tracking
- Implement and optimize evaluation systems for LLMs and multimodal models, including offline benchmarks and task-specific test harnesses
- Integrate human-in-the-loop and AI-augmented evaluation signals into model development workflows
- Build robust infrastructure and tooling for reproducible experimentation, metrics logging, and regression monitoring
- Diagnose model behavior and pipeline failures, including data issues, training instability, metric inconsistencies, and evaluation drift
- Collaborate with Language Data Scientists and Applied Research Scientists to translate evaluation frameworks into executable systems
- Work closely with customer technical stakeholders to understand goals, constraints, and success criteria; propose and implement technically sound solutions
- Contribute to internal research and platform development, including benchmark frameworks, evaluation tooling, and post-training workflow improvements
- Contribute to best practices and standards for LLM training, evaluation, and quality assurance across projects
- Mentor junior engineers and contribute to technical design reviews, documentation, and engineering rigor across the team
You’ll Thrive in This Role If You Have:
- BS/MS/PhD in Computer Science, Machine Learning, AI, Applied Mathematics, or a related quantitative technical field (MS/PhD preferred)
- 2-3 years of relevant industry or research engineering experience in ML/AI systems
- Hands-on experience with LLM training / fine-tuning / post-training, including at least one of:
- supervised fine-tuning (SFT)
- preference optimization (e.g., DPO or related methods)
- RLHF / RLAIF-style workflows
- task- or domain-adaptation of foundation models
- Strong programming skills in Python and experience building production-quality ML code
- Experience with modern ML frameworks (e.g., PyTorch, JAX, TensorFlow) and model libraries/tooling (e.g., Hugging Face ecosystem, vLLM, distributed training stacks)
- Experience designing and implementing evaluation pipelines for LLM/ML systems, including metrics computation, dataset handling, and experiment comparisons
- Strong understanding of data pipelines and ML systems engineering, including reproducibility, observability, and debugging
- Experience with large-scale distributed ML systems and performance optimization for training/evaluation workloads (GPU/accelerator environments preferred)
- Experience with large-scale data processing and workflow orchestration in support of model training/evaluation
- Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data engineers, and customer technical leads
- Strong written and verbal communication skills, including the ability to explain complex technical tradeoffs to both technical and non-technical audiences
Technical Skills
ML / LLM Engineering
- Experience training, fine-tuning, and evaluating transformer-based models
- Understanding of post-training workflows and model iteration loops
- Familiarity with inference-time considerations (latency, throughput, memory/performance tradeoffs) where relevant to evaluation or deployment
Evaluation & Experimentation
- Experience implementing automated evaluation pipelines and test harnesses
- Experience with experiment tracking, versioning, and reproducibility practices
- Ability to assess metric quality and ensure consistency across model comparisons
Software / Data Engineering
- Proficiency in Python and strong software engineering fundamentals
- Experience with data processing pipelines, storage formats, and scalable dataset workflows
- Familiarity with CI/CD, testing, and engineering quality practices for ML systems
The expected salary range for this position is $110,000 – $240,000 CAD per year, based on experience, skills, and qualifications.
Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at
If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at View email address on jobs.jobcopilot.com and consider reporting it to the FTC at ReportFraud.ftc.gov .
$150k - $180k per year
...rise to the occasion About the role: We are seeking an AI/ML Engineer to join BPM’s Enterprise Technology Solutions team. This role... ...Key Responsibilities: Model Development : Design and evaluate AI/ML models tailored to BPM use cases (e.g., tax AI bot, Copilot...TrainingRemote jobFull timeLocal areaImmediate startFlexible hoursShift work- ...Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and LLMs . The SOC-TH team builds... ..., and secure Continuously evaluate and improve model performance through... ...and designing regression testing for LLM pipelines ~ Ability to rapidly prototype...SuggestedFull timeRemote workWorldwide
- ...Architect / Principal, AI ML Accomplished Tech Visionary: Embark on an exciting journey... ...across multiple domains. You will evaluate emerging technologies, define architectural... ...systems including data pipelines, model training/inference, and MLOps. Serve as a strategic...TrainingFull timeWork at officeRemote workFlexible hours
- ...high-growth company delivering AI solutions that address some of... ...physics, mathematics, medicine, engineering, and other specialties. The company... ...R&D team builds leading-edge ML and physics-based models ("LQMs... ..., Tahoe-100M, and DILImap), train and evaluate expression-perturbation...TrainingPermanent employmentFull timeSeasonal workFlexible hours
- ...analytics, data warehousing, observability, and AI workloads. The company’s sustained,... ...journey! About the team The AI/ML Engineering team builds and operates ClickHouse's AI... ...monitoring, versioning, observability, and evaluation. Oncall : Participate in the daytime...SuggestedFull timeLocal areaRemote workHome officeFlexible hours
$186.37k - $230k per year
...Grafana Cloud's actually useful AI, organizations can see,... ...only at this time. Staff AI Engineer The Opportunity: At Grafana... ...including shipping and evolving LLM- or agent-powered workflows for... ...information provided in CVs to job postings. The recruitment team will...Remote jobFull timeLocal areaFlexible hours$207k - $242.5k per year
...specified location above. We are AI Native We are building an... ...— firmware, app, and cloud engineering working as one team to ship the... ...up through mass production and post-launch refinement. We're moving... ...is the scale this on-device ML platform should be built to handle...TrainingRemote jobFull timeContract workSummer workFlexible hours$50 per hour
...specialists with project-based AI opportunities for leading... ...companies, focused on testing, evaluating, and improving AI systems. Participation... ...create challenging tasks and evaluation criteria within realistic... ...data labeling Not prompt engineering Not writing code from...Permanent employmentFull timeTemporary work- ...At Pinpoint Global, we build training and compliance platforms that organizations... ...We're looking for true builders: engineers who orchestrate systems, ship production-grade AI, and modernize massive legacy... ...~ Integrate AI capabilities (LLM APIs, intelligent automation, personalization...TrainingFull timeInternshipShift work
$210k - $245k per year
...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our... ...intelligence by providing the data, evaluation frameworks, and human... ...validity of the datasets used to train, fine-tune, evaluate, and monitor...TrainingFull timeFixed term contractShift work- ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our... ...intelligence by providing the data, evaluation frameworks, and human... ...validity of datasets used to train, fine-tune, and evaluate health...TrainingFull timeFixed term contractShift work
$171.5k - $201k per year
...specified location above. We are AI Native We are building an... ...a product-minded, AI-Native engineering team. That means AI isn’t just... ...builds its experimentation and ML infrastructure, develops its services... ...agentic workflows, understands research-plan-implement cycle but doesn’...Full timeSummer workInternshipRemote workFlexible hours- ...agencies, product teams, and independent engineers that own original codebases and want to explore a new revenue stream through AI training licenses . It is a licensing... ...to connect and submit the repository for evaluation. Code must not include third-party material...TrainingFull time
- ...like China to destinations worldwide. As an AI Engineer, you will own the design, development,... ...intelligent automation and agentic workflows to LLM integrations embedded across our product... ...to full agentic workflows Build, evaluate, and iterate on AI systems using a...Full timeRemote workWorldwide
- ...cloud infrastructure for the global AI economy. We are building a full-stack... ...and enterprises from data and model training through to production deployment, without... ...of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU...TrainingLong term contractFull timeTemporary workImmediate startRemote work
- ...FirstPrinciples FirstPrinciples is a research organization building AI infrastructure for discovery in fundamental... ...-first team of builders, researchers, engineers, and thinkers working across Canada,... ...-as-code practices Partner with ML engineers, researchers, and software engineers...TrainingLong term contractFull timeRemote work
$16 per hour
...passion for quality? We are looking for AI Response Evaluators to work on projects aimed at advancing... ...detail Ability to perform sufficient research within allocated time, working within... ...protect client confidentiality Must pass training and a required quality test before...TrainingLong term contractFull timeContract workPart timeFreelanceSecond jobRemote workWeekend work3 days per week$85k - $120k per year
.... ABOUT YOU You are an AI practitioner who builds things,... ...product works. You have shipped LLM-powered systems into production... ...sits between product strategy and engineering architecture — with strong influence on both. This is not an ML research role. It is not a...TrainingLong term contractFull timeFlexible hours- ...Analytics is a global leader in AI and advanced analytics... ...continents, building cutting-edge ML and data solutions at scale. Join... ...enterprise AI. As an Innovation engineer, you will play a pivotal role... ...platform engineers, product, and researchers. This is a hands-on, highly...Full timeInternshipLocal area
$95k per year
...edge cloud platform, coupled with AI-driven analytics tools, unlocks... ...experienced Senior Software Engineer, AI to perform a key role in our... ..., with a strong focus on LLM applications, agentic AI systems... ...backend engineering, data pipelines, evaluation, and production operations. You...TrainingFull timeFlexible hours- ...Analytics is a global leader in AI and analytics, helping Fortune 1000... ...for a highly skilled AI Engineer with 7+ years of experience in software... ...integrating Generative AI/LLM APIs, AWS Bedrock, and other model... ...with DevOps, CI/CD pipelines, and ML pipelines within the AWS...Long term contractFull timeLocal area
$146.25k - $195k per year
...in these provinces. This is Engineering at Lattice Lattice's... ...and organizations to thrive. As AI becomes fundamental to every product... ...how AI quality is measured, evaluated, and continuously improved... ...experimentation platforms, or ML infrastructure. Knowledge of...Long term contractFull timeWork at officeRemote workWork from home$27 - $33 per hour
...Description Poste chez Dumas Position : Engineer in Training – Operations Reports To : Operation Manager / Area Manager Direct Reports : None Company : Dumas Contracting Ltd Business Address : 865 Mountjoy Street South, PO Box 1600, Timmins, ON...TrainingLong term contractPermanent employmentFull timeContract workTemporary workFor contractorsInternshipWork at officeShift work- ...Ericsson Ventures. We are seeking a Senior AI Security Engineer to focus on the emerging security... ...autonomous AI agents. In this role, you will research, design, and implement novel techniques... ...signals, and adversarial behaviors in LLM-powered agents. Agent Security...Full time
- ...Description Job Role: AI-First Lead Engineer Location: Ontario - Remote Who we are... ...patterns, and fallback mechanisms Own AI evaluation frameworks, establishing quality metrics... ...experience, including 5+ years delivering ML/AI solutions into production...Full timeRemote work
- ...looking for experienced Forward Deployment Engineer (Generative AI) with Gen AI experience to join our... ...has been recognized by various market research firms, including Forrester and Gartner... ...● Utilize Vertex AI for model training, tuning, and deployment, ensuring seamless...TrainingFull timeLocal area
- ...Question: Great Question is the all-in-one AI customer research platform for understanding your... ...redefining user research Product Engineer (Full-Stack) — Canada, Remote I’m... ...that: breaking work into sensible pieces, evaluating the output carefully, and most...TrainingFull timeLive InRemote workFlexible hoursShift workDay shift
- ...About The Role: We're looking for a rare kind of engineer: someone who thinks AI-first, ships end-to-end, and moves at a pace that makes a small... ...What You Bring: Hands-on experience building with LLM APIs and SDKs (OpenAI, Anthropic, or open models) - writing...Full timeFor contractorsLocal areaRemote workWeekend work
$120k - $155k per year
...connection. Do you love solving tough engineering challenges and want to shape how AI is used in real products? At... ...LlamaIndex, LangChain) Optimize LLM-powered features for performance, scalability... ...-augmented generation (RAG), and evaluation of system quality—not necessarily...Remote jobLong term contractFull timeFlexible hours- ...hedge fund professionals to support a high-impact research initiative with a leading AI lab focused on training advanced foundational models. In this role, you will... ...apply your expertise in long/short equity investing to evaluate and enhance financial reasoning content, helping...TrainingRemote jobHourly payWeekly payFull timeFor contractors10 hours per weekFlexible hours
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI/ML Research Engineer, LLM Post-Training & Evaluation. Be the first to apply!
- junior machine learning developer Canada
- machine learning engineer Canada
- ingénieur de recherche Canada
- mechanical research engineer Canada
- research engineer Canada
- deep learning research engineer Canada
- software engineer - ai machine learning Canada
- machine learning Canada
- machine learning researcher Canada
- machine learning part time Canada

