AI/ML Research Engineer, LLM Post-Training & Evaluation
$110k - $240k per yearInnodata Inc.
Innodata (Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our mission is to enable the responsible advancement of artificial intelligence by providing the data, evaluation frameworks, and human expertise required to build AI systems that can be trusted at scale. We provide a range of transferable solutions, platforms, and services for Generative AI / AI builders and adopters. In every relationship, we honor our 36+ year legacy delivering the highest quality data and outstanding outcomes for our customers.
Scope of the Role:
Innodata is expanding its team of technical experts in LLM training, post-training, and evaluation systems. As an AI/ML Research Engineer, LLM Training & Evaluation, you will build and optimize the technical foundations that power model improvement for foundation model builders and leading labs.
This role is ideal for someone who has hands-on experience fine-tuning and evaluating large language models (and ideally multimodal models), and who can bridge research and engineering in real-world customer environments. You will work closely with Language Data Scientists, Applied Research Scientists, data engineers, and client technical stakeholders to design and implement robust training/evaluation pipelines using both human-in-the-loop and AI-augmented methods.
The ideal candidate brings a strong computer science / machine learning engineering background, experience with modern LLM post-training workflows, and the ability to engage credibly with technical counterparts at leading AI organizations.
What You’ll Own:
As an AI/ML Research Engineer, LLM Training & Evaluation, you will design and implement the pipelines and tooling that connect data, evaluation, and post-training. You will help customers and internal teams move from evaluation findings to measurable model improvements.
Your work may include building fine-tuning workflows (e.g., supervised fine-tuning and preference-based optimization), integrating evaluation harnesses into model development loops, improving experiment reliability and throughput, and supporting advanced evaluation scenarios such as long-context, cross-modal, and dynamic multi-turn interactions.
You will also contribute to Innodata’s internal R&D efforts, including benchmark datasets, evaluation frameworks, and reusable infrastructure for model assessment and post-training experimentation. Additional responsibilities include (but are not limited to):
- Lead or co-lead technically complex ML engineering projects from initial customer discussions through implementation and delivery
- Design, build, and improve LLM training and post-training pipelines, including data ingestion, preprocessing, fine-tuning, evaluation, and experiment tracking
- Implement and optimize evaluation systems for LLMs and multimodal models, including offline benchmarks and task-specific test harnesses
- Integrate human-in-the-loop and AI-augmented evaluation signals into model development workflows
- Build robust infrastructure and tooling for reproducible experimentation, metrics logging, and regression monitoring
- Diagnose model behavior and pipeline failures, including data issues, training instability, metric inconsistencies, and evaluation drift
- Collaborate with Language Data Scientists and Applied Research Scientists to translate evaluation frameworks into executable systems
- Work closely with customer technical stakeholders to understand goals, constraints, and success criteria; propose and implement technically sound solutions
- Contribute to internal research and platform development, including benchmark frameworks, evaluation tooling, and post-training workflow improvements
- Contribute to best practices and standards for LLM training, evaluation, and quality assurance across projects
- Mentor junior engineers and contribute to technical design reviews, documentation, and engineering rigor across the team
You’ll Thrive in This Role If You Have:
- BS/MS/PhD in Computer Science, Machine Learning, AI, Applied Mathematics, or a related quantitative technical field (MS/PhD preferred)
- 2-3 years of relevant industry or research engineering experience in ML/AI systems
- Hands-on experience with LLM training / fine-tuning / post-training, including at least one of:
- supervised fine-tuning (SFT)
- preference optimization (e.g., DPO or related methods)
- RLHF / RLAIF-style workflows
- task- or domain-adaptation of foundation models
- Strong programming skills in Python and experience building production-quality ML code
- Experience with modern ML frameworks (e.g., PyTorch, JAX, TensorFlow) and model libraries/tooling (e.g., Hugging Face ecosystem, vLLM, distributed training stacks)
- Experience designing and implementing evaluation pipelines for LLM/ML systems, including metrics computation, dataset handling, and experiment comparisons
- Strong understanding of data pipelines and ML systems engineering, including reproducibility, observability, and debugging
- Experience with large-scale distributed ML systems and performance optimization for training/evaluation workloads (GPU/accelerator environments preferred)
- Experience with large-scale data processing and workflow orchestration in support of model training/evaluation
- Ability to collaborate directly with technical stakeholders including research scientists, ML engineers, data engineers, and customer technical leads
- Strong written and verbal communication skills, including the ability to explain complex technical tradeoffs to both technical and non-technical audiences
Technical Skills
ML / LLM Engineering
- Experience training, fine-tuning, and evaluating transformer-based models
- Understanding of post-training workflows and model iteration loops
- Familiarity with inference-time considerations (latency, throughput, memory/performance tradeoffs) where relevant to evaluation or deployment
Evaluation & Experimentation
- Experience implementing automated evaluation pipelines and test harnesses
- Experience with experiment tracking, versioning, and reproducibility practices
- Ability to assess metric quality and ensure consistency across model comparisons
Software / Data Engineering
- Proficiency in Python and strong software engineering fundamentals
- Experience with data processing pipelines, storage formats, and scalable dataset workflows
- Familiarity with CI/CD, testing, and engineering quality practices for ML systems
The expected salary range for this position is $110,000 – $240,000 CAD per year, based on experience, skills, and qualifications.
Please be aware of recruitment scams involving individuals or organizations falsely claiming to represent employers. Innodata will never ask for payment, banking details, or sensitive personal information during the application process. To learn more on how to recognize job scams, please visit the Federal Trade Commission’s guide at
If you believe you’ve been targeted by a recruitment scam, please report it to Innodata at View email address on jobs.jobcopilot.com and consider reporting it to the FTC at ReportFraud.ftc.gov .
$130k - $220k per year
...and the age of 5G and IoT, with world class engineering, best-in-class user experience, and... ...to join our team. We’re looking for an AI/ML Engineer who will develop, optimize, and... ...Design & Deploy Conversational / Multi-Agent LLM Solutions Craft multi-agent conversational...SuggestedFull timeWork from homeFlexible hours$150k - $180k per year
...rise to the occasion About the role: We are seeking an AI/ML Engineer to join BPM’s Enterprise Technology Solutions team. This role... ...Key Responsibilities: Model Development : Design and evaluate AI/ML models tailored to BPM use cases (e.g., tax AI bot, Copilot...TrainingRemote jobFull timeLocal areaImmediate startFlexible hoursShift work- ...Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and LLMs . The SOC-TH team builds... ..., and secure Continuously evaluate and improve model performance through... ...and designing regression testing for LLM pipelines ~ Ability to rapidly prototype...SuggestedFull timeRemote workWorldwide
- ...people, encourage their ideas and reward their results. As an AI/ML Research Intern , you will be an integral member of a team of... ...responsibility, mentored by industry-leading experts, and attend a robust training program to ensure your success at DRW. How will you make an...TrainingSummer workInternshipWork at officeImmediate startDay shift
$159k - $176k per year
...With the largest long-term and post-acute care dataset and a Marketplace... ...of our revenue back into research and development, ensuring our employees... ...human-first and accelerated by AI to create meaningful and... ..., and we continue to invest in training and development to nurture innovation...TrainingRemote jobLong term contractFull timeManual laborWork at officeFlexible hours- ...combines Strategy, Experience & Design, Engineering and Managed Services. We build digital solutions... ...avec les cadres d'évaluation de LLM et l'analyse statistique. ~ Capacité à... ...experiences and client partnerships. As a QA / AI Evaluation Engineer, you will join a highly...Full timeApprenticeshipImmediate start
- ...Architect / Principal, AI ML Accomplished Tech Visionary: Embark on an exciting journey... ...across multiple domains. You will evaluate emerging technologies, define architectural... ...systems including data pipelines, model training/inference, and MLOps. Serve as a strategic...TrainingFull timeWork at officeRemote workFlexible hours
- ...We are seeking a machine learning (ML) research developer to join our team working on a novel AI safety agenda. In this role, you will work closely with ML research scientists to solve difficult training and inference problems using very large models. Key responsibilities...TrainingFull timeFlexible hours
- ...analytics, data warehousing, observability, and AI workloads. The company’s sustained,... ...journey! About the team The AI/ML Engineering team builds and operates ClickHouse's AI... ...monitoring, versioning, observability, and evaluation. Oncall : Participate in the daytime...Full timeLocal areaRemote workHome officeFlexible hours
$161.5k - $191.5k per year
...customer connections using real-time, AI-driven insights. We’re now... .... Your Role As an AI Engineer: Voice Designer, you’ll own the... ...TTS vendor APIs while leading research and prototyping for open-source... ...Management: Design and manage LLM and TTS prompts and parameters...TrainingFull timeManual laborWork at officeShift work- ...Senior ML Engineer About Invoca Invoca is an AI-powered revenue execution platform that brings together marketing, commerce, and contact center teams... ...team owns the full ML lifecycle at Invoca, from model training and fine-tuning through inference optimization and production...TrainingLong term contractFull timeRemote workFlexible hours
$129.39k - $217.13k per year
...Grafana Cloud's actually useful AI, organizations can see,... ...only at this time. Senior AI Engineer The Opportunity: At Grafana... ...including shipping and evolving LLM- or agent-powered workflows for... ...information provided in CVs to job postings. The recruitment team will...Full timeLocal areaRemote workFlexible hours$207k - $242.5k per year
...specified location above. We are AI Native We are building an... ...— firmware, app, and cloud engineering working as one team to ship the... ...up through mass production and post-launch refinement. We're moving... ...is the scale this on-device ML platform should be built to handle...TrainingRemote jobFull timeContract workSummer workFlexible hours$50 per hour
...specialists with project-based AI opportunities for leading... ...companies, focused on testing, evaluating, and improving AI systems. Participation... ...create challenging tasks and evaluation criteria within realistic... ...data labeling Not prompt engineering Not writing code from...Permanent employmentFull timeTemporary work$160k - $180k per year
...building What we're working on AI agents are showing up in... ...tackle this space. We're a focused engineering team working on the hard parts:... ...systems, applied AI, and security research. We're early enough that the... ...enforceable policy. Be the AI/ML voice in our broader...Full timeFlexible hours$128.8k - $193.2k per year
...Opportunity When people talk about generative AI and other ML-powered solutions in today's conversation, they often refer to generative pre-trained transformers like ChatGPT that can respond... ...for strategic product areas including LLM Inference and the AI Gateway. We are at...Full timeInternshipWork at officeRemote workRelocation package3 days per week$110k - $130k per year
...KUBRA HQ is KUBRA’s unified, AI-powered, cloud-native platform... ...journey seamlessly. The AI Engineer is responsible for designing, building... ...(including generative AI and LLM-based systems), production-... ...feature/prompt engineering, model training, orchestration, deployment, and...TrainingLong term contractPermanent employmentFull timeCasual workLmiaWork at officeWork visa- ...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our... ...intelligence by providing the data, evaluation frameworks, and human... ...validity of datasets used to train, fine-tune, and evaluate health...TrainingFull timeFixed term contractShift work
$210k - $245k per year
...Nasdaq: INOD) is a global data engineering company. We believe that data and Artificial Intelligence (AI) are inextricably linked. Our... ...intelligence by providing the data, evaluation frameworks, and human... ...validity of the datasets used to train, fine-tune, evaluate, and monitor...TrainingFull timeFixed term contractShift work$127k - $225k per year
...Huawei Canada has an immediate 12-month contract opening for a Research Engineer. About the team: The Huawei Digital Trust Lab is on a mission... ...& authorization, kernel & hardware security, data security, ai agent security, advancement of agentic ai for product security....Full timeContract workInternshipImmediate startWorldwideShift work$85k - $120k per year
.... ABOUT YOU You are an AI practitioner who builds things,... ...product works. You have shipped LLM-powered systems into production... ...sits between product strategy and engineering architecture — with strong influence on both. This is not an ML research role. It is not a...TrainingLong term contractFull timeFlexible hours- ...transforming SecOps with Morpheus, our AI-driven Autonomous Security Operations Center... ...is looking for a hands-on Agentic AI Engineer — a software engineer who designs,... ...is a software engineering role, not a research or model-training position — you will build production agents...TrainingPermanent employmentFull timeWork at officeMonday to friday
- ...product company integrating cutting-edge AI capabilities into our core offering... .... Our AI work spans task-specific ML models, large language model (LLM) integration, and agentic systems that... ...We’re looking for a senior-level AI Engineer who is equally strong in backend engineering...TrainingFull time
- ...At Pinpoint Global, we build training and compliance platforms that organizations... ...We're looking for true builders: engineers who orchestrate systems, ship production-grade AI, and modernize massive legacy... ...~ Integrate AI capabilities (LLM APIs, intelligent automation, personalization...TrainingFull timeInternshipShift work
$209.3k - $313.8k per year
...Team: As a Staff Machine Learning Engineer focused on End-to-End (E2E) Model Development... ...leadership role focused on core model research and large-scale ML development, not feature-layer logic... ...cost functions. Drive large-scale training and evaluation for E2E learning,...TrainingRemote jobFull timeWork at officeRelocation$50k - $62k per year
...- $62,000 (Depending on prior ML experience, degree(s) obtained)... ...clients. We're hiring an ML Engineering Intern to join our ML Products... ..., design, and extend internal AI-assisted skills that automate... ...; you’ll have no bad habits to train you out of. PySpark or distributed...TrainingFull timeTemporary workInternshipWork at officeImmediate startWork from homeRelocationShift work- ...We are seeking a senior distributed machine learning (ML) research developer to join our team working on a novel AI safety agenda. In this role, you will work closely with ML research scientists to solve difficult training and inference problems using very large models....TrainingFull timeWork at office
- ...Machine Learning Engineer (NLP/Multi Agent Systems) About us: At Spin... ...at the intersection of AI and modern marketing, building... ...work with new technologies and research in a forward-thinking organization... ...Do Model Development: Train, fine-tune, and evaluate transformer...TrainingRemote jobLong term contractFull timeDirect hireImmediate startWork from homeMonday to friday
$171.5k - $201k per year
...specified location above. We are AI Native We are building an... ...a product-minded, AI-Native engineering team. That means AI isn’t just... ...builds its experimentation and ML infrastructure, develops its services... ...agentic workflows, understands research-plan-implement cycle but doesn’...Full timeSummer workInternshipRemote workFlexible hours- ...solutions in Strategy, Analytics, Digital Engineering, Cloud, Data & AI, Experience Design, and Marketing.... ...is a production-focused role — not research or prototyping. We are looking for senior... ..., and who understand failure modes, evaluation practices, and governance for mission...Full timeInternshipLocal areaWorldwideWork visa
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI/ML Research Engineer, LLM Post-Training & Evaluation. Be the first to apply!
- junior machine learning developer Remote
- machine learning engineer Remote
- mechanical research engineer Remote
- ingénieur de recherche Remote
- research engineer Remote
- deep learning research engineer Remote
- software engineer - ai machine learning Remote
- machine learning Remote
- machine learning part time Remote
- machine learning researcher Remote

