Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Software Engineer 2 - LLM Inference

$128.8k - $193.2k per year
Full-time

Nutanix Inc.


Hungry, Humble, Honest, with Heart.

 

The Opportunity
When people talk about generative AI and other ML-powered solutions in today's conversation, they often refer to generative pre-trained transformers like ChatGPT that can respond to queries from a position of deep learning. A GPT-in-a-box solution removes the burden of building or implementing these AI solutions yourself. It also makes overcoming the complexity, inefficiency, and security challenges of generative AI and AI/ML applications easy. Nutanix simplifies your learning curve on AI-ready infrastructure with Nutanix Cloud Platform for AI (GPT-in-a-Box). This high-performant Machine Learning full-stack cloud platform helps you optimise IT costs with a software-defined cloud operating model. Harness AI-ready capabilities right out of the box, simplified to build, fine-tune, and run models, including GPTs and LLMs, while you continue to use existing teams and skills.
Join the Nutanix AI team, responsible for the magic behind the scenes.

 

About the Team:
The Nutanix Enterprise AI team is responsible for strategic product areas including LLM Inference and the AI Gateway. We are at the forefront of Nutanix's mission to simplify AI deployment, recently showcasing our Agentic AI platform at NVIDIA GTC and NEXT 2026. This team is fast-paced, globally distributed, and focused on building the foundational layers of the AI stack.

 

You will report to a seasoned Technical Manager who will provide mentorship and guidance as you navigate through your responsibilities. The work setup at Nutanix AI is a hybrid model, offering a blend of in-office collaboration and remote work flexibility. As a new hire, you will be expected to be in the office for 3 days a week, ensuring that you have the opportunity to engage with your team and foster strong working relationships.

Your Role: 

  • Architect, design, and develop horizontally scalable, containerized, fault-tolerant services on Kubernetes.
  • Improve the performance of systems to deliver for low-latency and high-throughput use cases.
  • Optimize any part of the stack, including low-level systems.
  • Leverage and contribute to relevant open-source cloud native projects.
  • Develop scalable, efficient, and fault-tolerant observability architectures for collecting, analyzing, and reporting metrics for various platform services.
  • Collaborate closely with globally located product management and backend development teams to deliver high-quality products in a fast-paced environment.
  • Contribute to all stages of the product development cycle: technical design, development, test, experimentation, analysis, and launch.
  • Be a team player by reviewing code and design docs, giving feedback on product specs and mocks, and documentation.
  • Participate in an ongoing process definition and technology selection to ensure our technology stack is current with relevant trends.
  • Continuously learn and improve your technical and non-technical abilities.

  • What You Will Bring
  • 2-5 years of experience developing maintainable, modular, resilient, fail-safe, and long-lasting code from a Product Development company.
  • Have strong programming fundamentals, data structure, and algorithms.
  • Strong experience in Docker, Kubernetes, and Cloud native technologies
  • Experience building applications with Go and Python
  • Experience building and managing CI/CD pipelines
  • Strong understanding of datacenter design, including computing, storage, and networking.
  • Familiarity with on-prem, cloud, and hybrid software deployment architectures
  • Good experience in designing and tuning high-performance system software 
  • Strong understanding of distributed computing and storage architectures
  • Strong knowledge of OS internals, virtualization, application performance monitoring, compute storage, and networking management
  • Familiarity with machine learning concepts and popular frameworks (like TensorFlow, PyTorch, etc) is a strong plus
  • Experience with hardware accelerators, such as GPUs, is a strong plus.
  • Experience working with large codebases or contributing to open source is a strong plus.
  • Experience in building multi-tenant services on a virtualized infrastructure is a solid plus.
  • Detail-oriented with a strong focus on quality, design, and user experience.
  • Inquisitive and highly motivated self-starter and problem solver with a drive to integrate, communicate, and work well with large projects and teams.
  • Track record of being reliable, responsible, and thorough.
  • Bachelor's/Master's in Computer Science or equivalent work experience

Learn More About the Technology:  


Highlighted Benefits (Vancouver, Canada)
Retirement:   RRSP with dollar-for-dollar matching up to 7% of base salary
Mental Health:   Dedicated mental health coverage plus top-tier paramedical benefits
Family:   Fully paid maternity and parental leave and generous bereavement leave, including time for the loss of a pet
Equity:   RSUs and Employee Stock Purchase Plan at a 15% discount
Time Off:   Company holidays, sick days, company wellness days, and vacation starting at 10 days


Work Arrangement Hybrid:   This role operates in a hybrid capacity, blending the benefits of remote work with the advantages of in-person collaboration. In locations where our workplace policy applies (i.e. San Jose, Durham, Mexico City, Vancouver, Bangalore, Pune, Hoofddorp, Belgrade, Barcelona, Singapore, Sydney and Tokyo), employees are expected to work onsite a minimum of 3 days per week to foster collaboration, team alignment, and access to in-office resources. Workplace type may vary based on location and team requirements. Please speak with your recruiter for details. Additional team-specific guidance and norms will be provided by your manager.

 
Pay Transparency -  Role Location The pay range for this position at commencement of employment is expected to be between CAD $128,800 and CAD $193,200 per annual.
However, base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other elements, including a sign-on bonus, restricted stock units, and discretionary awards in addition to a full range of medical, financial and/or other benefits (including 401(k) eligibility and various paid time off benefits, such as vacation, sick time, and parental leave), dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment.

 
If hired, employee will be in an “at-will position” and the Company reserves the right to modify base salary (as well as any other discretionary payment or compensation program) at any time, including for reasons related to individual performance, Company or individual department/team performance, and market factors. Our application deadline is 40 days from the date of posting. In good faith, the posting may be removed prior to this date if the position is filled or extended in good faith.

 

--

Vacancy posted 10 hours ago
Similar jobs that could be interesting for youBased on the Software Engineer 2 - LLM Inference in Remote vacancy
  • $132k - $198k per year

     ...AI-ready capabilities out of the box ("GPT-in-a-Box") through a software-defined, full-stack infrastructure solution that simplifies AI...  ...forefront of this innovation, driving strategic products such as LLM Inference, the AI Gateway, and the Agentic AI Platform, recently... 
    Suggested
    Full time
    Internship
    Work at office
    Remote work
    Relocation package
    3 days per week

    Nutanix Inc.

    Remote
    10 hours ago
  •  ...Summary: Censys is seeking a Senior Software Engineer to join our SOC/TH team focused on AI and...  ...engineering experience, including 2+ years building and scaling AI-powered user...  ...LangSmith) and designing regression testing for LLM pipelines ~ Ability to rapidly... 
    Suggested
    Remote work
    Worldwide

    Censys

    Remote
    10 hours ago
  • $128.8k - $193.2k per year

     ...Honest, with Heart   The Opportunity We are looking for a Software Engineer to join our NKP (Nutanix Kubernetes Platform) team in Vancouver...  ...and optimizing existing services. What You Will Bring ~2-4 years of experience in a software development or engineering... 
    Suggested
    Long term contract
    Full time
    Work at office
    Remote work
    Relocation package
    3 days per week

    Nutanix Inc.

    Remote
    10 hours ago
  •  ...How will I make an impact?  We are looking for seasoned software developers who are passionate about developing software using engineering best practices. Senior software developers use past experiences & know-how to enable teams to be more productive & effective through... 
    Suggested
    Remote job
    Full time
    Work at office
    3 days per week

    D2l

    Remote
    10 hours ago
  • $74k - $87k per year

     .... 3 set days in the office and 2 WFH. About the role:  You...  ...design, build, and operate software that runs in production. This role...  ...it, and keep it healthy. Senior engineers will support your growth, and...  ...the products themselves: calling LLM APIs, wiring up tool use, and... 
    Suggested
    Long term contract
    Permanent employment
    Full time
    Internship
    Work at office
    Work from home
    Flexible hours

    Spring Financial

    Remote
    10 hours ago
  • $85k - $120k per year

     ...how a product works. You have shipped LLM-powered systems into production and have the...  ...who sits between product strategy and engineering architecture — with strong influence on both...  ...of engineering experience, with at least 2 years designing and deploying LLM-powered... 
    Long term contract
    Full time
    Flexible hours

    Xsolla

    Remote
    10 hours ago
  •  ...Competitive Intelligence Klue Engineering is hiring! We're looking for a Senior Software Engineer to join our team in...  ...build and optimize state-of-the-art LLM-powered agents at scale. You'll...  ...across the full stack: inference costs at scale, retrieval and query... 
    Full time
    Work at office

    Klue Inc.

    Remote
    10 hours ago
  •  ...date). Responsibilities Research and develop state-of-the-art LLM and/or video or multimodality models to accelerate open-source...  ...scenarios, like pre-training, supervised fine-tuning, post-training, inference etc. Continuously optimize the training framework to better... 
    Permanent employment
    Full time
    Internship
    Immediate start
    Remote work
    Worldwide

    Tether Operations Limited

    Remote
    15 days ago
  • $104k - $139k per year

     ...and distribute open-source software that enables people to enjoy the...  ...pipelines, high-throughput inference services, GPU orchestration, and...  ...looking for a Senior Software Engineer with a strong platform mindset...  ...Experience running a multi-provider LLM gateway, including routing... 
    Remote job
    Immediate start
    Home office

    Mozilla

    Remote
    10 hours ago
  • $196k - $207k per year

     ...together. We’re looking for an Engineering Manager to lead our Catalog Enrichment...  ...large language models, classical inference, workflow orchestration, and human...  ...Qualifications ~7+ years of software engineering experience, including 2+ years managing engineers as a... 
    Permanent employment
    Full time
    Work at office
    Remote work
    Work from home
    Flexible hours

    Instacart

    Remote
    10 hours ago
  •  ...Most of it has never been touched by modern software engineering, let alone AI. EviSmart is the dental...  ...for 10x').   Work hands-on alongside 2–3 engineers (30–50% IC) to bring system designs...  ...the full AI toolchain: Claude, Cursor, LLM-powered workflows in production, not a... 
    Long term contract
    Full time
    Temporary work
    Internship
    Work at office
    Immediate start

    Evismart

    Remote
    10 hours ago
  • $176.26k - $220.32k per year

     ...Make An Impact As a Senior Principal Software Engineer, you will be a technical leader driving...  ...Augmented Generation), prompt engineering, and LLM integration strategies. Cloud &...  ...fine tuning, model distillation, building inference infrastructure/framework, model... 
    Long term contract
    Full time

    Boomi

    Remote
    10 hours ago
  •  ...Abnormal AI is looking for a Senior Backend Engineer to join the App Foundations team. App...  ...Large Model Service (LMS, our internal LLM inference platform), and the MPP Go Service that sits...  ...Have Skills ~5+ years of backend software engineering experience building and... 
    Long term contract
    Full time
    Remote work

    Abnormal Company

    Remote
    10 hours ago
  •  ...About the Role We're seeking a Staff Software Engineer to lead the design and evolution of our...  ...systems - Familiarity with AI Agents, LLM-based systems, or AI orchestration platforms...  ...- LLM integration experience (streaming inference, prompt orchestration, RAG) -... 
    Internship
    Work at office
    Local area
    Remote work
    Work from home
    Home office

    Cresta

    Remote
    10 hours ago
  •  ...the role: You will own the inference backbone behind QVAC's local AI...  .... The role is centered on engineering quality at runtime level, including...  ...and enhancing inference engines like llama.cpp or similar, to...  ...know how to train/fine-tune a LLM You have productionized models... 
    Permanent employment
    Full time
    Local area
    Immediate start
    Remote work
    Worldwide

    Tether Operations Limited

    Remote
    2 days ago
  •  ...avenir. Aujourd'hui, nous recrutons un·e Software Engineer AI pour l'un de nos clients : une scale-...  ...as touché à des projets agentiques ou LLM - en production, en POC, ou en side project...  ...de télétravail partiel avec présentiel 2 jours/semaine (lundi + jeudi) à Montréal... 
    Daily paid
    Permanent employment
    Part time
    Remote work

    Maplr

    Remote
    10 hours ago
  •  ...for enterprises. We believe that software can radically transform the...  ...skilled and motivated Staff Software Engineer to join our team and work on...  ...using AI/ML technologies (LLM, SML, RAG, Prompt Engineering,...  ...global enterprises—and with over 3.2 billion downloads of Kyverno, Nirmata... 
    Long term contract
    Flexible hours

    Nirmata

    Remote
    10 hours ago
  • $100k - $120k per year

     ...growth-minded culture, you’ll belong here! We are hiring a Software Engineer to join our engineering team and work on the systems that power...  ...used by trading, operations, and research teams. Integrate AI/LLM workflows into internal processes where they add measurable... 
    Full time
    Remote work
    Flexible hours

    3iq

    Remote
    10 hours ago
  • $125k - $175k per year

     ...development of digital investigative software that acquires, analyzes, and...  ...Senior Software Development Engineer in Test (SDET) to join our...  ...Familiarity with testing LLM-powered applications and agentic...  ...as a machine-based system that infers from input to generate outputs... 
    Full time
    Contract work
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Remote
    10 hours ago
  • $20 per day

     ...can grow with us! As an Intermediate Engineer , you will be a hands-on contributor to...  ...of the ins and outs of data modeling and software architecture for developing performant and...  ..., including generating insights or inferences to assess job-related qualifications. This... 
    Internship
    Summer holiday
    Relocation

    Hiive

    Remote
    10 hours ago
  •  ...Baseten powers mission-critical inference for the world's most dynamic...  ...us and help build the platform engineers turn to to ship AI products....  ...are looking for early-career Software Engineers to join our team in...  ...HPC) and Large Language Model (LLM) engineering. You will be responsible... 
    Full time
    Flexible hours

    Baseten

    Remote
    10 hours ago
  • $80k - $130k per year

     ...development of digital investigative software that acquires, analyzes, and...  ...for a talented Software Engineer to join our growing team, responsible...  ...experience with .NET/C#; ~2+ years of ReactJS and/or...  ...as a machine-based system that infers from input to generate outputs... 
    Full time
    Work at office
    Local area
    Flexible hours

    Magnet Forensics

    Remote
    10 hours ago
  •  ...learn more, visit .  About the Role We are seeking a Software Engineering Manager to lead a team of engineers delivering high-quality...  ...experience Mobile device testing knowledge Awareness of AI/LLM capabilities including agentic frameworks, RAG, vector... 
    Long term contract
    Remote work
    Worldwide
    Home office

    Cority

    Remote
    10 hours ago
  • $135k - $150k per year

     ...of agentic AI in the accounting and finance space. The Senior Software Engineer will be the hands-on architect and builder of the new systems shaping...  ...where requirements evolve  Assets  ~ Experience with AI/LLM application development: prompt engineering, tool use, agentic... 
    Full time
    Flexible hours

    Treewalk Consulting Inc.

    Remote
    10 hours ago
  •  ...Overview   Giffen Consulting Ltd. (Giffen) is looking to hire a  Software Engineer Co-Op with experience in AI & Technology Integration t o  join...  ..., evaluate, and prototype AI tools and large language model (LLM) integrations that can improve engineering workflows, project... 
    Long term contract
    Temporary work
    For contractors
    Internship
    Work at office
    Work from home
    Flexible hours
    3 days per week

    Giffen Consulting

    Remote
    10 hours ago
  • $86.32k - $107.9k per year

     ...Position Overview: As a Software Engineer II at Diligent, you’ll take on a hands-on technical role in building secure, scalable, and high...  ...context length, embeddings, hallucinations), understands high-level LLM behavior, and recognizes safe vs. unsafe use cases (privacy,... 
    Full time
    Work at office
    Local area
    Flexible hours

    Diligent

    Remote
    10 hours ago
  •  ...Cambio is a software platform for world-class real estate decarbonization...  .... We're hiring a Software Engineer to accelerate the design and development...  ...: Next.js, Tailwind css LLM : AWS Bedrock, OpenAI, Claude...  ...is a hybrid role requiring 2 days of in-person collaboration... 
    Full time

    Cambio Ai Inc.

    Remote
    10 hours ago
  •  ...is the opportunity? We're looking for an engineer to teach our Flutter app how to talk —...  ...This role sits at the intersection of mobile software and embedded systems. You'll need to be equally...  ...pay and a Group RRSP (GRSP) with 2.5% employer matching (1-year vesting schedule... 
    Long term contract
    Full time
    Immediate start
    Remote work
    Flexible hours

    THRILLWORKS

    Remote
    10 hours ago
  •  ...really matter. Developers, DevOps engineers, and platform teams use Upsun...  ...rest. Our core belief is that software should power brighter...  ...maintaining our orchestration engine and git interface, ensuring they...  ...with AI coding assistants and LLM APIs, including applying them... 
    Full time
    Internship
    Work at office
    Remote work
    Flexible hours

    Remotewoman

    Remote
    10 hours ago
  • $15k per year

     ...and we have ambitious goals for the future. As a Software Engineer, you will shape how our business operates at its core...  ...explored Tableau, SQLMesh, MCP servers, and LLM-assisted workflows. Requirements ~2+ years of backend and/or data engineering experience.... 
    Full time
    Local area

    The Voleon Group

    Remote
    10 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Software Engineer 2 - LLM Inference. Be the first to apply!