Staff Engineer - AI/Inference /Gateway
$171k - $257k per yearNutanix Inc.
The Opportunity
At Nutanix, we're simplifying the future of AI with the Nutanix Cloud Platform for AI, enabling organizations to easily build, fine-tune, and run Generative AI, Large Language Models (LLMs), and next-generation Agentic AI applications without the complexity of managing AI infrastructure themselves. Our high-performance, full-stack machine learning cloud platform delivers AI-ready capabilities out of the box ("GPT-in-a-Box") through a software-defined, full-stack infrastructure solution that simplifies AI deployment across on-premises data centers, edge locations, and public clouds.
The Enterprise AI team is at the forefront of this innovation, driving strategic products such as LLM Inference, the AI Gateway, and the Agentic AI Platform, recently showcased at NVIDIA GTC and NEXT 2026. Join the team responsible for the foundational AI technologies powering the next wave of intelligent applications at Nutanix.
About the Team:
We are a fast-paced, globally distributed team building the foundational layers of the enterprise AI stack. By joining our team, you'll help shape the next generation of enterprise AI platforms, working at the intersection of large-scale distributed systems and machine learning infrastructure. This is an opportunity to solve complex technical challenges, influence the direction of key AI technologies, and build systems that power AI workloads at enterprise scale.
You will report to a seasoned Technical Lead/Manager who will provide mentorship and guidance as you navigate through your responsibilities. The work setup at Nutanix AI is a hybrid model, offering a blend of in-office collaboration and remote work flexibility. As a new hire, you will be expected to be in the office for 3 days a week, ensuring that you have the opportunity to engage with your team and foster strong working relationships.
Your Role
- Architect, design, and develop horizontally scalable, containerized, fault-tolerant services on Kubernetes for enterprise AI and LLM workloads.
- Build and operate high-performance inference and platform services that deliver low-latency, high-throughput experiences for Generative AI and Agentic AI applications.
- Design and optimize critical system components across the stack, including distributed systems, storage, networking, and low-level infrastructure layers.
- Develop and enhance multi-tenant platform services supporting on-premises, hybrid, and cloud-based AI deployments.
- Design and implement scalable observability architectures using technologies such as Prometheus, Grafana, Datadog, OpenTelemetry, and related cloud-native monitoring frameworks.
- Debug complex production issues, perform root-cause analysis, and improve reliability, resiliency, and operational efficiency of platform services.
- Build and maintain CI/CD pipelines and deployment automation to accelerate delivery of production-grade services.
- Design and implement foundational LLM serving capabilities including request routing, rate limiting, token streaming, load balancing, quota management, and usage budgeting.
- Collaborate closely with globally distributed product management, AI, and software engineering teams to deliver high-quality products in a fast-paced environment.
- Contribute to all stages of the product lifecycle, including architecture, design, development, testing, experimentation, performance analysis, deployment, and operations.
- Leverage and contribute to relevant open-source cloud-native and AI ecosystem projects.
- Review code and design documents, provide feedback on product requirements, and champion engineering excellence across the team.
- Continuously evaluate emerging technologies and help shape the technical direction of Nutanix's Enterprise AI Platform.
What You Will Bring
Required Qualifications
- 8+ years of experience developing maintainable, modular, resilient, fail-safe, and long-lived software products within a product development organization.
- Strong computer science fundamentals including data structures, algorithms, operating systems, networking, and distributed systems.
- Hands-on experience with Docker, Kubernetes, and cloud-native architectures.
- Production experience developing backend systems using Go, Python, C++, or Rust.
- Experience building, owning, and maintaining CI/CD pipelines and release automation end-to-end.
- Strong understanding of datacenter architecture including compute, storage, networking, and virtualization.
- Experience designing and deploying software across on-premises, cloud, and hybrid environments.
- Demonstrated experience designing and tuning high-performance, performance-sensitive system software.
- Solid understanding of distributed computing, distributed data stores, and large-scale service architectures.
- Experience diagnosing and resolving production performance issues using observability and monitoring platforms such as Prometheus, Grafana, Datadog, Open Telemetry, or similar tools.
- Familiarity with LLM serving concepts including rate limiting, token streaming, request scheduling, load balancing, quota management, and usage budgeting.
- Familiarity with modern LLM concepts including reasoning workflows, tool calling, prompt templates, and agent.
- Experience building multi-tenant services running on virtualized or containerized infrastructure.
- Strong communication, collaboration, and problem-solving skills with the ability to work effectively across globally distributed teams.
- Master's degree in Computer Science or equivalent practical experience.
Bonus Points If You Have Experience With
- Machine learning frameworks such as PyTorch or TensorFlow.
- GPU-based systems and acceleration technologies.
- Modern model-serving platforms such as vLLM, DeepSpeed, Hugging Face TGI, or Triton.
- Retrieval-Augmented Generation (RAG), vector databases, and AI orchestration frameworks.
- Open-source contributions or experience working in large distributed codebases.
- Production AI platforms, LLM APIs, agentic systems, or inference infrastructure.
- Building or scaling production LLM APIs, including streaming via SSE/WebSockets, prompt guardrails, rate limiting, and usage budgeting.
Learn More About the Technology:
#NAI
Highlighted Benefits (Vancouver, Canada)
Retirement: RRSP with dollar-for-dollar matching up to 7% of base salary
Mental Health: Dedicated mental health coverage plus top-tier paramedical benefits
Family: Fully paid maternity and parental leave and generous bereavement leave, including time for the loss of a pet
Equity: RSUs and Employee Stock Purchase Plan at a 15% discount
Work Arrangement Hybrid: This role operates in a hybrid capacity, blending the benefits of remote work with the advantages of in-person collaboration. In locations where our workplace policy applies (i.e. San Jose, Durham, Mexico City, Vancouver, Bangalore, Pune, Hoofddorp, Belgrade, Barcelona, Singapore, Sydney and Tokyo), employees are expected to work onsite a minimum of 3 days per week to foster collaboration, team alignment, and access to in-office resources. Workplace type may vary based on location and team requirements. Please speak with your recruiter for details. Additional team-specific guidance and norms will be provided by your manager.
Pay Transparency - Role Location The pay range for this position at commencement of employment is expected to be between CAD $171,000 and CAD $257,000 per annual.
However, base pay offered may vary depending on multiple individualized factors, including market location, job-related knowledge, skills, and experience. The total compensation package for this position may also include other elements, including a sign-on bonus, restricted stock units, and discretionary awards in addition to a full range of medical, financial and/or other benefits (including 401(k) eligibility and various paid time off benefits, such as vacation, sick time, and parental leave), dependent on the position offered. Details of participation in these benefit plans will be provided if an employee receives an offer of employment.
If hired, employee will be in an “at-will position” and the Company reserves the right to modify base salary (as well as any other discretionary payment or compensation program) at any time, including for reasons related to individual performance, Company or individual department/team performance, and market factors. Our application deadline is 40 days from the date of posting. In good faith, the posting may be removed prior to this date if the position is filled or extended in good faith.
--
$173.4k - $219.3k per year
...About Dialpad Dialpad is the AI platform for customer experience, built to resolve customer... ...Empathetic . Your role As a Sr. AI Engineer: Systems, you’ll serve as an embedded... ...productionizing models, enabling self-hosted inference, integrating external APIs, and building...SuggestedInternshipWork at office- ...Ashby's Co-Founder and VP of Engineering. We’re looking for a versatile... ...apply if: To you, a tech lead, staff, or principal engineer is... ...modeling and query language, policy engine, workflow engine, design... ...our app (short video below). AI-powered tooling. We think of...SuggestedFull timeInternshipWork at office
$150k - $170k per year
...transform insurance into a trusted and delightful experience using AI. We are building the AI operating system for North America’s... ...Insurtech 50 and Fintech 100 About the Role As a Full Stack Staff Engineer at Quandri, you will be directly responsible as an IC leader of...SuggestedLong term contractFull timeInternshipShift work$240k - $285k per year
...visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service... ...support you need to grow your career. Engineering at Brex Engineering at Brex is about... ...become leaders. What you’ll do As a Staff Software Engineer in Banking, you will...SuggestedLong term contractFull timeBank staffWork at officeRemote workWork from home3 days per week$135k - $170k per year
...DevOps Staff Platform Engineer Location: Remote (Canada) Compensation: $135,000 - $170,000 CAD + bonus... ...of how the business creates value, and AI runs as an enabling layer across the... ...disposable. Stand up and own our MCP and tool-gateway layer so agents reach infrastructure...SuggestedLong term contractRemote workFlexible hoursShift work$190k - $240k per year
...while deeply empathizing with real customer problems. You'll work across all teams at Durable to define and deliver what the future of AI UX looks like. You'll be responsible for execution, holding everyone to a high standard of quality. This is not a role where you'll spend...Work at office2 days per week$190k - $230k per year
...UC Berkeley, Cornell, SJTU, Cambridge, NUS, with broad industry experiences from Tencent, Alibaba, Amazon, ByteDance, and more. Our engineering team is composed of passionate, driven, and talented engineers who focus on building and innovating cloud technology solutions that...Hourly payFull timeWork at officeLocal areaWork from homeFlexible hours2 days per week- ...us forward. Across strategy, engineering, design, data, and operations,... ...explore, design, and implement AI strategies that are secure, scalable... ...an exceptional Chief of Staff for our Managing Director of AI... ...processes, building automated engines that run independently,...Full timeFor contractorsFor subcontractorRemote work
$204k - $258k per year
...About Dialpad Dialpad is the AI platform for customer experience, built to resolve customer problems in real time across voice... ...Optimistic, Persistent, and Empathetic . Your role As a Staff Software Engineer, you’ll own the platformization roadmap for shared services,...Work at office- ...uncharted. By combining our expertise across connectivity, AI, security and more, we’ll map a new way forward. Working together... ...to 60% at a typical company. Role Summary As a Staff Embedded Software Engineer, you will transform real-world ECU firmware into high-...Full timeContract workLocal area
- ...precision. With 465 billion automated optimizations per second, the AI-powered StackAdapt Marketing Platform seamlessly connects brand... ...and marketing channels. We are looking for a Staff Engineer, FinOps & Cost Platform to build and lead our engineering cost-management...Full timeLocal areaRemote workWork from homeHome office
$180k - $225k per year
...This role sits at the intersection of AI engineering and application development. You’ll build customer-facing features from scratch, while also helping us integrate and scale the AI infrastructure that power them. You're deeply fluent in TypeScript and React, strong enough...Full timeWork at officeRemote work- ...building a bespoke model per customer. Take inference to production scale. Move workloads from... .... • • Set technical direction for AI across the company, partnering closely... ...AI squads build, working across a Manila engineering organization, and spending real time with...Full timeInternshipWork at officeNight shift
$173.4k - $238.35k per year
...Databricks is on a mission to simplify and democratize data and AI — from making the next mode of transportation a reality to accelerating... ...can use deep data insights to improve their business. Founded by engineers — and customer obsessed — we leap at every opportunity to solve...Summer workWorldwide$130k - $180k per year
...decision. We bring Workforce AI to life for HR departments through... ...internal knowledge . As the Staff DevOps Developer on this... ...will bring a blend of software engineering discipline, deep cloud networking... ...Design, build, and optimize the inference-time retrieval service—from...Full timeWork at officeRemote work3 days per week- ...quality checks — none of it automated, most of it repeated daily. EviSmart's platform exists to replace that. The AI and automation infrastructure is the engine underneath everything: routing decisions, predictive analytics, intelligent process orchestration. The obstacle isn...Full timeWork at officeImmediate startWork from home
- ...Security D3 Security is transforming SecOps with Morpheus, our AI-driven Autonomous Security Operations Center (ASOC) platform.... ...Role Summary D3 Security is looking for a hands-on Agentic AI Engineer — a software engineer who designs, builds, and manages AI agents...Permanent employmentFull timeWork at officeMonday to friday
$184k - $205k per year
...Engineering Manager - Data & AI About Us Procurify is the Intelligent Spend Management company. We’re on a mission to give all organizations unprecedented visibility and control over their business spend. Our industry-leading platform unifies data across procurement...Long term contractFull timeImmediate startRemote workFlexible hours- ...Mythic is building the future of AI computing with breakthrough analog technology that delivers 100× the performance of traditional... ...large language models and CNNs to advanced signal processing, and is engineered to operate from –40 °C to +125 °C, making it ideal for industrial...Full time
- ...Intelligence that powers every people decision. We bring Workforce AI to life for HR departments through our award-winning, agentic AI... ...end-to-end—driving the overarching strategy, tooling engineering, detection content, and incident response processes. Operating within...Full timeWork at officeImmediate start3 days per week
$123k - $160k per year
...make growth measurable. Our unparalleled attribution, backed by AI-enhanced linking, is trusted to deliver seamless experiences that... ...About The Group We're hiring an AI DevOps & Reliability Engineer to own how software ships and runs at Branch. The role has two areas...Remote jobFull timeInternshipRelocation- ...Senior Software Engineer, AI Enablement The role & impact You'll join a small team that builds the tools, platforms and best practices helping engineers across Xero work with AI. Rather than shipping something once and moving on, you'll own it end to end from early experimentation...Full time
- ...We are seeking a talented and motivated early-career software engineer to join the Match Group Core in a brand new role enabling greater... ...impart your expertise on your stakeholders, and enable them with AI to achieve their own goals. This role is a paid, fixed-term 4-...Hourly payPermanent employmentFull timeFixed term contractInternshipWork at officeWorldwide3 days per week
$100k per year
...Sr. AI Engineer Who We Are Comm100 is an award-winning digital customer engagement platform, enabling organizations to better engage, convert and support their customers online. Established in 2009, Comm100 serves over 10,000 clients globally including HP, Rackspace, Government...Work at office$186k - $236k per year
...The team / how they connect The Payments Platform team is the engine powering money movement across the entire ecosystem, collaborating... ...engineering solutions. You are a champion of advanced usage in AI tools and practices across the team to accelerate development workflows...InternshipWork at officeWork from homeRelocation2 days per week1 day per week- ...We are seeking a talented and motivated early-career software engineer to join the Match Group Core in a brand new role enabling greater... ...impart your expertise on your stakeholders, and enable them with AI to achieve their own goals. This role is a paid, fixed-term 4-...Permanent employmentFull timeFixed term contractInternshipWork at officeWorldwide3 days per week
$238k - $270k per year
...We're building AI Teammates: agents that work like actual users in Asana and integrated apps. They triage bugs, respond to requests,... ...from tracking work to getting work done. We're looking for an Engineering Manager to lead the Agent Orchestration team — the team building...Long term contractFull timeWork at officeLocal areaWork from homeWorldwideShift work- ...moment, every day. We collaborate and co-create to build intelligent, AI-driven solutions that protect financial payment systems from... ...About the Role We are looking for a Senior Machine Learning / AI Engineer to help design, build, test and deploy machine learning and AI-...Long term contractFull time
$150k - $195k per year
.... This role is a split between data platform architecture and AI infrastructure. Half your time will go toward scaling our Snowflake... ...to set the technical standard for the team and mentor the engineers and analysts already here. You will report directly to the VP of...Full timeLocal areaFlexible hours$100k - $150k per year
...Category Engineering Hire Type Employee Job ID 17397 Base Salary Range $100000-$150000 Remote Eligible No Date Posted 05/19/2026... ...engineering solutions from silicon to systems, helping customers innovate AI-powered products faster through silicon design, IP, simulation,...Long term contractRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Staff Engineer - AI/Inference /Gateway. Be the first to apply!
