AI Inference Engineer
$165k - $330k per yearUrban Ridge Supplies
About Baseten
Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F, led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team:- Forward Deployed Engineering on the frontier of AI
- The fastest, most accurate Whisper transcription
- Deploy production-ready model servers from Docker images
- Deploy custom ComfyUI workflows as APIs
- Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects.
- Drive customer impact by designing, implementing, and deploying Baseten solutions end-to-end (problem framing → evaluation → production deployment → monitoring). This involves working with customers’ engineering teams at every stage of the customer journey including: sales, implementation, and expansion.
- Deliver with velocity: turn vague objectives into clear specs and well-defined PoCs so we can rapidly ship well-tested services and outcomes for our customers
- Optimize and enhance AI/ML projects, contributing to the continuous improvement of our technical stack. This includes developing features and PRDs with other engineering and product orgs.
- Own products and customer projects end-to-end, functioning as both an engineer, project manager, and product manager, with a focus on user empathy, project specification, and end-to-end execution.
- Navigate ambiguity and exercise good judgment on tradeoffs and tools needed to solve problems, avoiding unnecessary complexity.
- Demonstrate pride, ownership, and accountability for your work, expecting the same from your teammates.
- Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field.
- 2+ years of professional work experience in a fast-paced, high-growth environment.
- Demonstrated experience with one or more general-purpose programming languages in a production-level environment, with a strong preference for Python.
- Familiarity with AI/ML pipelines and the lifecycle of ML model development and deployment.
- Strong communication skills, particularly on complex technical topics.
- Experience in building or optimizing AI/ML projects is highly valued.
- Competitive compensation, including meaningful equity
- (U.S. only) 100% coverage of medical, dental, and vision insurance for employee and dependents
- Flexible PTO policy including company wide Winter Break (our offices are closed from Christmas Eve to New Year's Day!)
- Paid parental leave
- Fertility and family-building stipend through Carrot
- (U.S. only) Company-facilitated 401(k)
- Exposure to a variety of ML startups, offering unparalleled learning and networking opportunities.
Job Id: cOQ8RJB4VlJF9W7Qq59Z6Wnw9J1w+SrGBFf/YYs9szGdBLv7H04Ydy/aI2zJTR0nwQrzIvnfNqNediICXTY2JfFsof5YdFFDxA==
Vacancy posted 9 hours ago
Similar jobs that could be interesting for youBased on the AI Inference Engineer in Toronto, ON vacancy
$100k - $145k per year
...Thomson Reuters is seeking a Senior Inference Engineer, AI. This person will collaborate with platform teams to enhance capacity forecasting for AI workloads and work with Product, Data Science, Architecture, and Enterprise AI teams to onboard new research models into production...SuggestedFull timeWork at officeLocal areaFlexible hours- ...About the Role We are a small team of AI builders in Paytm Labs. As a Staff AI Platform Engineer, you will work across inference and agentic systems. You will contribute to Paytm's AI inference platform (Pi), serving internal teams and enterprise customers - running...SuggestedFull time
- ...About the Role We are a small team of AI builders in Paytm Labs. As a Staff AI Platform Engineer, you will work across inference and agentic systems. You will contribute to Paytm's AI inference platform (Pi), serving internal teams and enterprise customers - running...SuggestedFull time
$190k - $225k per year
...AssemblyAI builds the best-in-class Voice AI models powering the next generation of voice applications. Our models serve 600M+ inference calls monthly, process 1M+ hours of audio... ...About the role: We're hiring a Software Engineer to help turn cutting-edge AI research into...SuggestedInternship- ...Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This... ...to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based... ...ultra-fast decode on the Cerebras Wafer-Scale Engine. We are hiring a Software Engineer to...SuggestedFull time
- ...developers and enterprises who are building AI systems to power magical experiences like... .... Cohere is a team of researchers, engineers, designers, and more, who are passionate about... ...they influence latency and throughput of inference. ~ Strong understanding or working experience...Full timeWork at officeRemote workFlexible hours
$90k per year
...cutting-edge cloud platform, coupled with AI-driven analytics tools, unlocks the true essence... ...a highly skilled and experienced Software Engineer, AI to perform a key role in our digital... ..., and safety standards. Build data and inference pipelines Develop pipelines for...Full timeFlexible hours$150k - $160k per year
...Requisition Number: 105712 AI Engineer Location You will have the flexibility to work fully remotely. Insight at a Glance ~14,000+ engaged teammates globally ~$8.2 billion in revenue in 2025 ~ Certified as a Great Place to Work in 9 Countries in 2025 ~ Fortune...Full timeLocal areaRemote work- ...Nexxa is building the best AI systems for heavy industries — enabling machines, systems... ...Overview We're looking for Backend AI Engineers to design, build, and own the core AI infrastructure... ...at scale: model-serving pipelines, inference and orchestration layers, data pipelines,...Long term contractFull time
$110k - $150k per year
...across the capital markets with innovative, AI-driven technology and expert advisory... ...Opportunity We are looking for an Applied AI Engineer to join the AI Lab team within... ...engineering, RAG, fine-tuning, agents, or inference pipelines ~Data engineering: designing...InternshipLocal areaRemote workFlexible hours- ...with the teams and customers who depend on AI — risk, fraud, collections, payments,... ...build the platform underneath: Paytm’s AI inference platform (Pi) and the agentic runtime, orchestration... ...and standards for agentic systems; mentor engineers and partner with ML, product, and security...Full time
$135k - $170k per year
...General Information: Job Title: Senior Engineer, Applied AI Location: Toronto, ON (Onsite/Hybrid) Job Type: Full-Time Hiring Timeline... ...Nice-to-Have: Fine-tuning experience (LoRA, SFT, DPO) Inference stack experience (vLLM, TGI, llama.cpp) Observability...Long term contractFull timeInternshipWork at officeImmediate startRemote workFlexible hours$132.69k - $182.69k per year
...primarily focuses on the design, review of AI Security architecture and as well as the... ...-aware security culture across internal engineering and product teams. Knowledge/Skills/Competencies... ...and inversion attacks Membership inference attacks Prompt injection (direct and...Temporary workInternshipWork at officeLocal areaRemote workNight shift$200k - $350k per year
...based in Canada. Pay: $200,000–$350,000/yr Job Title: AI/ML Engineer Job Type: Full-Time About Us: the company is the end... ...ontologies to ensure high-quality data for AI training and inference. # Create and maintain REST APIs and SDK integrations to facilitate...Remote jobFull time- ...flexibility for meaningful work-life balance. Being a Senior AI Engineer at iManage Means… You are passionate about building and... ...production ~ Experience with GPU optimization for training and inference workloads ~ Experience with Kubernetes deployment on cloud infrastructure...Full timeWork at officeLocal areaWorldwideFlexible hours
$135k - $170k per year
General Information: Job Title: Senior Engineer, Applied AI Location: Toronto, ON (Onsite/Hybrid) Job Type: Full-Time Hiring Timeline... ...-to-Have: ~Fine-tuning experience (LoRA, SFT, DPO) ~Inference stack experience (vLLM, TGI, llama.cpp) ~Observability...Long term contractFull timeInternshipWork at officeImmediate startRemote workFlexible hours- ...Reference Number : R2868511 Position title : AI Engineer Department: Commercial Data Science Location: Toronto,ON ( Flexible... ...and machine learning, to analyze data and make predictions, inferences, decisions, or recommendations without direct human involvement...Work at officeWork from homeHome officeFlexible hours
$135k - $210k per year
...Overview: Guidepoint seeks an experienced AI Engineer as an integral member of the Toronto-based AI team. The Toronto Technology Hub serves as the base of our Data/AI/ML team, dedicated to building a modern data infrastructure for advanced analytics and the development...Full timeWorldwide$105k - $215k per year
...& Reporting The Team - We accelerate BMO’s AI journey by building cloud-native AI solutions. Our team combines engineering excellence with cutting-edge AI to deliver scalable... ...for AI infra: GPU rightsizing, caching, inference optimization, cost governance Application &...Permanent employmentFull timeContract workPart time- ...product development rooted in digital strategy, experience design, and engineering, with a unique combination of agility, scale, and maturity. We... ..., life sciences, ecommerce, and education. Job Title: AI Engineer Type: Hybrid Location: Mississauga, ON Project...Full timeRemote work
- ...Reference no. R2867104 Position title: Lead Data and AI Engineer Department: US Commercial Data Engineering Location: Toronto... ...algorithms and machine learning, to analyze data and make predictions, inferences, decisions, or recommendations without direct human involvement...Long term contractWork from home
- ...improve how work gets done. A key part of that vision is Reflow AI, agentic, automation-focused intelligence that adapts to how people... ...automation and workflow intelligence. Experiment with prompt engineering, context tuning, and model orchestration to improve performance and...Full timePart timeFlexible hours
- ...work alongside a great team in a high-growth environment, FacilityOS is the place to build your career. About the Role The AI Engineer reports to the VP of Engineering and serves as FacilityOS's hands-on technical anchor for AI agent development. This role sits...Full timeWork at office
$60 - $120 per hour
...check your spam folder. This is a fully remote position open to candidates based in Canada. Pay: $60$120/hr Job Title: AI Engineer Job Type: Contractor (15 hrs a week) Location: Remote Schedule: Flexible you pick the hours and days (including weekends...Full timeFor contractorsRemote workFlexible hours$120k - $170k per year
...About the Role We're looking for a Senior / Principal AI Platform Engineer to build the secure, scalable platform capabilities that enable... ...AI platform services including model gateways, API gateways, inference routing, vector and retrieval services, evaluation pipelines,...Long term contractPermanent employmentFull timeFlexible hours- ...proprietary generative media models and AI native creative workflows, tackling unsolved... ...taste, and craft as much as research and engineering – shipping experiences that creatives... ...numerical stability across data, training, and inference ~ You’ve taken ML features from idea to...Full timeWork at office
$85 per hour
...creative and technical talent with leading AI research labs. Headquartered in San... ..., and Jack Dorsey . Position: ML Engineer (Coding Agent Experience) Type: Contract... ...implementations involving model training , inference systems , MLOps , and LLM applications...Full timeContract workSummer workRemote work- ...Zafin is an AI platform company helping regulated institutions modernize how critical work is designed, governed, and delivered. Our... ...responsibly and at scale. What’s the Opportunity? As an A I Engineer , p art of Zafin’s AIOS Platform or Forward Deploy Unit s...Full time
- ...Position: Senior AI Engineer Location: Toronto, ON Job ID#: RQ11252 Duration: 8 Months Role Overview We are seeking a Senior AI Engineer to design, develop, and support advanced AI, Generative AI, and Conversational AI solutions. This role focuses...Full time
- ...Palona’s AI agents operate in real restaurant environments: noisy phone lines, varied... ...We are looking for an applied AI Modeling Engineer to improve the intelligence, accuracy, safety... ..., model services, or training and inference pipelines. ~ Strong software engineering...Long term contractFull timeTemporary workInternshipImmediate start
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to AI Inference Engineer. Be the first to apply!
Related searches
