Sign up to access all features of our service.
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

Full Stack LLM Engineer

Full-time

Cerebras Systems

Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. Our novel wafer-scale architecture provides the AI compute power of dozens of GPUs on a single chip, with the programming simplicity of a single device. This approach allows Cerebras to deliver industry-leading training and inference speeds and empowers machine learning users to effortlessly run large-scale ML applications, without the hassle of managing hundreds of GPUs or TPUs.  

Cerebras' current customers include top model labs, global enterprises, and cutting-edge AI-native startups.  OpenAI recently announced a multi-year partnership with Cerebras , to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference. 

Thanks to the groundbreaking wafer-scale architecture, Cerebras Inference offers the fastest Generative AI inference solution in the world, over 10 times faster than GPU-based hyperscale cloud inference services. This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.

About the Role
We are seeking a versatile and experienced engineer to join our Inference Core Model Bringup team. This team is responsible to rapidly bring up state-of-the-art open-source models (like LLaMA, Qwen, etc) or customer-provided proprietary models on our Cerebras CSX systems. Success in this role requires a system-minded generalist who thrives in fast-paced bringup environments and is comfortable working across the entire Cerebras software stack.
Your work will play a critical role in achieving unprecedented levels of performance, efficiency, and scalability for AI applications.

Responsibilities


  • Contribute to the end-to-end bring up of ML models on Cerebras CSX systems.

  • Work across the stack: model architecture translation, graph lowering, compiler optimizations, runtime integration, and performance tuning.

  • Debug performance and correctness issues spanning model code, compiler IRs, runtime behavior, and hardware utilization.

  • Propose and prototype improvements across tools, APIs, or automation flows to accelerate future bring ups.

Skills & Qualifications


  • Bachelor’s, Master’s, or PhD in Computer Science, Engineering, or a related field.

  • Comfort navigating the full AI toolchain: Python modeling code, compiler IRs, performance profiling, etc.

  • Strong debugging skills across performance, numerical accuracy, and runtime integration.

  • Experience with deep learning frameworks (e.g., PyTorch, TensorFlow) and familiarity with model internals (e.g., attention, MoE, diffusion).

  • Proficiency in C/C++ programming and experience with low-level optimization.

  • Proven experience in compiler development, particularly with LLVM and/or MLIR.

  • Strong background in optimization techniques, particularly those involving NP-hard problems.

What We Offer


  • Competitive salary and benefits package.

  • Opportunities for professional growth and career advancement.

  • A dynamic and innovative work environment.

  • The chance to work on cutting-edge technologies and make a significant impact on the future of AI.

 

Why Join Cerebras


People who are serious about software make their own hardware. At Cerebras we have built a breakthrough architecture that is unlocking new opportunities for the AI industry. With dozens of model releases and rapid growth, we’ve reached an inflection point in our business. Members of our team tell us there are five main reasons they joined Cerebras:


  1. Build a breakthrough AI platform beyond the constraints of the GPU.

  2. Publish and open source their cutting-edge AI research.

  3. Work on one of the fastest AI supercomputers in the world.

  4. Enjoy job stability with startup vitality.

  5. Our simple, non-corporate work culture that respects individual beliefs.

Read our blog:  Five Reasons to Join Cerebras in 2026.

Apply today and become part of the forefront of groundbreaking advancements in AI!


Cerebras Systems is committed to creating an equal and diverse environment and is proud to be an equal opportunity employer.  We celebrate different backgrounds, perspectives, and skills. We believe inclusive teams build better products and companies. We try every day to build a work environment that empowers people to do their best work through continuous learning, growth and support of those around them.

This website or its third-party tools process personal data. For more details, click here to review our CCPA disclosure notice.

Vacancy posted 13 hours ago
Similar jobs that could be interesting for youBased on the Full Stack LLM Engineer in Toronto, ON vacancy
  •  ...across early and growth-stage companies.StartupFuel is hiring a Full-Stack Software Engineer to build and scale the SaaS platform behind DiligenceGPTâ„...  ...platform Close collaboration with AI engineers shipping LLM features Exposure to enterprise customers and investor... 
    Suggested
    Full time
    Internship

    Startupfuel

    Toronto, ON
    13 hours ago
  • $140k - $160k per year

     ...and we’re ready to continue growing.   As a Senior Software Engineer - Full Stack with Perpetua, you will contribute to our core platform built...  ...analytics workloads.  Experience with machine learning systems, LLM integrations, or AI-powered products.  Experience mentoring... 
    Suggested
    Full time
    Work at office
    Local area
    Immediate start

    Flywheel Digital

    Toronto, ON
    13 hours ago
  • $145k - $165k per year

     ...organisations that want to plan smarter and move faster. As a Senior Engineer on the FP&A Plus platform, you will design, build, and maintain...  .../OIDC, JWT, SSO, and MFA. Experience building or consuming LLM/agent tooling (MCP or comparable) — an emerging and actively... 
    Suggested
    Full time
    Internship

    Prophix

    Toronto, ON
    13 hours ago
  • $90k - $110k per year

     ...Build enterprise systems for a globally recognized consumer health brand.   We’re hiring a  Full Stack Engineer  to join the enterprise applications team of a globally recognized Canadian organization in the consumer health and manufacturing space, with products distributed... 
    Suggested
    Long term contract
    Full time
    Work at office
    1 day per week

    Stack It Recruitment

    Toronto, ON
    13 hours ago
  •  ...health systems to use AI to its full potential and maximize impact...  ...directly with the Director of Engineering, and the technical decisions you...  ...production AI systems: you have taken LLM-powered or ML-powered features...  ...experience to work across the stack when a feature needs it   -... 
    Suggested
    Full time
    Manual labor
    Work at office
    2 days per week
    1 day per week

    Signal 1

    Toronto, ON
    13 hours ago
  •  ...the digital landscape. ROLE OVERVIEW We’re looking for full stack engineers who are not just strong technically, but are actively...  ...Design and implement AI-powered features , including: LLM integrations Intelligent workflows AI-assisted user experiences... 
    Full time
    Shift work

    Tribalscale

    Toronto, ON
    13 hours ago
  • $130k - $150k per year

     ...We're looking for a Senior Full Stack Software Engineer who is passionate about building high-quality software in an AI-native way. Someone who...  ...to ambiguous technical challenges and can evaluate when an LLM-in-the-loop is the right solution versus traditional software... 
    Full time
    Casual work

    Altaml

    Toronto, ON
    13 hours ago
  • $130k - $160k per year

     ...building innovative systems that scale. A network agnostic tech stack. Agile, cross-functional teams built on trust and mutual...  ...standards — we want to hear from you. We're looking for a full-stack JS engineer to help us scale and connect millions of devices to wireless... 
    Full time
    Work from home
    Worldwide
    Flexible hours

    Us Mobile

    Toronto, ON
    13 hours ago
  •  ...Description Basetwo provides manufacturing engineers with a low code AI platform that helps them troubleshoot and optimize their production...  ...Ingest and connect siloed databases The Role As a Full Stack Engineer, you will be working on one or more of our new product... 
    Long term contract
    Full time
    Work from home
    Flexible hours

    Basetwo Pty Ltd

    Toronto, ON
    13 hours ago
  • $85k - $225k per year

     ...a positive impact on its customers, employees, and communities. The Role Our teams are hiring multiple talented Full-Stack Software Engineers to build innovative products in Life Sciences. You'll work with the latest front-end and back-end technologies as we tackle... 
    Remote job
    Full time
    Work at office
    Local area
    Work from home

    Veeva Systems

    Toronto, ON
    13 hours ago
  •  ...We are currently looking for a Senior Full Stack Engineer to join our growing Web Applications team. Reporting to the Engineering Manager, you'll work on complex projects, actively contribute to Conversational AI, back-end and front-end initiatives, and grow along with other... 
    Full time
    Immediate start

    Benchsci

    Toronto, ON
    13 hours ago
  • $90k - $130k per year

     ...capacity. Several of our largest workstreams are carried by a single engineer each, and a significant wave of integration work is already...  ...say yes to what is coming. We're hiring an Intermediate Full-Stack Engineer for the Gravity pod, focused squarely on execution and... 
    Long term contract
    Full time
    Work at office
    Remote work
    Home office
    2 days per week
    3 days per week

    Medme Health

    Toronto, ON
    13 hours ago
  •  ...collect is becoming richer and more complex. We need a data platform engineer to help us design and operate the systems that turn that data...  ...of scale, performance, and strategy. We’re open to part or full-time. Ideal for builders who care about performance and precision... 
    Full time
    Part time
    Flexible hours

    Reflow

    Toronto, ON
    13 hours ago
  • $116k - $156k per year

     ...Job description The Opportunity We're hiring an Automation-Focused Full-Stack Engineer for the Gravity pod to spearhead the evolution of our core automation product. The gravity pod owns some of MedMe's most critical integration workstreams — connecting pharmacies and... 
    Long term contract
    Full time
    Internship
    Work at office
    Immediate start
    Remote work
    Work from home
    Home office
    2 days per week
    3 days per week

    Medme Health

    Toronto, ON
    13 hours ago
  • $120k - $150k per year

     ...scalable systems that help our games and teams perform at a higher level. About the Role Big Viking Games is hiring a Senior Full Stack Engineer to build AI-enabled products, workflows, tools, and systems that help turn creative direction into production-ready game... 
    Long term contract
    Full time
    Internship
    Work at office
    3 days per week

    Big Viking Games

    Toronto, ON
    13 hours ago
  • $90k - $125k per year

     ...management system for our customers so they can spend more time doing what they love. The Role: We’re looking for a Full Stack Software Engineer to join our growing Engineering team! In this role, you will play a key part in building and scaling our payroll platform,... 
    Full time
    Work at office
    Remote work

    Push Operations

    Toronto, ON
    13 hours ago
  •  ...who are underserved by the traditional financial system.   Position Summary: Haventree is looking for a Senior Software Engineer ( Full Stack ) to join our fast-growing team. In this role, you will design and build end-to-end solutions that power our customer-facing web... 
    Full time
    Flexible hours

    Haventree Bank

    Toronto, ON
    13 hours ago
  •  ...STAN is looking for a full-time Senior Software Engineer who is an enthusiastic problem solver to help create the technology that will change the way...  ...passionate about building a business ~5+ years of full-stack development experience with JavaScript and React libraries... 
    Full time
    Work at office

    STAN AI

    Toronto, ON
    13 hours ago
  • $100k - $125k per year

    1-800-GOT-JUNK? is seeking a Senior Software Engineer (Full Stack) to join our Product & Technology team in Vancouver or Toronto . You will sit on the Omni Channel Product Team, building commercial customer and field-facing products end to end: not only implementing tickets... 
    Full time
    Internship
    Work at office
    Remote work

    O2e Brands

    Toronto, ON
    13 hours ago
  • $50 - $80 per hour

     ...Benchmark , General Catalyst , Peter Thiel , Adam D'Angelo , Larry Summers , and Jack Dorsey . Position: Full-Stack Software Engineer Type: Contract Compensation: $50–$80/hour Location: Remote Role Responsibilities Write... 
    Full time
    Contract work
    Summer work
    Remote work

    Mercor

    Toronto, ON
    13 hours ago
  •  ...Position Description: Senior Full Stack S oftware Engineer   Reporting to: Director of Application Engineering     Signal 1 helps health systems accelerate AI adoption with a category defining technology platform. Signal 1’s first product, the AI Management System (AIMS... 
    Full time

    Signal 1

    Toronto, ON
    13 hours ago
  •  ...interview philosophy and how we use AI in our recruiting process here . We are looking for inquisitive, well-rounded Full-stack engineers to join our Core or Monetization Engineering teams. Working closely with product managers, designers, and backend engineers, you... 
    Full time
    Work at office
    Relocation
    Relocation package

    Pinterest

    Toronto, ON
    13 hours ago
  •  ...Design and implement subscription management features across our full-stack architecture. Build and maintain RESTful APIs using Restify/...  ...comprehensive unit tests using Jest. Collaborate with staff engineers on system design and architectural improvements. Participate... 
    Full time
    Work at office
    Remote work
    Worldwide
    Flexible hours

    E2x Is Now Part Of Apply Digital

    Toronto, ON
    13 hours ago
  •  ...learning never stops. As a Senior Product Engineer at Docebo , you operate at the...  ...needs and product delivery. You'll own the full lifecycle of the features you build — from...  ...real data. You'll combine strong full stack engineering skills, with a focus on front-... 
    Full time
    For contractors
    Work at office
    Worldwide
    3 days per week

    Docebo

    Toronto, ON
    13 hours ago
  •  ...Snowflake's security, governance, lineage, and access controls by default. See Deploy Faster with Snowflake Apps . As a Senior Full Stack Engineer, you will drive initiatives that span our product areas and tech stack. It’s a high-impact role - you will work on an early-... 
    Full time
    Internship

    Snowflake

    Toronto, ON
    13 hours ago
  •  ...us Appnovation is a global, full-service digital partner that combines...  ...Strategy, Experience & Design, Engineering and Managed Services. We build...  ...ROLE OVERVIEW: As a Full-Stack Developer, you will architect...  ...applications: integrating LLM APIs for validation, decision-making... 
    Full time

    Appnovation Technologies

    Toronto, ON
    13 hours ago
  •  ...About the Position:   As a member of the LLM inference team, you will help build state-...  ...and implementing the best inference stacks in the LLM world? Work and collaborate with...  ...orchestration, distributed systems, inference engine optimization, and writing high-performance... 
    Full time
    Internship
    Flexible hours

    Centml

    Toronto, ON
    13 hours ago
  •  ...you in? Your Impact Starts Here We’re looking for a Full Stack Software Engineer II to join our Hiring Assistant team —one of the most...  ...(screening, matching, automation) Integrate and apply AI/LLM tools, agents, and automation frameworks to improve product... 
    Hourly pay
    Long term contract
    Full time
    Temporary work
    Internship
    Work at office
    Local area
    Work from home
    Flexible hours

    Homebase

    Toronto, ON
    6 hours ago
  • $116k - $156k per year

     ...prototype to production, and we need a senior engineer who can own high-impact work end-to-end....  ...assistant experiences, or a core LLM/voice platform subsystem) within your first...  ...API design, data modelling Practical full-stack fluency — comfortable contributing in React... 
    Long term contract
    Full time
    Work at office
    Remote work
    Work from home
    Home office
    2 days per week
    3 days per week

    Medme Health

    Toronto, ON
    13 hours ago
  • $115k - $125k per year

     ...experiment, and not as a replacement for engineering judgment. We move faster because of these...  ...— an open architectural decision spanning LLM query layers, Atlas vector search, and existing...  ...We're Looking For You move across the stack without losing altitude — Angular MFE one... 
    Remplacement
    Full time
    Internship
    Work at office
    Local area
    2 days per week

    Triparc

    Toronto, ON
    13 hours ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to Full Stack LLM Engineer. Be the first to apply!