Data Engineer RD-017
$75k per yearIrth Solutions Inc
- Remote job
Data Engineer (Mid-level)
Location: Remote (Quebec/Canada)
About Irth Solutions
Irth Solutions is a leading provider of cloud-based SaaS software for damage prevention, asset integrity, stakeholder engagement and land management, helping energy, utility, telecom, and infrastructure companies protect their critical network infrastructure. With nearly three decades of industry experience, Irth serves customers across North America and continues to expand its platform with new data-driven and AI-powered capabilities.
About the Role
We are looking for a Data Engineer to design, build, and maintain data ingestion and processing pipelines in Databricks, transforming high-volume external data sources into clean, structured, reliable inputs for downstream analysis and intelligence. This role will primarily support our Stakeholder Engagement offering, working closely with our Data Scientist and application teams.
Key responsibilities
1. Data Pipeline Development (Primary Responsibility)
- Design, build, and maintain ingestion pipelines from high-volume external APIs, capable of running continuously and reliably at scale.
- Implement ingestion and transformation workflows using Databricks (Spark/PySpark, SQL, Delta Live Tables), applying medallion architecture patterns (Bronze → Silver → Gold) to move from raw ingested content to clean, structured, analysis-ready data.
- Build the infrastructure for deduplication and relevance filtering of incoming content, implementing filtering logic and quality criteria defined in collaboration with the Data Scientist.
- Implement schema evolution handling and data validation rules as data sources and formats change over time.
2. Platform & Storage Implementation
- Configure and manage Delta Lake storage structures, tables, partitions, and optimization routines (OPTIMIZE, Z-ORDER, VACUUM).
- Design and evolve data schemas that balance query performance, cost, and maintainability as data volume grows.
- Maintain clear metadata and documentation of table structures to support easy consumption by the Data Science and application teams.
3. Reliability, Monitoring & Operational Support
- Ensure pipeline reliability and observability: error handling, retries, monitoring, and alerting for a continuously running system.
- Adapt pipelines to evolving external API contracts, rate limits, authentication changes, and new data sources.
- Troubleshoot pipeline failures, perform recovery, and tune performance as needed.
4. Orchestration & Automation
- Build, schedule, and monitor workflows using Databricks Workflows, Delta Live Tables, or similar orchestration tools.
- Contribute to CI/CD pipelines for code deployment, versioning, and environment management.
5. Collaboration & Documentation
- Work closely with the Data Scientist to expose clean, well-structured data feeding LLM/NLP pipelines and downstream models.
- Participate in technical decisions around data architecture and propose structuring solutions as the team's needs evolve.
- Document pipelines, data dictionaries, job schedules, and transformation logic.
- Support the onboarding of new data sources and pipelines as the product expands to additional solution areas.
Ingénieur(e) de données (Intermédiaire)
Lieu : Télétravail (Québec/Canada)
À propos d'Irth Solutions
Irth Solutions est un fournisseur de premier plan de logiciels SaaS infonuagiques pour la prévention des dommages, l'intégrité des actifs, la mobilisation des parties prenantes et la gestion foncière, aidant les entreprises des secteurs de l'énergie, des services publics, des télécommunications et des infrastructures à protéger leurs réseaux d'infrastructure critiques. Avec près de trois décennies d'expérience dans l'industrie, Irth dessert des clients à travers l'Amérique du Nord et continue d'élargir sa plateforme avec de nouvelles capacités axées sur les données et l'intelligence artificielle.
À propos du poste
Nous sommes à la recherche d'un(e) ingénieur(e) de données pour concevoir, construire et maintenir des pipelines d'ingestion et de traitement de données dans Databricks, transformant des sources de données externes à haut volume en données propres, structurées et fiables pour l'analyse et l'intelligence en aval. Ce poste appuiera principalement notre offre de mobilisation des parties prenantes (Stakeholder Engagement), en collaboration étroite avec notre data scientist et les équipes de développement applicatif.
Responsabilités principales
1. Développement de pipelines de données (responsabilité principale)
- Concevoir, construire et maintenir des pipelines d'ingestion à partir d'API externes à haut volume, capables de fonctionner de façon continue et fiable à grande échelle.
- Implémenter des flux d'ingestion et de transformation avec Databricks (Spark/PySpark, SQL, Delta Live Tables), en appliquant les patrons d'architecture medallion (Bronze → Silver → Gold) pour transformer le contenu brut ingéré en données propres, structurées et prêtes pour l'analyse.
- Construire l'infrastructure de déduplication et de filtrage de pertinence du contenu entrant, en implémentant la logique de filtrage et les critères de qualité définis en collaboration avec le data scientist.
- Implémenter la gestion de l'évolution des schémas et des règles de validation des données, à mesure que les sources et formats de données évoluent.
2. Implémentation de la plateforme et du stockage
- Configurer et gérer les structures de stockage Delta Lake, les tables, le partitionnement et les routines d'optimisation (OPTIMIZE, Z-ORDER, VACUUM).
- Concevoir et faire évoluer des schémas de données qui équilibrent performance des requêtes, coûts et facilité de maintenance à mesure que le volume de données croît.
- Maintenir une documentation claire des métadonnées et des structures de tables afin de faciliter leur utilisation par les équipes de data science et de développement applicatif.
3. Fiabilité, surveillance et soutien opérationnel
- Assurer la fiabilité et l'observabilité des pipelines : gestion des erreurs, mécanismes de reprise (retries), surveillance et alertes pour un système fonctionnant en continu.
- Adapter les pipelines aux changements des contrats d'API externes, des limites de taux (rate limits), des méthodes d'authentification et à l'ajout de nouvelles sources de données.
- Diagnostiquer les défaillances des pipelines, effectuer les reprises nécessaires et optimiser la performance au besoin.
4. Orchestration et automatisation
- Construire, planifier et surveiller des flux de travail avec Databricks Workflows, Delta Live Tables ou des outils d'orchestration équivalents.
- Contribuer aux pipelines CI/CD pour le déploiement du code, la gestion des versions et des environnements.
5. Collaboration et documentation
- Collaborer étroitement avec le data scientist pour fournir des données propres et bien structurées alimentant les pipelines LLM/NLP et les modèles en aval.
- Participer aux décisions techniques liées à l'architecture des données et proposer des solutions structurantes à mesure que les besoins de l'équipe évoluent.
- Documenter les pipelines, les dictionnaires de données, les calendriers d'exécution et la logique de transformation.
- Appuyer l'intégration de nouvelles sources et pipelines de données à mesure que le produit s'étend à d'autres domaines de solutions.
Exigences
Forte préférence pour les candidat(e)s résidant au Québec, la maîtrise du français (parlé et écrit) constituant un atout important en plus de l'anglais.
RequirementsStrong preference for candidates residing in Quebec, with fluency in French (spoken and written) as a strong asset in addition to English
Required qualifications
- 3 to 5 years of experience in data engineering, with solid experience building and operating production-grade data pipelines.
- Familiarity with data modeling, data quality, and schema evolution.
- Solid understanding of data pipeline reliability practices: monitoring, alerting, and handling failures gracefully in a continuously running system.
- Hands-on experience with Databricks (or an equivalent Spark-based environment): schema design, Delta Lake, performance tuning, and pipeline orchestration.
- Experience with at least one major cloud (Azure preferred; AWS/GCP also beneficial).
- Experience integrating with external APIs at scale: authentication, pagination, rate limiting, retries, error handling.
- Strong proficiency in Python and SQL
- Comfortable working with unstructured/semi-structured text data at scale.
Nice to have Qualifications
- LLM prompting experience and/or basic understanding of AI/NLP concepts
- Exposure to medallion architecture or lakehouse best practices.
- Experience with orchestration frameworks (ADF, Workflows, Airflow, DBX, etc.).
- Experience with CI/CD tools and version control (Git, GitHub Actions or equivalent).
- Basic understanding of security practices: RBAC, encryption, credential management.
- Databricks certification (Data Engineer Associate or equivalent).
AI Use in Hiring
As part of our hiring process, this role may use artificial intelligence or automated tools to assist with reviewing and screening applications. These tools support, but do not replace, human judgment in making hiring decisions.
Compensation
The salary range for this role is CAD $75000
This range reflects the base salary only and does not include any additional compensation components. Any offer of employment is dependent on several factors, including, but not limited to, the candidate’s experience, skills, qualifications, and location.
Qualifications requises
- 3 à 5 ans d'expérience en ingénierie de données, avec une solide expérience dans la construction et l'exploitation de pipelines de données en production.
- Connaissance de la modélisation de données, de la qualité des données et de l'évolution des schémas.
- Bonne compréhension des pratiques de fiabilité des pipelines de données : surveillance, alertes et gestion des défaillances dans un système fonctionnant en continu.
- Expérience pratique avec Databricks (ou un environnement équivalent basé sur Spark) : conception de schémas, Delta Lake, optimisation de la performance et orchestration de pipelines.
- Expérience avec au moins un fournisseur cloud majeur (Azure de préférence; AWS/GCP également un atout).
- Expérience dans l'intégration d'API externes à grande échelle : authentification, pagination, limites de taux, mécanismes de reprise, gestion des erreurs.
- Forte maîtrise de Python et SQL.
- À l'aise avec le traitement de données textuelles non structurées ou semi-structurées à grande échelle.
Atouts
- Expérience en prompt engineering avec des LLM et/ou connaissances de base en IA/NLP.
- Connaissance de l'architecture medallion ou des meilleures pratiques lakehouse.
- Expérience avec des frameworks d'orchestration (ADF, Workflows, Airflow, DBX, etc.).
- Expérience avec des outils CI/CD et de contrôle de version (Git, GitHub Actions ou équivalent).
- Connaissances de base des pratiques de sécurité : RBAC, chiffrement, gestion des identifiants.
- Certification Databricks (Data Engineer Associate ou équivalent).
Utilisation de l'IA dans le processus de recrutement
Dans le cadre de notre processus de recrutement, ce poste peut faire appel à l'intelligence artificielle ou à des outils automatisés pour appuyer l'examen et la présélection des candidatures. Ces outils appuient le jugement humain dans la prise de décisions d'embauche, mais ne le remplacent pas.
Rémunération
L'échelle salariale pour ce poste se situe entre CAD $75000
Cette échelle reflète uniquement le salaire de base et n'inclut aucune autre composante de rémunération. Toute offre d'emploi dépend de plusieurs facteurs, incluant, sans s'y limiter, l'expérience, les compétences, les qualifications et le lieu de résidence du ou de la candidat(e).
BenefitsBenefits
- Competitive Salary – A competitive compensation package based on experience and qualifications.
- Medical, Dental, and Vision Insurance – Comprehensive insurance coverage to support you and your family.
- 401(k) Plan with Company Match.
- Generous Paid Time Off (PTO)– Time off to support work-life balance and personal needs.
- Company-Paid Holidays – Paid holidays throughout the year.
- Flexible Work Options – Work-from-home opportunities are available, depending on role and business needs.
- On-Call Compensation – Additional pay for eligible on-call shifts.
Avantages sociaux
- Salaire compétitif — Une rémunération concurrentielle basée sur l'expérience et les qualifications.
- Assurance médicale, dentaire et visuelle — Une couverture d'assurance complète pour vous et votre famille.
- Régime de retraite avec contribution de l'employeur .
- Congés payés généreux — Du temps libre pour favoriser l'équilibre travail-vie personnelle.
- Jours fériés payés par l'entreprise — Congés payés tout au long de l'année.
- Modalités de travail flexibles — Possibilité de télétravail selon le poste et les besoins de l'entreprise.
- Compensation pour disponibilité — Rémunération additionnelle pour les quarts de garde admissibles.
- ...practice resiliency, demonstrate leadership, go the extra mile for our customers, and empower our people to be their best. As a Data Engineer, you’ll help expand, optimize, and maintain data pipelines and infrastructure that power our IoT-enabled safety ecosystem. You’ll...SuggestedLong term contractFull timeInternshipFlexible hours
- ...with a vision to be more than an advertising platform, it’s a hub of innovation, imagination and creativity. We're looking to add Data Engineers to our data science team! This team works on solving complex problems for StackAdapt's digital advertising platform. You'll be...SuggestedRemote jobFull timeWork from homeHome office
$100k - $130k per year
...by placing patients at the center of the record release process, data moves more ethically, easily and securely throughout the care journey... ...for millions more patients across North America. As a Data Engineer, you will help build and maintain the data infrastructure that...SuggestedFull timeInternshipManual laborRemote work$140k - $160k per year
...Silver Lake Waterman, Moody’s, Sequoia Capital, GV and Riverwood Capital. About the Team: Be part of an exciting and growing Data Engineering team at the core of everything we do. SecurityScorecard's ratings platform processes enormous volumes of data — and the...SuggestedFull timeInternshipWorldwide$137k - $205k per year
...want to be part of the team that defines how the world eats, there is no better time to join us. About the Role As a Senior Data Engineer, you’ll play a key role in scaling and improving how we integrate and process customer data. You will design and implement ETLs that...SuggestedFull timeLive InRemote work- ...what makes your company great Job Description Key Responsibilities: Snowflake Expertise: Design and develop schemas and data models within the Snowflake platform. Write and optimize complex SQL queries for data extraction and transformation. Guidewire Claims...Full time
- ...doing this alone. About this role We are looking for a Data Engineer Intern to join our growing team for the Fall 2026 Term. In this... ...powered by many micro-services, including the bot platform, the NLP engine and the pricing/recommendation engines. We use a modern data...InternshipWork at officeRemote workHome officeNight shift
$160k - $205k per year
...want to challenge conventions in gaming, media and entertainment, we want to talk to you. About the Role & Team As a Staff Data Engineer, you'll partner closely with Data Services, Analytics Engineering, Product, and Business stakeholders to design and deliver...Remote jobFull time- ...unparalleled commitment to build faster, safer software and harness AI and data intelligence to mitigate risk, maximize efficiencies, and drive... ...software supply chains. We’re looking for a Staff Data Engineer to join our growing Data Platform team. You’ll play a key role...Long term contractFull timeFlexible hours
- ...Le Lead Data Engineer sera responsable de diriger la conception et le développement de solutions de données en environnements cloud (AWS, Azure). PLUS PRÃCISÃMENT ⢠Assurer la qualité, la gouvernance et la sécurité des données incluant les int...Full time
- ...AI-native professional services firm building operational AI and data systems. Our foundation beliefs are simple: optimization lowers... ...certain passion for bending time . The role As the Data Engineering Lead , you will build the critical data foundation that makes...Full timeManual laborRemote work
- ...What the Role Is The Data team at Babylist powers decision-making across the entire business — from product growth to operations to AI. This is a senior individual contributor role on the Data Engineering team, sitting at the intersection of platform thinking and AI-native...Full timeInternshipWork at officeRemote workFlexible hoursWeekend work
- ...supercharge it through the people, the work, and the programs that fuel who we are. About the role This role is part of our Data Engineering team. Being Data Driven is a core value of Super.com and this team plays a huge part in this, enabling the organization to...Full timeTemporary workRemote workFlexible hours
- ...our employees say about working with us on DOU . Domain: Data Infrastructure & Analytics / AdTech Client Location: Israel... ...is fast, product‑focused, and highly collaborative, with strong engineering leadership and a culture of ownership. We're looking for a...Full timeFor contractorsFlexible hours
$152k - $237.5k per year
...Jane. For the last year I've been responsible for building the engineering side of Jane's marketing platform from the ground up. Now we're... ...small, founding engineering team responsible for Jane's customer data infrastructure. Jane is used by thousands of clinics and hundreds...Long term contractFull timeInternshipRemote work- ...depend on physical operations to harness Internet of Things (IoT) data to develop actionable insights and improve their operations. At... ...long term. About the role: Samsara is seeking a Senior Data Engineer to join our Data team, comprising both Data Scientists, Data Analysts...Long term contractFull timeRemote workRelocation package
$169.2k per year
...Do you get excited to work with data & lead impactful teams? Then Jobber might be the... ...you! We’re looking for a Manager, Data Engineering to be part of our Data org. Jobber exists... ...including data stores, compute engines, and orchestration systems. Ensure data...Long term contractFull timeInternshipWork at officeLocal area$57.5k - $90k per year
...and investment remains anchored in long-term benefits to customers, members, and communities. Come join us. As an Associate Data Engineer at Gore Mutual Insurance, you have a strong technical background in software engineering / computer science and will be responsible...Long term contractFull time$79k - $85k per year
...Join us to do the best work of your career, solving meaningful problems with remarkable teams. Greenhouse is looking for a Data Migrations Engineer I to join our Data Migrations team! Reporting to the Manager, Data Migrations, you will work in tandem with Customers,...Full timeFor contractorsWork at officeLocal areaRemote work- ...on Deloitte's Technology Fast 500 again in 2024 – but we're just getting started. How You'll Make an Impact: As a Staff Data Engineer , you'll design and implement large-scale distributed data processing systems using technologies like Apache Hadoop, Spark, and Flink...Long term contractFull timeWork at officeLocal areaRemote workWork from homeHome officeWeekend work
- ...Our Why At Dotmatics At Dotmatics, we believe science, data, and decision-making must be deeply intertwined for innovation to thrive... ...What Do We Need? We're looking for a Senior Data Analytics Engineer who thrives in a fast-moving environment and enjoys solving complex...Full timeImmediate startRemote workFlexible hours
$142k - $189k per year
...distribute open-source software that enables people to enjoy the internet on their terms. About this team and role: Mozilla’s Data Engineering team is looking for a Senior Staff Data Engineer / Data Platform Engineer to help build the present and future of our Data...Remote jobFull timeImmediate startHome office- ...operations to harness Internet of Things (IoT) data to develop actionable insights and improve... ...About the role The Integrations, Data Engineering and AI (IDEA) team within Samsara’s... ...monitoring, and security (e.g., S3, IAM, RDS, Lambda, API-Gateway, VPC, EC2, ECS/EKS)....Long term contractFull time
- ...terms. Our HelixAI platform and Helix Pods delivery model put our engineers at the center of real agentic transformation — doing work that is... ...We are building the future of enterprise AI. Lead, Azure Data Engineer Embark on an exciting journey into the realm of...Full timeWork at officeRemote workFlexible hours
$170k - $220k per year
...Identity is where modern companies break. Too much access leaks data. Too little grinds the business to a halt. Lumos is the industry'... ...rich primitives that power every Lumos product. As a Software Engineer on our Data Platform team you will drive the evolution of our identity...Long term contractFull timeTemporary workLocal areaRemote workFlexible hours- Job Description (Key Responsibilities) Assist clients in delivering the overall security design and approach across complex SAP environments Lead the implementation of SAP security designs within client organizations Own and lead development and testing of SAP...Remote jobFull time
- ...and social events, so you’ll never feel like you’re doing this alone. About this role We are looking to hire a Software Engineer In Data intern to join our growing Data team. As a Software Engineer In Data at Super.com , your work will span from application development...InternshipWork at officeRemote workHome officeNight shift
- ...operations to harness Internet of Things (IoT) data to develop actionable insights and improve... ...About the role The Integrations, Data Engineering and AI (IDEA) team within Samsara’s... ...API management tools. ~ RDBMS: MySQL, AWS RDS/Aurora MySQL, PostgreSQL, Oracle or equivalent...Long term contractFull timeRemote work
$30 per hour
...what’s best for our customers. Cohere is a team of researchers, engineers, designers, and more, who are passionate about their craft. Each... .... 12 month contract. Performance incentives included! As a Data Annotation Specialist, you will: Evaluate the model's ability...Remote jobHourly pay16 hoursPermanent employmentFull timeContract workTemporary workPart timeFor contractorsWork at officeFlexible hours$37 per hour
...Description Mindrift is looking for highly skilled Python Data Scraping Engineers to join the Tendem project and drive specialized data scraping workflows within our hybrid AI + human system. In this role, as an AI Pilot – that’s how we refer to this role at Mindrift...Hourly payFull timePart timeFreelanceRemote work
Do you want to receive more vacancies?
Subscribe and receive similar vacancies to Data Engineer RD-017. Be the first to apply!
