Fabric Data Engineer
Job Description
Arrivia, Inc. is seeking an experienced Fabric Data Engineer to own complex data pipeline architecture within the Microsoft Fabric ecosystem. This hybrid role in Scottsdale, AZ focuses on modernizing data platforms, building governed lakehouse and warehouse pipelines, and enabling AI-ready retrieval workflows.
Responsibilities
- Architect and optimize end-to-end pipelines using Microsoft Fabric Data Factory, Dataflows Gen2, and PySpark and Spark SQL notebooks for scale and performance.
- Lead Lakehouse and Warehouse design using the medallion pattern (Bronze, Silver, Gold) and establish best practices for the team.
- Support migration of on-premise relational data into OneLake while helping retire legacy data-warehouse systems.
- Build low-latency streaming pipelines with Fabric Eventstream from sources such as Azure Event Hubs, IoT Hub, and custom applications.
- Write and optimize T-SQL, Spark SQL, and PySpark, owning incremental loads, refresh scheduling, and SLA monitoring.
- Design pipelines that support Retrieval-Augmented Generation workflows, including chunking, embeddings, and vector search, and collaborate using LLMs and Model Context Protocol (MCP) servers to improve team productivity.
- Champion governance via sensitivity labels, role-based access controls, and cataloging in Microsoft Purview.
- Drive CI/CD using Fabric deployment pipelines, Git branching strategies, and automated testing for data assets.
- Maintain Power BI semantic models where required to keep enterprise reporting consistent and accurate.
- Coach Fabric Data Engineer I team members through code reviews, pair programming, and knowledge sharing, and support architectural reviews and continuous improvement.
- Design and optimize enterprise-scale data solutions, set technical standards, and mentor junior engineers.
- Partner with analysts, data scientists, and business stakeholders to translate complex requirements into reliable, performant, and well-governed data platforms.
Requirements
- 3 to 5 years of data engineering, ETL/ELT development, or a related analytics engineering role.
- Strong SQL across T-SQL and Spark SQL, and strong Python with PySpark.
- Working knowledge of Scala is a plus.
- Several years working with Apache Spark and lakehouse-scale data, including performance-tuning experience and notebook development with PySpark and Spark SQL over large datasets.
- Solid grounding in lakehouse architecture, data warehousing, dimensional modeling, and data vault methodology.
- Hands-on experience with Microsoft Fabric or a comparable cloud data platform such as Azure Synapse or Databricks.
- Experience with real-time and streaming data at scale, including event-driven architectures and tools such as Azure Event Hubs or Kafka.
- Familiarity with vector search, embeddings, and Retrieval-Augmented Generation patterns, plus hands-on use of LLMs and AI-assisted development tools.
- Strong CI/CD practices across Fabric deployment pipelines, Git, and automated testing.
- Demonstrated experience mentoring junior engineers and leading technical initiatives.
- Microsoft Certified: Fabric Data Engineer Associate (DP-700) is highly preferred.
- A bachelor’s degree in a related field, or equivalent practical experience.
Technology Stack
- Microsoft Fabric, Microsoft Fabric Data Factory, Dataflows Gen2
- PySpark, Spark SQL, T-SQL, OneLake
- Fabric Eventstream, Azure Event Hubs, IoT Hub
- Apache Spark, Azure Synapse, Databricks, Kafka
- Retrieval-Augmented Generation, LLMs, Model Context Protocol (MCP) servers, vector search, embeddings
- Microsoft Purview, Power BI, Power BI semantic models
- CI/CD, Fabric deployment pipelines, Git, automated testing
- Sensitivity labels, role-based access controls
Benefits
- Unlimited PTO
- Exclusive employee travel rates
- Travel discounts through arrivia programs
- Medical, dental, and vision insurance
- 401(k) with company participation
About the Role
Arrivia is looking for an experienced Fabric Data Engineer to take ownership of complex data pipeline architectures within the Microsoft Fabric ecosystem. The position includes designing and optimizing enterprise-scale data solutions, setting the technical standards the team follows, and mentoring junior engineers. You will help migrate an existing on-premise relational database into the modern cloud environment and support the decommissioning of legacy data-warehouse systems.
You will collaborate with analysts, data scientists, and business stakeholders to deliver reliable, performant, and well-governed data platforms. Arrivia is committed to AI-driven innovation, and you will work with LLMs, Model Context Protocol (MCP) servers, and AI-assisted tools to accelerate delivery and raise engineering quality. The role offers flexible and hybrid work and certification support, with a clear path to grow into Senior, Lead, and Principal roles.
Who We Are
- We power travel loyalty and rewards programs for some of the world’s leading brands.
- Our teams help millions of travelers book memorable experiences while delivering innovative technology and travel solutions for partners.
- We operate with a global workforce and a culture built on curiosity, ownership, authenticity, and collaboration.