Fabric Data Engineer
Job Description
arrivia builds technology that powers travel loyalty and rewards for some of the world’s leading brands, supporting millions of travelers with booking and experience platforms. In this hybrid role based in Scottsdale, AZ, the team is looking for an experienced Fabric Data Engineer to take ownership of complex data pipeline architectures within the Microsoft Fabric ecosystem and help shape how the platform evolves.
You will lead modernization efforts by migrating on-prem relational data into OneLake while retiring legacy data-warehouse systems. The work spans batch and streaming pipelines, governance, performance tuning, and AI-enabled data experiences that support Retrieval-Augmented Generation patterns.
Key Responsibilities
- Architect and optimize end-to-end pipelines using Microsoft Fabric Data Factory, Dataflows Gen2, and PySpark and Spark SQL notebooks for scale and performance.
- Design Lakehouse and Warehouse structures using the medallion approach (Bronze, Silver, Gold) and define best practices for the team.
- Support the move of on-premise relational data into OneLake and help decommission legacy data-warehouse systems to accelerate the transition to the cloud.
- Build low-latency streaming pipelines with Fabric Eventstream from sources including Azure Event Hubs, IoT Hub, and custom applications.
- Write highly optimized T-SQL, Spark SQL, and PySpark, owning incremental loads, refresh scheduling, and SLA monitoring.
- Enable pipelines for Retrieval-Augmented Generation by implementing chunking, embeddings, and vector search, and using LLMs and Model Context Protocol (MCP) servers alongside AI-assisted tooling.
- Champion governed data access using sensitivity labels, role-based access controls, and cataloging with Microsoft Purview.
- Drive CI/CD using Fabric deployment pipelines, Git branching strategies, and automated testing for data assets.
- Maintain Power BI semantic models where required to support consistent, accurate enterprise reporting.
- Coach Fabric Data Engineer I team members through code reviews, pair programming, knowledge sharing, and architecture reviews.
- Design and optimize enterprise-scale data solutions, set technical standards, mentor junior engineers, and lead continuous improvement.
- Partner with analysts, data scientists, and business stakeholders to translate complex requirements into reliable, performant, well-governed data platforms.
Required Qualifications
- 3 to 5 years in data engineering, ETL and ELT development, or a related analytics engineering role.
- Strong SQL across T-SQL and Spark SQL, and strong Python with PySpark.
- Several years working with Apache Spark and lakehouse data at scale, including performance tuning and notebooks built for large datasets.
- Solid grounding in lakehouse architecture, data warehousing, dimensional modeling, and data vault methodology.
- Hands-on experience with Microsoft Fabric or a comparable cloud data platform such as Azure Synapse or Databricks.
- Experience with real-time and streaming data at scale, including event-driven architectures and tools like Azure Event Hubs or Kafka.
- Familiarity with vector search, embeddings, and Retrieval-Augmented Generation patterns, plus hands-on use of LLMs and AI-assisted development tools.
- Strong CI/CD habits across Fabric deployment pipelines, Git, and automated testing.
- Demonstrated experience mentoring junior engineers and leading technical initiatives.
- Microsoft Certified: Fabric Data Engineer Associate (DP-700) is highly preferred.
- A bachelor’s degree in a related field, or equivalent practical experience.
Technologies
- Microsoft Fabric, Microsoft Fabric Data Factory, Dataflows Gen2, PySpark, Spark SQL, T-SQL
- OneLake, Fabric Eventstream
- Azure Event Hubs, IoT Hub, Apache Spark, Azure Synapse, Databricks, Kafka
- Retrieval-Augmented Generation, LLMs, Model Context Protocol (MCP) servers, vector search, embeddings
- Microsoft Purview, Power BI, Power BI semantic models
- CI/CD, Fabric deployment pipelines, Git, automated testing
- Sensitivity labels, role-based access controls
Benefits
- Unlimited PTO
- Exclusive employee travel rates
- Travel discounts through arrivia programs
- Medical, dental, and vision insurance
- 401(k) with company participation
About the Role: arrivia is looking for an experienced Fabric Data Engineer to take ownership of complex data pipeline architectures in the Microsoft Fabric ecosystem. You will design and optimize enterprise-scale data solutions, set the technical standards the team works to, and mentor junior engineers along the way, while migrating an existing on-premise relational database into the modern cloud environment and decommissioning legacy systems.
Role Logistics: Scottsdale, AZ (hybrid). Minimum experience: 3 years.