Fabric Data Engineer
Backend Developer
Analytics
Azure
Azure Data Engineer
Azure Data Factory
Azure Data Platform
Azure Platform
Big Data
Business Analytics
Business Intelligence
Cloud
Cloud Data Engineering
Cloud Data Platform
Cloud Platform
Cloud Platforms
Cloud Platforms Cloud Platforms
Data
Data Analysis
Data Analytics
Data Analytics Tools
Data Architecture
Data Engineer
Data Engineering
Data Engineering Lead
Data Factory
Data Factory Azure
Data Integration
Data Pipeline
Data Platform
Data Processing
Data Science
Data Visualization
Data Warehouse
Database
Databases
Design
Digital Marketing
ETL
Hr Technology
Informatica
Information Technology (IT)
Integration
Microsoft
Microsoft Azure
Microsoft Fabric
Microsoft Office
Power BI
Power Platform
Reporting and Analytics
SQL
Visual Design
Job Description
Arrivia, Inc. is looking for a Fabric Data Engineer to help shape complex, enterprise-grade data pipeline architectures within the Microsoft Fabric ecosystem. This hybrid role based in Scottsdale, AZ will support migration away from legacy systems by designing governed lakehouse and warehouse solutions, building reliable streaming ingestion, and enabling analytics teams with consistent reporting platforms.
What you will do
- Architect and optimize end-to-end pipelines using Microsoft Fabric Data Factory, Dataflows Gen2, and PySpark and Spark SQL notebooks for scale and performance.
- Lead Lakehouse and Warehouse design using the medallion pattern (Bronze, Silver, Gold) and establish the best practices the team follows.
- Move on-premises relational data into OneLake and help retire legacy data-warehouse systems to accelerate the transition to a modern cloud platform.
- Build low-latency streaming pipelines with Fabric Eventstream, integrating sources such as Azure Event Hubs, IoT Hub, and custom applications.
- Write and optimize T-SQL, Spark SQL, and PySpark, owning incremental loads, refresh scheduling, and SLA monitoring.
- Design data pipelines that support Retrieval-Augmented Generation (RAG), including chunking, embeddings, and vector search, and leverage LLMs and Model Context Protocol (MCP) servers to improve team workflows.
- Support governance with Microsoft Purview, including sensitivity labels, role-based access controls, and cataloging.
- Drive CI/CD using Fabric deployment pipelines, Git branching strategies, and automated testing for data assets.
- Maintain Power BI semantic models when needed to keep enterprise reporting consistent and accurate.
- Coach Fabric Data Engineer I team members through code reviews, pair programming, and knowledge sharing, and contribute to architectural reviews and continuous improvement.
Qualifications
- 3 to 5 years of experience in data engineering, ETL/ELT development, or a related analytics engineering role.
- Strong SQL skills across T-SQL and Spark SQL, plus strong Python development with PySpark.
- Working knowledge of Scala is a plus.
- Several years building lakehouse-scale solutions with Apache Spark, including performance tuning with PySpark and Spark SQL notebooks on large datasets.
- Solid grounding in lakehouse architecture, data warehousing, dimensional modeling, and data vault methodology.
- Hands-on experience with Microsoft Fabric or a comparable platform such as Azure Synapse or Databricks.
- Experience with real-time and streaming data at scale, including event-driven architectures and tools such as Azure Event Hubs or Kafka.
- Familiarity with vector search, embeddings, and RAG patterns, including hands-on use of LLMs and AI-assisted development tools.
- Strong CI/CD practices using Fabric deployment pipelines, Git, and automated testing.
- Proven experience mentoring junior engineers and leading technical initiatives.
- Microsoft Certified: Fabric Data Engineer Associate (DP-700) is highly preferred.
- A bachelor’s degree in a related field (or equivalent practical experience).
Benefits
- Unlimited PTO
- Exclusive employee travel rates
- Travel discounts through arrivia programs
- Medical, dental, and vision insurance
- 401(k) with company participation