EngineerJobs.io
← Back to all jobs

Job Description

Arrivia, Inc. is looking for a Fabric Data Engineer to help shape complex, enterprise-grade data pipeline architectures within the Microsoft Fabric ecosystem. This hybrid role based in Scottsdale, AZ will support migration away from legacy systems by designing governed lakehouse and warehouse solutions, building reliable streaming ingestion, and enabling analytics teams with consistent reporting platforms.

What you will do

  • Architect and optimize end-to-end pipelines using Microsoft Fabric Data Factory, Dataflows Gen2, and PySpark and Spark SQL notebooks for scale and performance.
  • Lead Lakehouse and Warehouse design using the medallion pattern (Bronze, Silver, Gold) and establish the best practices the team follows.
  • Move on-premises relational data into OneLake and help retire legacy data-warehouse systems to accelerate the transition to a modern cloud platform.
  • Build low-latency streaming pipelines with Fabric Eventstream, integrating sources such as Azure Event Hubs, IoT Hub, and custom applications.
  • Write and optimize T-SQL, Spark SQL, and PySpark, owning incremental loads, refresh scheduling, and SLA monitoring.
  • Design data pipelines that support Retrieval-Augmented Generation (RAG), including chunking, embeddings, and vector search, and leverage LLMs and Model Context Protocol (MCP) servers to improve team workflows.
  • Support governance with Microsoft Purview, including sensitivity labels, role-based access controls, and cataloging.
  • Drive CI/CD using Fabric deployment pipelines, Git branching strategies, and automated testing for data assets.
  • Maintain Power BI semantic models when needed to keep enterprise reporting consistent and accurate.
  • Coach Fabric Data Engineer I team members through code reviews, pair programming, and knowledge sharing, and contribute to architectural reviews and continuous improvement.

Qualifications

  • 3 to 5 years of experience in data engineering, ETL/ELT development, or a related analytics engineering role.
  • Strong SQL skills across T-SQL and Spark SQL, plus strong Python development with PySpark.
  • Working knowledge of Scala is a plus.
  • Several years building lakehouse-scale solutions with Apache Spark, including performance tuning with PySpark and Spark SQL notebooks on large datasets.
  • Solid grounding in lakehouse architecture, data warehousing, dimensional modeling, and data vault methodology.
  • Hands-on experience with Microsoft Fabric or a comparable platform such as Azure Synapse or Databricks.
  • Experience with real-time and streaming data at scale, including event-driven architectures and tools such as Azure Event Hubs or Kafka.
  • Familiarity with vector search, embeddings, and RAG patterns, including hands-on use of LLMs and AI-assisted development tools.
  • Strong CI/CD practices using Fabric deployment pipelines, Git, and automated testing.
  • Proven experience mentoring junior engineers and leading technical initiatives.
  • Microsoft Certified: Fabric Data Engineer Associate (DP-700) is highly preferred.
  • A bachelor’s degree in a related field (or equivalent practical experience).

Benefits

  • Unlimited PTO
  • Exclusive employee travel rates
  • Travel discounts through arrivia programs
  • Medical, dental, and vision insurance
  • 401(k) with company participation

Similar Jobs