Fabric Principal Data Engineer / Solution Architect

CAPITAL VIEW IV ASSOCIATION, INC.
New York, NY, United States
9 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure System Configuration Data Architecture Information Engineering Data Infrastructure Data Systems Data Warehousing Digital Assets Identity and Access Management Python (Programming Language) Machine Learning
+20 more
Microsoft SQL Server Performance Tuning Cloud Services Tensorflow Azure Data Lake SQL Stored Procedures SQL Databases Systems Integration Data Processing Azure Data Factory Pytorch Indexer Microsoft Fabric Pyspark Kubernetes Data Management Machine Learning Operations Cloud Integration Data Pipelines Databricks

Job description

This is a full-time position for a Principal Data Engineer / Solution Architect specializing in Microsoft Fabric and Azure data architectures to support a fast-moving project for a key client in the Life Sciences / Biotech industry. In this role, you will serve as a senior technical resource responsible for supporting, enhancing, and executing data engineering solutions across Microsoft Fabric, Fabric Data Factory, Microsoft SQL, ADLS Gen2, Databricks Unity Catalog, PySpark in Fabric, and related Azure ecosystem services. The project operates at high velocity, requiring someone who can ramp up quickly, navigate technical ambiguity, and keep active delivery moving seamlessly. You will help stabilize, troubleshoot, and advance an active migration from traditional Azure infrastructure to Microsoft Fabric. Location & Logistics: Candidates must be based in Boston, MA or fully capable of traveling to Boston as required by project milestones, maintaining 100% working time overlap with US East Coast business hours. Qualification / Skill Set Requirement:

  • Core Technical Expertise, * Fabric & Pipeline Engineering: Design, build, enhance, and support scalable data pipelines and integrations using Microsoft Fabric Data Factory, PySpark in Fabric, and Azure Data Factory.
  • Migration Support & Stabilization: Drive active migration efforts to Microsoft Fabric, including pipeline transition, validation, performance tuning, and operational stabilization.
  • Databricks Governance: Manage and organize Databricks Unity Catalog structures, permissions, and access control policies.
  • Data Asset Management: Develop and optimize Microsoft SQL data assets, complex stored procedures, queries, and ADLS Gen2 storage layers.
  • Troubleshooting & Quality Assurance: Quickly diagnose and resolve issues across data pipelines, SQL queries, Fabric PySpark notebooks, Unity Catalog, and cloud integration layers.
  • Technical Leadership & Knowledge Transfer: Assist with knowledge transfer from outgoing project resources, document technical designs and operational workflows, and mentor delivery team members.
  • Stakeholder Collaboration: Partner closely with technical, project, and business stakeholders to translate complex requirements into reliable, compliant data solutions during US Eastern business hours., Lead Machine Learning Engineer (MLOps, KServe + building Kubernetes Clusters, PyTorch, TensorFlow on AWS) As a Capital One Machine Learning Engineer (MLE), you’ll be part of an Agi…
  • 1 day ago

Requirements

  • 8+ years of hands-on data engineering, data architecture, or complex data platform development experience.
  • Strong, practical experience building, supporting, and troubleshooting data pipelines in Microsoft Fabric and Azure Data Factory (ADF).
  • Advanced proficiency in Microsoft SQL (query development, stored procedures, schema design, indexing, performance tuning) and Python / PySpark in Fabric.
  • Hands-on experience managing and supporting Databricks Unity Catalog (catalog/schema/table organization, access controls, permissions, and governance administration).
  • Deep familiarity with Azure Data Lake Storage Gen2 (ADLS Gen2), including data organization, security, and cloud service integrations.
  • Architecture & Migration Experience:

  • Proven experience supporting or leading platform migrations, modernizations, or transitions (specifically Azure-to-Microsoft Fabric).
  • Familiarity with medallion architecture, lakehouse patterns, and modern data warehouse design.
  • Experience with Azure administration activities supporting data platforms (access management, service configuration, and monitoring).
  • Industry & Domain Experience:

  • Required: Prior hands-on experience working with Life Sciences, pharmaceutical, biotechnology, medical device, clinical, or regulatory data environments.
  • Knowledge of appropriate data handling, quality, compliance awareness, and business context in regulated settings.
  • Logistics & Operating Style:

  • Based in Boston, MA or willing/able to travel to Boston, MA for client engagements and alignment meetings.
  • Full working time overlap with US East Coast business hours (ET).
  • Demonstrated ability to operate effectively in a startup-speed environment with shifting priorities, tight timelines, and minimal ramp-up time.
  • Certifications Preferred

About the company

Fusemachines is a 12+ year old AI company, dedicated to delivering state-of-the-art AI products and solutions to a diverse range of industries. Founded by Sameer Maskey, Ph.D., an Adjunct Associate Professor at Columbia University, our company is on a steadfast mission to democratize AI and harness the power of global AI talent from underserved communities. With a robust presence in four countries and a dedicated team of over 400 full-time employees, we are committed to fostering AI transformation journeys for businesses worldwide. At Fusemachines, we not only bridge the gap between AI advancement and its global impact but also strive to deliver the most advanced technology solutions to the world.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

2:36 min

Analyzing limitations with PostgreSQL bitmap heap scans

Dharin Shah Dharin Shah · World Congress 2025

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

1:10 min

Introduction to Microsoft Fabric and data agents

Dr. Alexander Wachtel Dr. Alexander Wachtel +1 · World Congress 2025

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

2:50 min

Exploring the Microsoft Fabric workspace foundation

Alexandra Mihai Alexandra Mihai +2 · Europe 2026 Virtual

Videos

See all

Related articles

See all