Data Scientist

VACO LLC
Cranberry Township, PA, United States
2 days ago
Apply on jobs.vaco.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source

Tech stack

Sql Data Warehouse Application Programming Interfaces (APIs) Artificial Intelligence Business Analytics Applications Data Analysis Microsoft Azure Big Data Software as a Service Information Systems Information Engineering Extract Transform Load (ETL) Data Warehousing
+26 more
JSON Python (Programming Language) Log Files Machine Learning Microsoft SQL Server SQL Azure Oracle (Applications) Power BI Tensorflow Sentiment Analysis SQL Databases Data Streaming Speech Recognition Extensible Markup Language (XML) Enterprise Data Management Data Processing Azure Data Factory Apache Spark Microsoft Fabric AI Platforms Pyspark Information Technology Data Analytics Apache Kafka Restful APIs Stream Analytics

Job description

Vaco by Highspring and its parents, affiliates, and subsidiaries (“we,” “our,” or “Vaco by Highspring”) respects your privacy and are committed to providing transparent notice of our policies.

  • California residents may access Vaco by Highspring HR Notice at Collection for California Applicants and Employees here.
  • Virginia residents may access our state specific policies here.
  • Residents of all other states may access our policies here.
  • Canadian residents may access our policies in English here and in French here.
  • Residents of countries governed by GDPR may access our policies here.

Requirements

Work Authorization Candidates must currently reside in the Pittsburgh area and be available for a hybrid schedule. This opportunity is for direct hire only and is not available for C2C, third-party submissions, or visa sponsorship Data Scientist We are seeking a versatile Data Scientist to join our team in Pittsburgh, PA. This is a direct hire position offering a hybrid work schedule for candidates who currently reside in the Pittsburgh area. This role is not open to C2C, visa sponsorship, or third parties. Role Overview In this role, you will design, build, and optimize data pipelines while also delivering advanced analytics, predictive modeling, and business insights. You will work across modern cloud platforms and leverage Microsoft Fabric to unify data engineering, analytics, and business intelligence efforts. Microsoft Fabric’s integrated ecosystem supports OneLake, Lakehouse, Data Warehouse, Data Factory, Dataflows Gen2, Eventstream, Spark notebooks, and Power BI, making it well suited for this type of end-to-end analytics work. Key Responsibilities Design, develop, and maintain interactive dashboards and analytical solutions in Power BI, including Direct Lake semantic models, Copilot-enabled experiences, paginated reports, and executive scorecards. Power BI Copilot provides chat-based analysis and support for tasks such as DAX generation and on-the-fly analysis. Build KPI-driven dashboards for pricing analytics, POS sales trends, inventory optimization, supply chain performance, customer purchasing behavior, and branch/DC operational analytics. Enable self-service analytics through governed semantic models, curated datasets, and reusable reporting assets. Direct Lake semantic models are optimized for large data volumes in Fabric. Develop predictive and prescriptive models for dynamic pricing, sales forecasting, demand planning, customer segmentation, lost sales analysis, and inventory replenishment. Apply machine learning and AI techniques using Python, Spark notebooks, Fabric Data Science workloads, and Azure AI / OpenAI capabilities. Fabric supports Azure OpenAI and related AI services through integrated tooling. Support AI-driven initiatives such as Copilot-enabled analytics, conversational BI, sentiment analysis, speech-to-text analytics, and agentic AI use cases. Work extensively within the Microsoft Fabric ecosystem, including OneLake, Lakehouse, Data Warehouse, Data Factory, Dataflows Gen2, Eventstream / Real-Time Intelligence, Spark notebooks, semantic models, and Power BI. Support ingestion and transformation of enterprise data from Oracle, SQL Server, APIs, SaaS platforms, XML/JSON log files, and event streaming platforms. Partner with Pricing, Sales, Supply Chain, Operations, and IT teams to translate business needs into technical analytics solutions and present findings to leadership. Contribute to enterprise data modernization and cloud transformation initiatives. Required Qualifications Bachelor’s or Master’s degree in Data Science, Computer Science, Information Systems, Engineering, Statistics, Mathematics, or a related field. 3-5 years of experience in Data Analytics, Business Intelligence, Data Science, or a related analytics role. Strong experience with Power BI, SQL, data modeling, and dashboard development. Experience with Microsoft Azure and Microsoft Fabric. Proficiency in Python for analytics and data manipulation. Experience working with large enterprise datasets. Preferred Qualifications Experience with Microsoft Fabric services such as Lakehouse, Spark, Data Factory, Real-Time Analytics, and semantic models. Familiarity with machine learning frameworks and core AI/ML concepts. Knowledge of pricing analytics, supply chain analytics, inventory optimization, and wholesale/distribution analytics. Experience with REST APIs, event streaming, Kafka/Event Hub, and XML/JSON data processing. Familiarity with cloud data warehouse modernization initiatives. Microsoft certifications in Azure, Power BI, Fabric, or Data Engineering are a plus. Technical Skills Required Preferred Required Preferred Power BI Microsoft Fabric SQL Azure Synapse Python Spark / PySpark Data Modeling Dataflows Gen2 DAX OneLake ETL/ELT Concepts Azure Data Factory Dashboard Development Azure Event Hub REST APIs AI/ML Tooling

Benefits & conditions

Determining compensation for this role (and others) at Vaco by Highspring depends upon a wide array of factors including but not limited to:

  • the individual’s skill sets, experience and training;
  • licensure and certification requirements;
  • office location and other geographic considerations;
  • other business and organizational needs.

With that said, as required by local law, Vaco by Highspring believes that the following salary range referenced above reasonably estimates the base compensation for an individual hired into this position in geographies that require salary range disclosure. The individual may also be eligible for discretionary bonuses.

About the company

Additionally, submissions to this position are subject to the use of AI to perform preliminary candidate screenings, focused on ensuring minimum job requirements noted in the position are satisfied. More details about Vaco by Highspring’s use of AI can be found here (https://www.highspring.com/ai-use-notices/). Further assessment of candidates beyond this initial phase will be conducted by recruiters and hiring managers. Vaco by Highspring does not know and cannot opine on if its client’s use of AI products in hiring.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on jobs.vaco.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

1:46 min

Traditional data architecture before Microsoft Fabric

Dr. Alexander Wachtel Dr. Alexander Wachtel +1 · World Congress 2025

2:03 min

Distinguishing type definition constructs from data validation routines

Clemens Vasters Clemens Vasters · World Congress 2025

Videos

See all

Related articles

See all