Data Scientist

FRACTAL LLC
United States
1 day ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Compensation
$140,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Data Analysis Microsoft Azure Information Engineering UN Electronic Data Interchange for Administration Commerce and Transport Python (Programming Language) Machine Learning NumPy Rapid Prototyping Process Recommender Systems Tensorflow Standard Sql
+20 more
Azure Machine Learning Azure Data Lake Software Deployment Cloud Platform System Feature Engineering Azure Data Factory GitHub Copilot Office365 Large Language Models Prompt Engineering Deep Learning Generative AI Pandas Containerization Data Lakes Scikit Learn Information Technology Machine Learning Operations Software Version Control Databricks

Job description

Team is seeking a talented and mission driven Data Scientist to design, develop, and deploy AI and machine learning solutions that transform the Healthcare Revenue Cycle Management (RCM) process. In this role, you will partner closely with Operations, Product, Data Engineering, and Engineering teams to build predictive models, generative AI solutions, and intelligent automation that improve efficiency across claims processing, denials and appeals, clinical documentation, coding, customer service, and related workflows.

You will work hands on with cloud-native technologies-primarily in Azure Machine Learning and Databricks-to bring models from concept to production. This position is ideal for someone who is self-motivated and self-driven, thinks innovatively about how to unlock value from RCM data, and consistently presents solutions rather than being blocked by challenges. You should be comfortable operating in a fast-paced environment with rapid prototyping and iterative delivery, and be willing to research and find answers independently rather than expecting step-by-step guidance.

Responsibilities

  • Develop machine learning, deep learning, and generative AI models to support RCM use cases such as claim outcome prediction, denials classification, next best action recommendation, clinical appeal summarization, and workflow optimization.

  • Build end to end ML pipelines including feature engineering, model training, hyperparameter tuning, validation, and monitoring.

  • Research and apply advanced techniques in LLMs, embeddings, retrieval augmentation (RAG), prompt engineering, and document intelligence.

  • Translate operational and product requirements into measurable model objectives, data specifications, and evaluation frameworks.

  • Conduct exploratory data analysis (EDA) to understand data patterns, anomalies, and business insights.

  • Implement and follow best practices around MLOps, including model versioning, reproducibility, feature stores, and drift monitoring.

  • Partner with Data Engineers to ensure feature availability, data quality, and scalable ML/AI deployment within Azure Databricks and enterprise platforms.

  • Collaborate with cross-functional teams to communicate insights, present model results, and drive adoption of ML solutions.

  • Maintain awareness of emerging AI technologies and propose enhancements or new opportunities for Provider Engineering platforms.

Requirements

  • Bachelor’s or Master’s degree in Data Science, Computer Science, Statistics, Applied Mathematics, or a related quantitative field.

  • 8+ years of hands-on experience developing ML or AI models in an applied industry setting (healthcare experience strongly preferred).* Strong proficiency in:

Databricks

Python (pandas, scikit-learn, NumPy, TensorFlow, MLflow)

SQL

  • Preferred: Knowledge of US healthcare, ideally with Revenue Cycle Management (RCM) experience, including familiarity with EDI transactions such as 837 and 835.

  • Ability to thrive in a fast-paced environment with rapid prototyping, ambiguity, and iterative delivery.

  • Required: Strong working knowledge of GitHub Copilot (or equivalent enterprise-approved AI coding assistant) to accelerate development.

  • Self-motivated and solution-oriented; able to independently research, troubleshoot, and identify paths forward when facing technical or data challenges.

  • Experience working in Azure or similar cloud environments:

Azure Databricks (MLflow, Delta Lake)

Azure Machine Learning

Azure Data Lake / Azure Data Factory (in partnership with Data Engineering)

  • Hands-on experience in at least one of the following:

Classification, regression, time series, or recommendation systems

Agentic AI / Large Language Models (LLMs) / RAG

  • Strong statistical modeling foundation and experience building production-quality ML systems.* Ability to communicate technical findings to nontechnical stakeholders clearly and effectively.

Required AI Skills:

All contractor resources are expected to demonstrate baseline proficiency in enterprise-approved AI tools as part of their day-to-day responsibilities. This includes, but is not limited to:

Consistent Use: Maintain a minimum of 90% weekly usage of AI tools such as GitHub Copilot, Microsoft 365 Copilot, and other GenAI platforms approved by the enterprise.

Applied Productivity: Leverage AI tools to enhance coding, documentation, data analysis, and decision-making workflows.

Benefits & conditions

The wage range for this role takes into account the wide range of factors that are considered in making compensation decisions including but not limited to skill sets; experience and training; licensure and certifications; and other business and organizational needs. The disclosed range estimate has not been adjusted for the applicable geographic differential associated with the location at which the position may be filled. At Fractal, it is not typical for an individual to be hired at or near the top of the range for their role and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current range is: $140,000 - $150,000. In addition, you may be eligible for a discretionary bonus for the current performance period., As a full-time employee of the company or as an hourly employee working more than 30 hours per week, you will be eligible to participate in the health, dental, vision, life insurance, and disability plans in accordance with the plan documents, which may be amended from time to time. You will be eligible for benefits on the first day of employment with the Company. In addition, you are eligible to participate in the Company 401(k) Plan after 30 days of employment, in accordance with the applicable plan terms. The Company provides for 11 paid holidays and 12 weeks of Parental Leave. We also follow a “free time” PTO policy, allowing you the flexibility to take time needed for either sick time or vacation.

About the company

Fractal Analytics is a strategic AI partner to Fortune 500 companies with a vision to power every human decision in the enterprise. Fractal is building a world where individual choices, freedom, and diversity are the greatest assets. An ecosystem where human imagination is at the heart of every decision. Where no possibility is written off, only challenged to get better. We believe that a true Fractalite empowers imagination with intelligence. And that it will be such Fractalites that will continue to build the company for the next 100 years.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · World Congress 2025

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · World Congress 2021

Videos

See all

Related articles

See all