Full Stack Data Science Engineer

Salute Inc.
United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Compensation
$120,000.0 - $150,000.0
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Business Analytics Applications Apache HTTP Server Audit Trail Big Data BigQuery Computerized Maintenance Management Systems Software Documentation Data Centers Information Engineering Data Governance
+32 more
Extract Transform Load (ETL) Dataspaces Data Virtualization Data Center Infrastructure Management (CIM) Event Logging Supervisory Control and Data Acquisition (SCADA) Python (Programming Language) NumPy Operational Data Store Power BI Tensorflow Standard Sql Tableau (Software) Datadog Data Processing Building Management System (BMS) Feature Engineering Data Ingestion Pytorch ReactJS Snowflake Apache Spark Pandas Data Lakes Scikit Learn Information Technology Data Lineage Apache Kafka Operational Systems Machine Learning Operations Presto Databricks

Job description

  • Design and operate sovereign data lake and warehouse architectures: schema design, data contracts, lineage tracking, freshness SLAs, governance frameworks, and multi-source ingestion pipelines.
  • Build end-to-end data science systems: feature engineering, model training and evaluation, inference pipeline deployment, monitoring, and feedback loops.
  • Develop novel statistical models, predictive algorithms, and optimization frameworks applied to operational data - labor efficiency, asset performance, energy consumption, and SLA adherence.
  • Identify patentable innovations in data processing architectures, model designs, and analytical methods; author invention disclosures and support patent prosecution alongside legal counsel.
  • Create production-grade analytical products: executive KPI dashboards, real-time operational intelligence layers, forecasting tools, and anomaly detection systems.
  • Translate raw operational data from physical infrastructure (sensors, CMMS, BMS, field logs) into structured, queryable, and model-ready data products.
  • Establish data engineering best practices: dbt transformations, data quality tests, observability tooling, and documentation standards across the data platform.
  • Collaborate with AI/ML engineers on feature stores, embeddings, and model-ready data products; partner with software engineers to integrate analytical outputs into user-facing applications.
  • Drive data governance: access controls, PII handling, audit trails, and compliance with data sovereignty requirements.
  • Architect and execute the migration from fragmented, siloed operational data systems to a decentralized federated data model - implementing federated query engines (e.g., Trino, Spark, or equivalent) and data virtualization layers that enable cross-domain analytics without requiring full data centralization; design domain-oriented data products aligned with data mesh principles, preserving source-system ownership while enabling platform-wide discoverability and governed access.

Requirements

Do you have experience in Technical writing within technology?, * 8+ years combining data science and data engineering in production environments, including at least 3 years at a senior IC level.

  • Deep expertise in the Python data ecosystem: pandas, NumPy, scikit-learn, PyTorch or TensorFlow, and statistical modeling libraries.
  • Proven experience designing and operating large-scale data warehouses or data lakes (Snowflake, BigQuery, Databricks, or equivalent).
  • Strong SQL and transformation tooling (dbt, Spark, or similar); experience with streaming data pipelines (Kafka, Kinesis, or equivalent).
  • Full-stack capability: able to take a data product from raw source through pipeline, model, API, and user-facing interface without hand-offs.
  • Experience with intellectual property in the data or software domain: invention disclosures, prior art research, or patent application involvement.
  • Strong technical writing skills - able to articulate novel methodologies clearly for patent disclosures and analytical documentation, * Named inventor on data science, ML, or software patents.
  • Domain experience in operations-heavy industries: data center management, facilities, manufacturing, energy, logistics, or industrial IoT.
  • Experience building client-facing BI or analytics products; familiarity with tools such as Power BI, Tableau, or custom React-based dashboard frameworks.
  • Knowledge of CMMS systems, BMS data formats, SCADA historian data, or similar operational technology data sources. Familiarity with xAPI / LRS standards for operational event logging and workforce analytics.
  • Experience with MLOps platforms (MLflow, Weights & Biases, SageMaker) and model governance practices.
  • Hands-on experience with federated query engines (Trino, Presto, DuckDB), data virtualization platforms, or data mesh implementations that unify access across siloed operational systems without full data movement.
  • Familiarity with open table formats (Apache Iceberg, Delta Lake, Apache Hudi) as the foundation for interoperable, decentralized data lakes.
  • Graduate degree in Data Science, Statistics, Computer Science, Applied Mathematics, or related field.

If you are a motivated and results-driven individual with a passion for data center services and a knack for building strong client relationships, we want to hear from you. Join us in revolutionizing the data center industry and apply today!

Benefits & conditions

3.53.5 out of 5 stars United States Remote $120,000 - $150,000 a year - Full-time, Pulled from the full job description

  • AD&D insurance
  • Parental leave
  • Health insurance
  • 401(k) matching
  • Paid time off
  • Vision insurance
  • Health savings account, The base salary range for this role can go up to $150,000/per year. This reflects a good-faith estimate of what we reasonably expect to pay upon hire, based on factors such as skills, experience, education, and market/location. Our comprehensive benefits include health, dental, and vision insurance, Health Savings Account (HSA), gym discount, mental health, Discounted Group Life & AD&D, Discounted Group Short & Long-term Disability, 401(k) retirement matching, PTO/paid holidays, and parental leave. Final compensation will be determined by job-related factors consistent with applicable law.

About the company

Salute is a leading provider of cutting-edge Data Center Infrastructure Services, dedicated to serving data center clients worldwide. We pride ourselves on delivering sustainable solutions, unparalleled reliability, and outstanding customer service. As we continue to grow, we are seeking a dynamic and experienced Full Stack Data Science Engineer to join our team and drive our relationships with hyperscale clients to new heights.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · World Congress 2021

Videos

See all

Related articles

See all