Lead / Sr Analyst Data Scientist

Boardwalk Pipelines
Houston, TX, United States
4 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
3 years minimum
Working hours
Regular working hours

Tech stack

Geographic Information Systems Amazon Web Services Amazon S3 Data Analysis Big Data Cloud Computing Cloud Database Continuous Delivery Continuous Integration Data Governance Data Security Data Visualization
+20 more
Monitoring of Systems Supervisory Control and Data Acquisition (SCADA) Python (Programming Language) Machine Learning Power BI Tensorflow SQL Databases Systems Integration Tableau (Software) Cloud Platform System Feature Engineering Pytorch Git Pandas Scikit Learn Information Technology Real Time Data Machine Learning Operations Software Version Control Databricks

Job description

The Data Scientist plays a key role in advancing digital innovation by developing data-driven models and insights that support operational efficiency, asset optimization, and strategic decision-making. This role will work closely with data engineers, business stakeholders, and digital leadership to design and deploy machine learning (ML) models, predictive analytics, and advanced visualizations that drive measurable business outcomes., Model Development & Advanced Analytics

  • Design, build, and deploy predictive models and machine learning (ML) algorithms to support asset performance, reliability, and commercial optimization.
  • Conduct exploratory data analysis (EDA), feature engineering, and statistical modeling using Python, R, or similar tools.
  • Develop time series forecasting, anomaly detection, and classification models for operational and business use cases.
  • Apply geospatial and sensor data analytics to support pipeline monitoring, flow optimization, and risk assessment.

Cloud & Platform Integration

  • Leverage Databricks and Amazon Web Services (AWS) tools such as SageMaker, Redshift, Simple Storage Service (S3), Lambda, and Glue for model training, deployment, and data access.
  • Collaborate with data engineers to ensure models are integrated into production pipelines and dashboards.
  • Use version control (e.g., Git), MLFlow and continuous integration/continuous deployment (CI/CD) practices to manage model lifecycle and reproducibility.

Business Collaboration & Impact

  • Partner with operations, engineering, and commercial teams to identify high-impact use cases and translate business needs into analytical solutions.
  • Present findings and recommendations through compelling data visualizations and storytelling using tools like Power BI.
  • Support the development of self-service analytics and promote data literacy across the organization.
  • Document model assumptions, limitations, and performance metrics to ensure transparency and trust.

Model Governance & Continuous Improvement

  • Monitor model performance and retrain as needed to maintain accuracy and relevance.
  • Implement MLOps practices to support scalable, automated model deployment and monitoring.
  • Ensure compliance with data governance, privacy, and ethical AI standards.
  • Stay current with industry trends and emerging technologies to continuously improve analytical capabilities., Job Title Data Scientist Education Bachelor’s Degree Location Houston, TX 77032 US (Primary) Career Level Professional Category Operations Job Type Permanent …
  • 14 days ago

Requirements

The ideal candidate combines strong analytical skills with deep technical expertise in cloud-based data science tools, particularly within the Databricks and Amazon Web Services (AWS) ecosystem. This role requires a passion for solving complex problems, a collaborative mindset, and the ability to translate data into actionable insights for a midstream energy environment., * Bachelor’s or Master’s degree in Data Science, Computer Science, Engineering, Statistics, or a related field.

  • 3-5+ years of experience in applied data science, preferably in the energy, utilities, or industrial sectors.
  • Proficiency in Python, SQL, and data science libraries such as pandas, scikit-learn, TensorFlow, or PyTorch.
  • Experience working with Databricks and AWS services.
  • Strong understanding of statistical modeling, machine learning, and data visualization techniques.
  • Ability to communicate complex technical concepts to non-technical stakeholders.
  • Experience working with large datasets and real-time or near-real-time data environments.

PREFERRED SKILLS, KNOWLEDGE, AND EXPERIENCE:

  • Experience in the natural gas midstream or broader oil & gas industry.
  • Familiarity with MLOps practices and tools for model monitoring and retraining.
  • Exposure to geospatial data, sensor data, or SCADA systems.
  • Experience with Power BI, Tableau, or similar BI tools.
  • Knowledge of data governance, security, and compliance in cloud environments.

REQUIRED EDUCATION:

  • Bachelor’s degree in Computer Science, Data Science, Engineering, or related field

PREFERRED EDUCATION:

  • Masters Degree

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

Videos

See all

Related articles

See all