Data Engineer

National Oilwell Varco
Houston, TX, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Working hours
Regular working hours

Tech stack

Agile Methodology Amazon Web Services Data Analysis Confluence JIRA Automation of Tests Software Documentation Computer Engineering Continuous Integration Information Engineering Data Infrastructure Extract Transform Load (ETL)
+24 more
Data Warehousing Relational Databases Database Queries Distributed Computing Environment Supervisory Control and Data Acquisition (SCADA) Python (Programming Language) Machine Learning Microsoft SQL Server NoSQL Operational Data Store Reliability Engineering SQL Databases System Availability Delivery Pipeline Apache Spark Git Containerization Pyspark Information Technology Deployment Automation Data Pipelines Docker Databricks Programming Languages

Job description

In this role, you will design, develop, and support a robust data ecosystem that powers real-time insights for condition-based maintenance, drilling optimization, and operational efficiency. You will work closely with engineers, data scientists, and operational experts to translate business needs into technical solutions that enhance reliability and performance across NOV’s global operations. By joining our team, you’ll help transform how drilling and rig equipment data drives operational excellence. Your work will directly support predictive maintenance, drilling optimization, and operational safety, ensuring our customers can achieve maximum efficiency and uptime., * Collaborate with drilling engineers, reliability experts, and data scientists to turn business problems into technical solutions.

  • Design, develop, and operationalize data pipelines and analytics to support NOV’s drilling optimization and condition-based maintenance programs.
  • Prepare, transform, and validate structured and unstructured operational data and high-frequency sensor data for analytics.
  • Productionize and deploy data pipelines and analytical models in cloud and hybrid environments.
  • Build and support automated testing, CI/CD workflows, and deployment pipelines.
  • Implement data quality validation and monitoring.
  • Adhere to and maintain architecture, coding, and documentation standards.
  • Proactively identify technology risks and propose solutions to ensure data availability and reliability in mission-critical systems.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, Data Engineering, or related field.
  • 4+ years of hands-on development experience in Python or another modern programming language.
  • Strong SQL skills and experience with modern data platforms and technologies (e.g., Databricks, SQL Server, SQL/NoSQL databases).
  • Strong understanding of relational database design, normalization, and dimensional data modeling.
  • Experience building and maintaining batch and streaming ELT/ETL pipelines.
  • Experience preparing, transforming, and cleaning data for analytics or machine learning applications.
  • Experience with cloud platforms (AWS preferred).
  • Experience with Git, automated testing, CI/CD workflows, and deployment automation.
  • Familiarity with containerization technologies (e.g., Docker).
  • Familiarity with Agile software development practices and tools (e.g., JIRA, Confluence).
  • Ability to collaborate effectively with cross-functional teams.

Preferred:

  • Experience with Spark, PySpark, or similar distributed data processing frameworks.
  • Databricks or cloud data engineering certification (or equivalent hands-on experience).
  • Experience working with industrial IoT, telemetry, SCADA, WITSML, or other operational technology (OT) data sources.
  • Background in Mechanical Engineering or related field is a plus.
  • Familiarity with drilling, rig operations, condition monitoring, condition-based maintenance, and equipment reliability concepts is a plus.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on egay.fa.us6.oraclecloud.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:05 min

Integrating an assistant application with Jira software

Felix Augenstein · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all