Data Engineer

Habitat Energy
Austin, TX, United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Training Data Application Programming Interfaces (APIs) Airflow Amazon Web Services Amazon S3 Data Analysis Cloud Computing Information Engineering Data Infrastructure Data Visualization Data Warehousing Relational Databases
+25 more
Database Design Python (Programming Language) PostgreSQL Machine Learning NumPy RabbitMQ Software Tools Prometheus Server Administration Software Engineering Working Model 2D Parquet Scripting Feature Engineering Grafana Git Pandas Containerization Kubernetes Xgboost Data Management Machine Learning Operations Terraform Data Pipelines Docker

Job description

  • Supporting Data Engineering Infrastructure: *

  • Contribute to the design, development, implementation and continuous improvement of our data engineering tools, workflows, processes, and platforms. This includes enhancing the architectural foundations and integrating new data management technologies.
  • Writing Well-Structured Code: *

  • Develop clean, maintainable, well-documented code that adheres to best practices. Support best coding practices within Habitat’s software, machine-learning, and data science teams.
  • Enhance data engineering knowledge: *

  • Improve expertise within the software team and ensure their ability to support and collaborate on the data infrastructure infrastructure.
  • Data Quality Management: *

  • Continuously enhance data quality across multiple dimensions such as accuracy, availability, performance, and accessibility to ensure a clear understanding of data within the company.
  • Providing backup/escalation to the tech-on-call team.
  • Communicating effectively across Software and Data Science teams.

Requirements

Preferred skills and experience:

  • 3+ years of Python experience.
  • 3+ years of working in technical teams, building data pipelines, delivering productionised code, building/maintaining live applications, developing tooling and improving backtesting frameworks.
  • Experience in applying relational database design.
  • Proficiency with Orchestration and IaC in AWS (e.g. Terraform, Kubernetes, RabbitMQ, Airflow, Prefect), Git, containerisation (Docker), database management (e.g. Postgres, Alembic).
  • Fluent in Python and its wider numerical ecosystem (e.g. Pandas, NumPy, Polars, Pydantic).

‘Nice to have’ skills and experience:

  • 2+ years of orchestrating machine learning workflows.
  • Experience with OSS data warehousing tooling and management.
  • Cloud infrastructure experience.
  • Experience with monitoring frameworks (e.g. Prometheus).
  • Experience archiving data to Parquet on S3 and creating tools for API/Grafana queries.
  • Experience centralising diverse datasets for analytics, visualisation and machine learning.
  • Familiarity with time-series forecasting and/or optimisation.
  • Experience with data visualisation and dashboards (e.g. Grafana, Superset).
  • Familiarity with machine learning and associated techniques (feature engineering, boosting methods, LightGBM).

Ultimately we are looking for someone who is a great fit for our company so we encourage you to apply even if you may not meet every requirement in this posting. We value diversity and our environment is supportive, challenging and focused on the consistent delivery of high quality, meaningful work.

In return, we’ll give you a competitive salary, flexible working arrangements and a lot of personal development opportunities. We operate a hybrid working model with at least 2 days in our office in Austin., Alternative Energy, Amazon Simple Storage Service (S3), Amazon Web Services (AWS), Application Programming Interface (API), Architectural Services, Best Practices, Building Codes, Cloud Computing, Communication Skills, Continuous Improvement, Data Management, Data Quality, Data Science, Data Visualization, Data Warehousing, Database Administration, Database Design, Diversity, Docker, Ecosystems, Energy Management, Forecasting, Git, Machine Learning, Machine Tool, On Call, Operational Support Systems (OSS), PostgreSQL, Python Programming/Scripting Language, Quality Management, RabbitMQ, Relational Databases (RDBMS), Reporting Dashboards, Software Administration, Software Engineering, Team Player, Training Data Sets

About the company

Habitat Energy is a fast growing technology company focussed on the physical and financial optimisation of energy storage and renewable generation assets globally through complex models and trading. By maximising the returns from these assets we aim to drive investment in renewable energy and accelerate the transition to a low carbon world. Our rapidly growing team of 130+ people in Austin, TX, Oxford, UK, and Melbourne, Australia brings together exceptionally talented and passionate people in the domains of energy trading, data science, software engineering and renewable energy management.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerbuilder.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all