Data Engineer

Calpine
Houston, TX, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
8 years minimum
Working hours
Regular working hours
Job source

Tech stack

Query Performance Business Intelligence Development Cloud Engineering Information Engineering Data Governance Data Transformation Query Languages Dimensional Modeling Document-Oriented Databases Machine Learning Performance Tuning Power BI
+22 more
Azure Machine Learning Azure Data Lake Software Deployment SQL Server Reporting Services SQL Server Integration Services SQL Server Analysis Services Data Streaming System Testing Systems Integration Cloud Platform System Azure Data Factory Sql Optimization Large Language Models Microsoft Fabric Data Lakes Pyspark Data Analytics Star Schema Azure Synapse Analytics Data Pipelines Powerapps Databricks

Job description

As a Data Engineer at Calpine Inc., you will contribute to the BI development to optimize energy operations, enhance forecasting, and improve operational efficiency. This role is ideal for professionals with 10+ years of experience who are passionate about comprehensive data and BI solutions development and solving real-world energy sector challenges using PySpark and Azure Data Factory., * Engage with business users to gather requirements and translate them into comprehensive data and BI solutions.

  • Design and implement scalable data pipelines using PySparkand Azure Data Factory, integrating with Azure Data Lake, Synapse Analytics, DataBricks, and Microsoft Fabric.
  • Develop and optimize data models and visualizations using Power BI, ensuring performance and usability across business units.
  • Lead modernization efforts by migrating legacy BI systems to cloud-based platforms.
  • Collaborate with cross-functional teams to ensure alignment with architectural standards and business objectives.
  • Oversee system testing, user acceptance testing (UAT), and production deployments.
  • Maintain and enhance existing BI solutions, ensuring reliability and scalability.
  • Document data flows, transformation logic, and architectural decisions.
  • Communicate effectively with stakeholders regarding upcoming changes and coordinate change management activities.

Requirements

  • Minimum 8-10 years of experience in data engineering and BI development.
  • Proven expertise in PySpark for data transformation and pipeline development.
  • Extensive experience with Azure Data Services, including Data Factory, Data Lake, Synapse,Microsoft Fabricand DataBricks.
  • Strong proficiency in Power BI dashboard development and data modeling.
  • Solid understanding of cloud architecture, data governance, and performance optimization.
  • Advanced SQL skills with a focus on query performance and scalability.
  • Experience with dimensional modeling (Star and Snowflake schemas).
  • Excellent communication and collaboration skills in cross-functional environments.
  • Ability to manage multiple priorities in a dynamic and fast-paced setting.
  • Familiarity with SSIS, SSRS, and SSAS(tabular and multi-dimensional models).
  • Experience with DAX and MDX query languages.
  • Exposure to Power Apps, Data Flows
  • Knowledge of machine learning (ML)and artificial intelligence (AI) concepts and their application in data analytics.
  • Have worked with Azure ML Services, and understand OpenAI LLM Model.
  • Experience implementing role-based and row-level security in BI platforms.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

3:48 min

Standardizing data access schemas with OData

Florian Bader Florian Bader · World Congress 2026 Europe

Videos

See all

Related articles

See all