Data Engineer 2

Cps, Inc.
San Antonio, TX, United States
about 1 month ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours

Tech stack

RESTful API Modeling Language Application Programming Interfaces (APIs) Amazon Web Services Data Analysis Unit Testing Microsoft Azure Big Data Cloud Computing Databases Data Architecture Data Integration Extract Transform Load (ETL)
+37 more
Data Warehousing Desktop Computing Apache Hadoop MapReduce Apache HBase Apache Hive JSON Python (Programming Language) PostgreSQL MongoDB MySQL NumPy Oracle (Applications) Performance Tuning Anypoint Studio Tensorflow Simple Object Access Protocol (SOAP) PL-SQL SQL Databases Data Streaming Systems Integration Test Case Web Services Extensible Markup Language (XML) Enterprise Software Applications Informatica Powercenter Apache Spark Pandas Scikit Learn Information Technology Druid Apache Kafka Apache Nifi Data Objects Data Pipelines Api Management Mulesoft

Job description

Provide the development and automation of computing processes to detect, predict and respond to opportunities in business operations. Working with a variety of disparate datasets that encompass many disciplines and business units including weather, transmission and distribution grid infrastructure, power generation, gas delivery, commercial market operations, safety and security and customer engagement. Strive to transform and implement true business integration, leveraging top-notch data integration best practices. Merging and securing data in a way that reduces the cost to maintain and increases the utilization of enterprise-wide data as an asset. Developing business intelligence.

Tasks and Responsibilities

  • Design, Develop, and unit test new or existing ETL/Data Integration solutions to meet business requirements.
  • Daily production support for Enterprise Data Warehouse including ETL/ELT jobs.
  • Design and Develop data integration/engineering workflows on big data technologies and platforms (Hadoop, Spark, MapReduce, Hive, HBase, MongoDB, Druid)
  • Develop data streams using Apache Spark, Nifi and/or Kafka. Strong Python development for data transfers and extractions (ELT or ETL)
  • Develop workflows in the cloud environment using Cloud base architecture (Azure or AWS)
  • Develop dataflows and processes for the Data Warehouse using SQL (Oracle, Postgres, HIVEQL, SparkSQL & Dataframes)
  • Perform data analysis & model prototyping using Spark/Python/SQL and common data science tools & libraries (e.g. NumPy, Pandas, scikit-learn, TensorFlow)
  • Develop Data integration workflows using Web services in XML, JSON, flat file format, SOAP
  • Participate in troubleshooting and resolving data integration issues such as data quality.
  • Deliver increased productivity and effectiveness through rapid delivery of high-quality applications.
  • Provide work estimates and communicate status of assignments.
  • Assist in QA efforts on tasks by providing input for test cases and supporting test case execution.
  • Analyze transaction errors, troubleshoot issues in the software, develop bug-fixes, involved in performance tuning efforts.
  • Makes some independent decisions and recommendations which affect the section, department and/or division.
  • Participates and provides input to area budget. Works within financial objectives/budget set by management.
  • Develops alternative solutions for decision-making which support organizational goals/objectives and budget constraints.
  • Works with minimum supervision, conferring with superior on unusual matters. Incumbents have considerable freedom to decide on work priorities and procedures to be followed. May include limited supervisory responsibilities.
  • Provide reporting and analytics functionality to monitor API usage and load (overall hits, completed transactions, number of data objects returned, amount of compute time and other internal resources consumed, volume of data transferred).
  • Use results from API reporting/analytics to guide API Developer offering within an organization’s overall continuous improvement process and for defining software Service-Level Agreements for APIs.
  • Performs other duties as assigned.

Requirements

Experience in a data integration role. Experience using Apache Spark, Nifi and/or Kafka. Experience using Python. Experience integrating enterprise software using ETL modules. Knowledge of data architecture, structures and principles with the ability to critique data and system designs. Ability to design, create and/or modify data processes that meet key timelines while conforming to predefined specifications utilizing the Informatica and/or Mulesoft platform. Understanding of big data technologies and platforms (Hadoop, Spark, MapReduce, Hive, HBase, MongoDB). Ability to integrate data from Web services in XML, JSON, flat file format, SOAP. Knowledge of core concepts of RESTful API Modeling Language (RAML 1.0) and designing with MuleSoft solutions., * Relevant Certifications

  • Ability to write code to build ETL / ELT data pipelines on premise/Cloud
  • Experience in API Management
  • Proficiency with the following databases/technologies: Mulesoft Anypoint Studio, Informatica PowerCenter, Oracle RDMS, PL/SQL, MySQL
  • Professional experience in a technology organization, Demonstrating Initiative Communicates Effectively Using Computers and Technology Driving for Results

Minimum Education

Bachelor’s degree in Computer Science, Engineering, or related field from an accredited university.

Required Certifications

Working Environment

Indoor work, operating computer, manual dexterity, talking, hearing, repetitive motion. Use of personal computing equipment, telephone, multi-functioning printer and calculator. Ability to travel to and from meetings, training sessions or other business related events.

Physical Demands

Exerting up to 10 pounds of force occasionally, and/or a negligible amount of force frequently or constantly to lift, carry, push, pull or otherwise move objects, including the human body. Sedentary work involves sitting most of the time. Jobs are sedentary if walking and standing are required only occasionally, and all other sedentary criteria are met.

About the company

We are engineers, high line workers, power plant managers, accountants, electricians, project coordinators, risk analysts, customer service operators, community representatives, safety and security specialists, communicators, human resources partners, information technology technicians and much, much more. We are 3,500 people committed to enhancing the lives of the communities we serve. Together, we are powering the growth and success of our community progress every day!

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on careers.cpsenergy.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:18 min

Scaling MySQL databases for massive user growth

Johannes Nicolai Johannes Nicolai +1 · LIVE

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:48 min

Analyzing network packets with database protocol tools

Daniël van Eeden Daniël van Eeden · WWC Europe 2026

2:03 min

Distinguishing type definition constructs from data validation routines

Clemens Vasters Clemens Vasters · WWC 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all