Data Engineer and Analyst - Hybrid

Genesis10
Charlotte, NC, United States
3 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$174,845.0 - $191,485.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Airflow Amazon Web Services Data Analysis Application Release Automation Microsoft Azure Big Data Cloud Computing Code Review Continuous Integration Data as a Services
+34 more
Data Deduplication Information Engineering Relational Databases JSON Python (Programming Language) Machine Learning Microsoft SQL Server Modular Design MongoDB OpenShift Parsing Performance Tuning Systems Development Life Cycle SQL Databases Data Streaming Test Case Extensible Markup Language (XML) Parquet Google Cloud Feature Engineering Sql Optimization Large Language Models Git Pandas Containerization Data Lakes Pyspark Semi-structured Data Kubernetes Apache Kafka Feature Extraction GPT Data Pipelines Docker

Job description

We are hiring a hands-on Senior Data Engineer and Analyst who can both build data pipelines and analyze large datasets. You will design and deliver scalable services in Python and Java on Red Hat OpenShift (OCP) and cloud platforms. This role involves operating within agile backlogs, independently breaking down complex data problems, and serving as a technical leader for SQL, Python, data modeling, ML application, and practical GenAI use., * Own full-cycle data problem solving: profile large datasets, design pipelines, engineer transformations, and perform deep analysis to identify patterns, outliers, and root causes

  • Implement production-grade code: develop data services and utilities in Python (primary) and Java (for service implementations) with strong testing, observability, and reliability
  • Write advanced SQL for data exploration, deduplication, quality checks, and performance-tuned analytics across data lakes/warehouses
  • Parse, normalize, and analyze XML payment objects to identify unique examples, apply deduplication strategies, and generate representative test cases
  • Apply LLM-powered techniques (Gemini/GPT) to accelerate data triage, test generation, and anomaly detection with careful evaluation and guardrails
  • Apply machine learning knowledge (e.g., feature extraction, similarity measures, dedup/record-linkage techniques) to large-scale data problems
  • Set engineering standards for architecture, SDLC excellence, CI/CD automation, and code review quality

Requirements

  • 5-10+ years of combined data engineering/analysis experience
  • Expert Python skills for data engineering & analysis (pandas, PySpark or similar, modular design, testing)
  • Advanced SQL skills (analytical window functions, performance optimization, CTEs, partitioning, large-scale joins)
  • Java experience for service implementations (APIs, data services, utilities), with strong SDLC discipline
  • Proven experience with large datasets: profiling, cleaning, deduping, and synthesizing insights; comfort with semi-structured data (XML/JSON)
  • Experience with data lake / warehouse concepts (e.g., parquet, object storage, lakehouse patterns)
  • Hands-on CI/CD (Git, pipelines, build/test/release automation) and containerized deployments (Docker/Kubernetes; OpenShift/OCP highly preferred)
  • Independent problem solver with the ability to break down ambiguous data issues, form hypotheses, validate with code, and communicate outcomes clearly
  • Practical GenAI usage: ability to craft prompts and evaluate LLM outputs for data triage, test case generation, and analysis acceleration
  • Foundational ML knowledge: familiarity with applying models/techniques relevant to data quality (e.g., clustering, similarity, dedup/record linkage, anomaly detection)

Desired skills:

  • Payments domain experience (especially wire payments)
  • Experience with streaming (Kafka), workflow/orchestration (Airflow), and feature engineering for ML
  • Familiarity with Microsoft SQL Server or similar enterprise RDBMS
  • Cloud platform experience (Google Cloud Platform, AWS, Azure)
  • MongoDB experience

About the company

Genesis10 is currently seeking a Senior Data Engineer and Analyst - Hybrid for a contract position with a Global Financial Institution located in Iselin, NJ or Charlotte, NC or Boston, MA or Dallas, TX. This is an 18+ month contract opportunity., Ranked a Top Staffing Firm in the U.S. by Staffing Industry Analysts for six consecutive years, Genesis10 puts thousands of consultants and employees to work across the United States every year in contract, contract-for-hire, and permanent placement roles. With more than 300 active clients, Genesis10 provides access to many of the Fortune 100 firms and a variety of mid-market organizations across the full spectrum of industry verticals.

For contract roles, Genesis10 offers the benefits listed below. If this is a perm-placement opportunity, our recruiter can talk you through the unique benefits offered for that particular client. Benefits of Working with Genesis10:

  • Access to hundreds of clients, most who have been working with Genesis10 for 5-20+ years.
  • The opportunity to have a career-home in Genesis10; many of our consultants have been working exclusively with Genesis10 for years.
  • Access to an experienced, caring recruiting team (more than 7 years of experience, on average.)
  • Behavioral Health Platform
  • Medical, Dental, Vision
  • Health Savings Account
  • Voluntary Hospital Indemnity (Critical Illness & Accident), For multiple years running, Genesis10 has been recognized as a Top Staffing Firm in the U.S., as a Best Company for Work-Life Balance, as a Best Company for Career Growth, for Diversity, and for Leadership, amongst others. To learn more and to view all our available career opportunities, please visit us at our website.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

Videos

See all

Related articles

See all