Data Engineer

Spectraforce
Richmond, VA, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Amazon Web Services Amazon Elastic Compute Cloud Amazon S3 Automation of Tests Cloud Computing Security Code Review Continuous Integration Information Engineering Extract Transform Load (ETL) Data Security
+23 more
Dataspaces Data Systems Software Debugging DevOps Distributed Computing Environment Identity and Access Management Python (Programming Language) Operational Databases Performance Tuning Cloud Services SQL Databases Scripting Freeform SQL Snowflake Apache Spark State Machines Backend Git Pyspark Cloudwatch Data Pipelines Serverless Computing Golang

Job description

*Design, develop, and maintain data pipelines leveraging Python, Spark/PySpark, and cloud-native services. *Build and optimize data workflows, ETL processes, and transformations for large-scale structured and semi-structured datasets. *Write advanced and efficient SQL queries against Snowflake, including joins, window functions, and performance tuning. *Develop backend and automation tools using Golang and/or Python as needed. *Implement scalable, secure, and high-quality data solutions across AWS servies such as S3, Lambda, Glue, Step Functions, EMR, and CloudWatch. *Troubleshoot complex production data issues, including pipeline failures, data quality gaps, and cloud environment challenges. *Perform root-cause analysis and implement automation to prevent recurring issues. *Collaborate with data scientists, analysts, platform engineers, and product teams to enable reliable, high-quality data access. *Ensure compliance with enterprise governance, data quality, and cloud security standards. *Participate in Agile ceremonies, code reviews, and DevOps practices to ensure high engineering quality.

Requirements

*Skills-Data Engineer- Python , Spark/PySpark, AWS, Golang, Able to write complex SQL queries against Snowflake tables / Troubleshoot issues, Java/Python, AWS (Glue, EC2, Lambda). *Proficiency in Python with experience building scalable data pipelines or ETL processes. *Strong hands-on experience with Spark/PySpark for distributed data processing. *Experience writing complex SQL queries (Snowflake preferred), including optimization and performance tuning. *Working knowledge of AWS cloud services used in data engineering (S3, Glue, Lambda, EMR, Step Functions, CloudWatch, IAM). *Experience with Golang for scripting, backend services, or performance-critical processes. *Strong debugging, troubleshooting, and analytical skills across cloud and data ecosystems. *Familiarity with CI/CD workflows, Git, and automated testing.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on leoforce.us

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · WWC Europe 2026

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon · WWC Europe 2026

Videos

See all

Related articles

See all