Senior Data Engineer - Remote (USA)

ICF Incorporated, L.L.C.
Reston, VA, United States
4 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$89,649.0 - $152,404.0
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Agile Methodology Airflow Amazon Web Services Amazon S3 Big Data C++ (Programming Language) Cloud Computing Software Quality Code Review Data Governance Data Infrastructure
+37 more
Extract Transform Load (ETL) Data Transformation Data Security Data Structures Data Systems Software Design Documents DevOps Github Apache Hive IP Addressing Virtual Private Networks (VPN) Python (Programming Language) PostgreSQL Metadata NoSQL Scala (Programming Language) Software Engineering SQL Databases Visual Analytics Web Services Jupyter Notebook Data Processing Scripting Data Storage Technologies Test-Driven Development (TDD) Delivery Pipeline Apache Spark Electronic Medical Records Integration Tests Luigi Cassandra AWS Glue Build Process Spark Streaming Terraform Data Pipelines Databricks

Job description

  • Design, develop, and maintain scalable data pipelines using Spark, Hive, and Airflow
  • Develop and deploy data processing workflows on the Databricks platform
  • Develop API services to facilitate data access and integration
  • Create interactive data visualizations and reports using AWS QuickSight
  • Builds required infrastructure for optimal extraction, transformation and loading of data from various data sources using AWS and SQL technologies
  • Monitor and optimize the performance of data infrastructure and processes
  • Develop data quality and validation jobs
  • Assembles large, complex sets of data that meet non-functional and functional business requirements
  • Write unit and integration tests for all data processing code
  • Work with DevOps engineers on CI, CD, and IaC
  • Read specs and translate them into code and design documents
  • Perform code reviews and develop processes for improving code quality
  • Improve data availability and timeliness by implementing more frequent refreshes, tiered data storage, and optimizations of existing datasets
  • Maintain security and privacy for data at rest and while in transit
  • Other duties as assigned

Requirements

  • Bachelor’s degree
  • 7+ years of hands-on software development experience
  • 4+ years of experience building data pipelines using Python, Java, and cloud technologies, with hands-on experience leveraging Spark and Hive for large-scale data processing.
  • Candidate must be able to obtain and maintain a Public Trust clearance
  • Candidate must reside in the US, be authorized to work in the US, and work must be performed in the US
  • Must have lived in the US 3 full years out of the last 5 years, * Experience building job workflows with the Databricks platform
  • Strong understanding of AWS products including S3, Redshift, RDS, EMR, AWS Glue, AWS Glue DataBrew, Jupyter Notebooks, Athena, QuickSight, EMR, and Amazon SNS
  • Familiar with work to build processes that support data transformation, workload management, data structures, dependency and metadata
  • Experienced in data governance process to ingest (batch, stream), curate, and share data with upstream and downstream data users.
  • Experienced in data pipeline builder and data wrangler who enjoys optimizing data systems and building them from the ground up.
  • Demonstrated understanding using software and tools including relational NoSQL and SQL databases including Cassandra and Postgres; workflow management and pipeline tools such as Airflow, Luigi and Azkaban; stream-processing systems like Spark-Streaming and Storm; and object function/object-oriented scripting languages including Scala, C++, Java and Python.
  • Familiar with DevOps methodologies, including CI/CD pipelines (Github Actions) and IaC (Terraform)
  • Ability to obtain and maintain a Public Trust; residing in the United States
  • Experience with Agile methodology, using test-driven development.

Job Location: This position requires that the job be performed in the United States. If you accept this position, you should note that ICF does monitor employee work locations and blocks access from foreign locations/foreign IP addresses, and also prohibits personal VPN connections.

About the company

ICF is a mission-driven company filled with people who care deeply about improving the lives of others and making the world a better place. Our core values include Embracing Difference; we seek candidates who are passionate about building a culture that encourages, embraces, and hires dimensions of difference.

Our Health Engineering Systems (HES) team works side by side with customers to articulate a vision for success, and then make it happen. We know success doesn’t happen by accident. It takes the right team of people, working together on the right solutions for the customer. We are looking for a seasoned Senior Data Engineer who will be a key driver to make this happen., ICF is a global advisory and technology services provider, but we’re not your typical consultants. We combine unmatched expertise with cutting-edge technology to help clients solve their most complex challenges, navigate change, and shape the future.

We can only solve the world’s toughest challenges by building a workplace that allows everyone to thrive. We are an equal opportunity employer. Together, our employees are empowered to share their expertise and collaborate with others to achieve personal and professional goals. For more information, please read our EEO (https://www.icf.com/legal/equal-employment-opportunity) policy.

We will consider for employment qualified applicants with arrest and conviction records., At ICF, we are committed to ensuring a fair interview process for all candidates based on their own skills and knowledge. As part of this commitment, the use of artificial intelligence (AI) tools to generate or assist with responses during interviews (whether in-person or virtual) is not permitted. This policy is in place to maintain the integrity and authenticity of the interview process.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dejobs.org

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters ¡ WWC 2023

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva ¡ JS Congress

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt ¡ LIVE

2:40 min

Using GitHub primitives for internal documentation and corporate operations

Kyle Daigle ¡ Coffee With Developers

3:16 min

Terminology differences between relational and NoSQL databases

Tim Faulkes ¡ LIVE

2:57 min

Core technical practices for robust data engineering

Sandhya Menon Sandhya Menon ¡ WWC Europe 2026

Videos

See all

Related articles

See all