Big Data Engineer

Aivra Health Llc
San Jose, CA, United States
25 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Airflow Amazon Web Services Big Data Information Engineering Data Integrity Extract Transform Load (ETL) Data Warehousing Distributed Computing Environment Apache Hadoop Apache Hive Python (Programming Language) Machine Learning
+13 more
Standard Sql Unstructured Data Data Storage Technologies Cloud Platform System Snowflake Apache Spark Git Kubernetes Information Technology Apache Kafka Data Pipelines Docker Databricks

Job description

We are seeking a highly skilled Big Data Engineer to design, develop, and maintain scalable big data solutions that process large volumes of structured and unstructured data. You will collaborate with Data Engineers, Data Scientists, and Software Engineers to build reliable data pipelines and support advanced analytics initiatives., * Design, build, and maintain scalable big data pipelines.

  • Develop ETL/ELT workflows for large-scale data processing.
  • Optimize data storage, retrieval, and processing performance.
  • Work with distributed computing frameworks and cloud-based data platforms.
  • Integrate data from multiple sources into enterprise data lakes and warehouses.
  • Monitor data quality and ensure data integrity.
  • Collaborate with cross-functional teams to support business intelligence and machine learning initiatives.
  • Troubleshoot and resolve data pipeline and infrastructure issues.
  • Maintain technical documentation and follow best engineering practices.

Requirements

  • Bachelor’’s degree in Computer Science, Information Technology, Data Engineering, or a related field.
  • Strong knowledge of SQL and Python or Scala.
  • Experience with distributed data processing frameworks.
  • Understanding of ETL/ELT concepts and data warehousing.
  • Strong analytical, problem-solving, and communication skills.

Preferred Skills

  • Apache Spark
  • Hadoop
  • Kafka
  • Hive
  • Scala
  • Python
  • SQL
  • Databricks
  • Snowflake
  • Apache Airflow
  • AWS
  • Docker
  • Kubernetes
  • Git

Benefits & conditions

  • Health Insurance
  • Paid Time Off
  • Flexible Work Schedule
  • 401(k)
  • Learning & Certification Support
  • Career Growth Opportunities
  • Employee Assistance Program (EAP)
  • Professional Development Opportunities

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon · WWC Europe 2026

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy · WWC 2024

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all