Big Data Engineer - Remote

YO AI Labs View all jobs
New York, NY, United States
10 days ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Compensation
$104,000.0 - $166,400.0
Working hours
Regular working hours

Tech stack

Amazon Web Services Microsoft Azure Big Data Cloud Computing Data Integration Extract Transform Load (ETL) Data Systems Data Warehousing Distributed Computing Environment Distributed Data Store Apache Hadoop Python (Programming Language)
+8 more
NoSQL Data Processing Google Cloud System Availability Apache Spark Apache Flink Machine Learning Operations Data Pipelines

Job description

Job Summary: In this role, you’ll apply your expertise to help train next-generation AI systems. Your work will shape how models learn, reason, and perform through high-quality, real-world input., * Design, build, and maintain scalable big data pipelines and architectures to support robust data solutions.

  • Collaborate with cross-functional teams to understand and deliver on data requirements and business objectives.
  • Implement data integration, transformation, and processing solutions using Python and relevant big data technologies.
  • Develop, manage, and optimize distributed databases and storage systems for efficiency and reliability.
  • Monitor, troubleshoot, and enhance data systems to ensure high availability and performance.
  • Enforce data quality, security, and governance standards across all solutions.
  • Document solutions and communicate complex technical concepts effectively, both in writing and verbally.

Requirements

  • Proven expertise in big data engineering with hands-on experience building and maintaining large-scale data pipelines.
  • Advanced proficiency in Python for data processing, automation, and integration.
  • Deep understanding of relational and NoSQL databases, including optimization and management techniques.
  • Experience with distributed data processing frameworks (e.g., Hadoop, Spark, Flink).
  • Strong foundation in data modeling, ETL processes, and data warehousing principles.
  • Excellent written and verbal communication skills, with the ability to convey technical ideas clearly to both technical and non-technical stakeholders.
  • Detail-oriented, proactive, and self-motivated, thriving in remote and autonomous work environments., * Prior experience in fast-paced or startup-like environments supporting global teams.
  • Expertise with cloud-based big data platforms (e.g., AWS, GCP, Azure).
  • Familiarity with machine learning operations and data science workflows is a plus.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:55 min

Infrastructure challenges when combining Kafka with Apache Flink

Bobur Umurzokov · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

2:37 min

Comparing traditional SQL tables versus NoSQL non-tabular databases

Stanimira Vlaeva · JS Congress

1:59 min

Evolving roles in AI driven software teams

Ignacio Riesgo Ignacio Riesgo +1 · World Congress 2024

1:09 min

Evaluating mature stream processing frameworks for production systems

Soroosh Khodami Soroosh Khodami · World Congress 2024

Videos

See all

Related articles

See all