Python/Big Data Developer

Matlen Silver
New York, NY, United States
about 1 month ago
Apply on www.indeed.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Compensation
$124,800.0 - $141,440.0
Working hours
Regular working hours
Job source

Tech stack

Apache HTTP Server Big Data Cloudera Impala Databases Data Infrastructure Distributed Data Store Distributed Systems Apache Hadoop Python (Programming Language) PostgreSQL MongoDB Object-Oriented Software Development
+8 more
Software Construction Enterprise Data Management Data Processing Fastapi Data Lakes Pyspark Information Technology Data Management

Job description

  • Develop and maintain Python applications supporting enterprise Data Lake platforms.
  • Participate in Data Lake migration and modernization efforts.
  • Build and optimize scalable data processing pipelines using PySpark.
  • Work with Apache Iceberg, Hadoop, and Impala to develop and maintain large-scale data solutions.
  • Collaborate with engineering teams to improve performance, reliability, and scalability of the data platform.
  • Troubleshoot production issues and support ongoing platform maintenance.
  • Write clean, maintainable, and reusable Python code following software engineering best practices.

Requirements

Overview We are seeking a Python Developer with strong Big Data and Data Lake engineering experience to support Data Lake migration, modernization, and ongoing platform maintenance initiatives. This role is ideal for an engineer who is passionate about building scalable data solutions and has experience working with modern Big Data technologies.

We are open to candidates with either:

  • Strong Big Data/Data Lake engineering experience complemented by Python development, or
  • Strong Python development experience with solid exposure to Big Data engineering.

Required Skills

  • Strong Python development experience
  • Experience working with Big Data/Data Lake technologies
  • Hands-on experience with PySpark
  • Experience with Apache Iceberg
  • Experience with Hadoop
  • Experience with Impala
  • Experience supporting Data Lake migration, engineering, and maintenance initiatives
  • Strong understanding of data processing, distributed computing, and scalable data platforms

Preferred Qualifications

  • Experience with object-oriented databases such as:
  • MongoDB
  • PostgreSQL, * Bachelor’s or Master’s degree in Computer Science, Engineering, or a related technical discipline.
  • Experience developing scalable applications using Python.
  • Hands-on experience with Big Data ecosystems and Data Lake technologies.
  • Strong analytical and problem-solving skills.
  • Excellent communication and collaboration abilities.

Preferred Candidate Profile We’re looking for a software engineer who combines strong Python development skills with experience building and maintaining enterprise-scale Data Lake platforms. The ideal candidate has worked with distributed data technologies such as PySpark, Apache Iceberg, Hadoop, and Impala, and has experience supporting Data Lake migrations. Experience with object-oriented databases-including MongoDB, PostgreSQL, or financial industry platforms

Benefits & conditions

$60 - $68 an hour - Contract, Pulled from the full job description

  • 401(k)
  • Health insurance
  • Vision insurance
  • Dental insurance, * Health, vision, and dental insurance (single and family coverage)
  • 401(k) plan (employee contributions only)

About the company

Experience Matters. Let your experience be driven by our experience. For more than 40 years, Matlen Silver has delivered solutions for complex talent and technology needs to Fortune 500 companies and industry leaders. Led by hard work, honesty, and a trusted team of experts, we can say that Matlen Silver technology has created a solutions experience and legacy of success that is the difference in the way the world works.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.indeed.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:33 min

Connecting frontends via a FastAPI proxy backend layer

Saoussen Chaabnia Saoussen Chaabnia · Europe 2026 Virtual

2:01 min

Migrating existing applications from MongoDB to Postgres

Nikita Shamgunov Nikita Shamgunov · World Congress 2024

4:43 min

Building an anti-money laundering production architecture with Python

Stefan Donsa Stefan Donsa +1 · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all