Python Developer

Propertyvalue Pacific Consultancy Services
United States
2 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Data Analysis Microsoft Azure Big Data Cloud Computing Databases Continuous Integration Data Validation Data Cleansing Information Engineering Extract Transform Load (ETL) Data Transformation
+16 more
Software Debugging Python (Programming Language) Software Engineering SQL Databases Test Data Automated Data Processing (ADP) Data Processing Scripting Google Cloud Apache Spark Git Pyspark Data Analytics Restful APIs Databricks Data Generation

Job description

We are looking for an experienced Python Developer with Databricks expertise to design, develop, and maintain Python-based data processing solutions. The ideal candidate will have strong Python programming skills, hands-on experience with Databricks notebooks and pipelines, and practical knowledge of data processing, analytics, and synthetic test-data generation., * Develop robust and scalable applications, utilities, and data-processing solutions using Python.

  • Build and maintain Databricks notebooks for data transformation, processing, validation, and analysis.
  • Develop and manage Databricks pipelines and workflows for automated data processing.
  • Create Python scripts and frameworks for synthetic data generation and test-data preparation.
  • Generate realistic datasets covering normal, boundary, and exception scenarios for testing.
  • Perform data cleansing, transformation, validation, and analysis using Python and Databricks.
  • Work with large datasets and troubleshoot data-processing and pipeline issues.
  • Develop reusable Python modules, utilities, and automation scripts.
  • Analyze data to identify inconsistencies, anomalies, and quality issues.
  • Collaborate with QA, data engineering, analytics, and application development teams.
  • Optimize Python code and Databricks processing workflows for performance and reliability.
  • Troubleshoot failures in notebooks, pipelines, and data-processing jobs.
  • Independently manage development activities in a fast-paced environment.

Requirements

  • Strong hands-on Python development experience.
  • Hands-on experience with Databricks.
  • Strong experience developing and working with Databricks notebooks.
  • Experience creating and managing Databricks pipelines/workflows.
  • Experience with synthetic data generation or test-data generation.
  • Good understanding of data processing and data transformation concepts.
  • Good understanding of data analytics and data validation.
  • Strong debugging and problem-solving skills.
  • Ability to work independently and take ownership of assigned deliverables.

Preferred Skills

  • PySpark / Apache Spark
  • SQL and database technologies
  • ETL/ELT development
  • Data quality and validation
  • Python automation
  • REST APIs
  • Cloud platforms such as AWS, Azure, or Google Cloud Platform
  • CI/CD and Git
  • Agile/Scrum development

Must-Have Skills

Python + Databricks + Databricks Notebooks + Databricks Pipelines + Synthetic/Test Data Generation + Data Processing

Ideal Candidate

A strong candidate should be primarily a Python Developer with substantial hands-on Databricks experience, rather than a purely analytics-focused professional. The person should be comfortable writing production-quality Python code, developing Databricks notebooks and pipelines, and creating test datasets for development and validation activities.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

1:31 min

Baseline developer skills for software data science

Markus Harrer Markus Harrer · World Congress 2021

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

2:34 min

Capabilities of the Apache Spark processing engine

Ayon Roy · LIVE

Videos

See all

Related articles

See all