> Markdown version of [/jobs/ext/2629765-cloudera-developer](https://www.wearedevelopers.com/jobs/ext/2629765-cloudera-developer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Cloudera Developer - **Company:** Genesis10 - **Location:** Charlotte, NC, United States - **Experience:** Expert - **Contract:** Temporary contract - **Skills:** Data Analysis, CA Workload Automation Ae, Big Data, Data Architecture, Data Governance, Data Integration, Data Integrity, Extract Transform Load (ETL), Data Visualization, Database Queries, Distributed Data Store, Distributed Systems, Apache Hadoop, Apache Hive, Python (Programming Language), MySQL, Performance Tuning, Cloudera, Software Deployment, Data Streaming, Data Processing, Apache Spark, Pandas, Pyspark, Data Pipelines - **Published:** August 4, 2026 - **Apply:** https://www.dice.com/job-detail/de0d404d-19f7-4bc9-92db-3cc8c4426088 ## About the Role * 5-7 years of experience with distributed data/computing tools such as Hadoop, Hive, MySQL, etc. * Strong problem-solving skills with an emphasis on product development * Experience working with and creating data architectures * Proficient in Python programming for data processing, automation, and analytical solutions * Experience using PySpark for large-scale data processing and distributed computing environments * Strong SQL skills with experience developing complex queries, joins, aggregations, and performance optimization * Knowledge of ETL processes, data integration, and data quality best practices * Experience using pandas to cleanse, transform, and analyze large datasets * Experience working with AutoSys * Ability to analyze data, identify trends and anomalies, and deliver actionable business insights * Experience creating clear and effective data visualizations and dashboards * Excellent written and verbal communication skills for coordinating across teams ## Description As a Cloudera Developer, you will be responsible for developing and maintaining robust, scalable data solutions using the Cloudera platform. This role involves designing and implementing data processing pipelines and ingestion frameworks, primarily using PySpark, while working with cross-functional teams to meet data requirements and deliver actionable business insights., * Design, develop, and maintain data processing pipelines using Cloudera technologies such as Apache Hadoop, Apache Spark, Apache Hive, and Python * Utilize multiple architectural components in the design and development of client requirements * Maintain, improve, clean, and manipulate data for operational and analytical data systems * Develop and maintain data ingestion frameworks for efficiently extracting, transforming, and loading data from various sources * Collaborate with data engineers and data scientists to understand data requirements and translate them into technical specifications * Optimize and tune data processing jobs to ensure high performance and scalability * Implement data governance and security policies to ensure data integrity and compliance * Monitor and troubleshoot data processing jobs to identify and resolve issues in a timely manner * Document technical specifications, data flows, and data architecture diagrams * Adhere to team delivery/release process and cadence for code deployment and release ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Everything a Developer Needs to Know About MCP with Neo4j](https://www.wearedevelopers.com/magazine/604-everything-a-developer-needs-to-know-about-mcp-with-neo4j) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)