> Markdown version of [/jobs/ext/3022521-big-data-engineer](https://www.wearedevelopers.com/jobs/ext/3022521-big-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Big Data Engineer - **Company:** ROBERTS RECRUITING, LLC - **Location:** Framingham, MA, United States - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Agile Methodology, Amazon Web Services, Data Analysis, Microsoft Azure, Big Data, Unix, Cloud Computing, Cloudera Impala, Computer Programming, Computer Literacy, Data Architecture, Information Engineering, Data Files, Data Governance, Data Infrastructure, Linux, R (Programming Language), Apache Hadoop, Apache HBase, Apache Hive, Information Management, Python (Programming Language), Machine Learning, Apache Oozie, Scrum Methodology, Standard Sql, Cloudera, Scala (Programming Language), Software Engineering, SQL Databases, Tableau (Software), Wi-Fi Technology, Data Processing, Scripting, Cloud Platform System, Data Ingestion, Apache Spark, Information Technology, Spark Streaming, Data Management, Text Analysis, Data Pipelines - **Published:** September 21, 2026 - **Apply:** https://www.careerbuilder.com/job-details/senior-big-data-engineer-framingham-ma--1d90f6a3-f1d6-4ef0-8e2a-fb66cafdf5d5 ## About the Role We are launching exciting products such as our new Wireless Headphone and WIFI Speakers that are Connected to the Cloud. Come be part of a new team focused on driving business value from the insights gained from these Connected Products!, * Deep expertise is working with data - all kinds, clean, dirty, unstructured, semi-structured * Have strong expertise Apache Spark (batch and Spark streaming) * Experience in real-time and batch data processing and associated technologies * Able to demonstrate strong skills in programming/scripting languages such as Python, Scala, Java and R * Ability do design, develop and deploy end to end data pipelines that meet business requirements and use cases * Experience with all Cloudera components with focus on Impala, Hive, Hbase, Scoop and Hue * Experience with automation technologies such as Oozie * Strong experience with large cloud-compute infrastructure solutions such as Amazon Web Services, Google, Azure * Knowledge of UNIX/Linux * Experience with Text Analytics * Strong knowledge of SQL * Experience in Hadoop Platform Security and Hadoop Data Governance topics * Experience in triaging production issues to understand and resolve the issue * Experience in technical computing (optimization, statistics and machine learning) * Experience with analytics visualization software such as Tableau * Experience leading development teams, defining development processes, evaluating new technologies and practices to enhance solution delivery, * Minimum of BS in Computer Science or similar field required * Must have at least 5-8+ years' experience in information management and application development * Must have a minimum of 3-5 years working hands on with Big Data technologies Skills: Agile Programming Methodologies, Amazon Web Services (AWS), Apache HBase, Apache Hadoop, Apache Hive, Architectural Analysis, Automation, Best Practices, Big Data, Business Case, Business Strategy, Cloud Computing, Cloudera, Coaching, Computer Programming, Computer Science, Computer Skills, Customer Relations, Data Analysis, Data Management, Data Processing, Data Science, Data Sets, Housekeeping/Cleaning, Information Architecture, Java, Leadership, Linux Operating System, Machine Learning, Mentoring, Microsoft Windows Azure, Problem Solving Skills, Product/Service Launch, Python Programming/Scripting Language, R Programming Language, SQL (Structured Query Language), Scala Programming Language, Scripting (Scripting Languages), Scrum Project Management and Software Development, Software Development, Statistics, Tableau, Team Lead/Manager, Unix Operating Systems, Use Cases, User Groups, Wi-Fi, Wireless Communications ## Description We are looking for a Senior Big Data Engineer to join our Big Data Analytics Platform team. In this role, you will be responsible for designing, developing, deploying and supporting the data ingestion pipeline from our Connected Products. In this role, you will be partnering with our data science community by supplying highly performant data sets for advanced analytics and will provide leadership and experience in the data engineering space. The candidate must be results driven, customer focused, technologically savvy, and skilled at working in an agile development environment. Job Responsibilities: * Design and develop the data ingestion pipelines into the Big Data & Analytics Platform for a variety of big data use cases * Deploy and support highly optimized solutions with a focus on automation * Participate in the end to end delivery of business use-cases including data architecture to deliver results * Deliver as part of an agile scrum team on the highest business value use cases that will drive our business strategy * Partner with IT architects to define analytic architecture that best leverages the Big Data Platform to enable advanced analytics capabilities * Provide leadership to the ongoing maturity of the development process, coach/mentor the development team on best practices and methodologies for enhanced solution development. * Stay up to date on the relevant technologies, plug into user groups, understand trends and opportunities ## Related Videos - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [WeAreDevelopers LIVE - Node and Package Security](https://www.wearedevelopers.com/videos/2138-wearedevelopers-live-node-and-package-security) - [Docker network without Docker](https://www.wearedevelopers.com/videos/1418-docker-network-without-docker) - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [Why and when should we consider Stream Processing frameworks in our solutions](https://www.wearedevelopers.com/videos/1085-why-and-when-should-we-consider-stream-processing-frameworks-in-our-solutions) - [Docker exec without Docker](https://www.wearedevelopers.com/videos/1094-docker-exec-without-docker) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [7 Cloud Computing Trends Coming in 2025 for Developers](https://www.wearedevelopers.com/magazine/412-7-cloud-computing-trends-coming-in-2025-for-developers) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again)