> Markdown version of [/jobs/ext/1890321-big-data-engineer](https://www.wearedevelopers.com/jobs/ext/1890321-big-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Big Data Engineer - **Company:** Vision Technologies, LLC - **Location:** Sunnyvale, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $90,000.0 - $110,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Airflow, Amazon Web Services, Apache HTTP Server, Microsoft Azure, Big Data, Cloud Computing, Continuous Delivery, Continuous Integration, Information Engineering, Data Governance, Data Stores, Database Queries, Software Debugging, Distributed Systems, Fault Tolerance, Apache Hadoop, Hadoop Distributed File System, Apache HBase, Apache Hive, Python (Programming Language), Machine Learning, NoSQL, Apache Oozie, Software Engineering, SQL Databases, Sqoop, Data Streaming, Unstructured Data, Data Processing, Scripting, Apache Spark, Hdinsight, Electronic Medical Records, Kubernetes, Information Technology, Collibra, Apache Flink, Apache Kafka, Spark Streaming, Data Management, Amazon Elastic Mapreduce (EMR), Databricks, Programming Languages - **Published:** July 31, 2026 - **Apply:** https://www.careerbuilder.com/job-details/big-data-engineer-sunnyvale-ca--8ccfddd2-479c-4753-97d8-4304fb7fdaeb ## About the Role ingesting, transforming, and analyzing massive volumes of structured and unstructured data to support enterprise analytics, machine learning, and reporting workloads. The ideal candidate will combine deep technical expertise across the Hadoop ecosystem with strong software engineering fundamentals and a clear understanding of how to deliver reliable, performant, and cost-effective data platforms in production environments.Required Qualifications * Bachelor's degree in Computer Science, Engineering, or a related technical discipline. * Five or more years of professional experience designing and operating big-data pipelines on Hadoop. * Strong hands-on expertise with Apache Spark (Scala, Python, or Java) in production environments. * Solid experience with Hive, HDFS, Sqoop, HBase, and the broader Hadoop ecosystem. * Hands-on experience with streaming data platforms such as Kafka, Spark Streaming, or Flink. * Strong SQL skills and experience working with both relational and NoSQL data stores. * Experience with workflow orchestration tools such as Airflow or Oozie. * Solid understanding of distributed systems concepts, including partitioning, replication, and fault tolerance. * Strong scripting skills in Python or Shell. * Excellent troubleshooting, debugging, and documentation skills. Preferred Qualifications * Experience operating Hadoop on cloud platforms such as AWS EMR, Azure HDInsight, or Databricks. * Familiarity with modern lakehouse formats (Delta, Iceberg, Hudi). * Exposure to data governance tooling such as Apache Atlas or Collibra. * Experience with Kubernetes-based data platforms (Spark-on-K8s, Trino). * Hands-on experience with CI/CD and infrastructure-as-code in data engineering workflows., Amazon Web Services (AWS), Apache, Apache HBase, Apache Hadoop, Apache Hive, Apache Spark, Apache Sqoop, Artificial Intelligence (AI), Big Data, Cloud Computing, Computer Science, Consulting, Continuous Deployment/Delivery, Continuous Integration, Data Management, Data Processing, Debugging Skills, Distributed Computing, Documentation, EAD, Ecosystems, Electronic Medical Records, HDFS (Hadoop Distributed File System), High Tech Industry, Identify Issues, Java, Machine Learning, Machine Tool, Microsoft Windows Azure, NoSQL, Production Systems, Python Programming/Scripting Language, Replication and Remote Mirroring, SQL (Structured Query Language), Scala Programming Language, Software Development, Software Engineering, Structured Data, Unstructured Data ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Leveraging Real time data in FSIs](https://www.wearedevelopers.com/videos/806-leveraging-real-time-data-in-fsis) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) - [NoSQL Data Modeling for Front-end Developers](https://www.wearedevelopers.com/videos/297-nosql-data-modeling-for-front-end-developers) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)