Data Engineer
Cornerstone OnDemand
Santa Monica, CA, United States
13 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
3 years minimum
Working hours
Regular working hours
Job source
Tech stack
Sql Data Warehouse
Java (Programming Language)
Airflow
Amazon Web Services
Microsoft Azure
Bash Shell
Big Data
C++ (Programming Language)
Cloud Computing
Data Control
Data Infrastructure
Extract Transform Load (ETL)
+28 more
Data Systems
Data Warehousing
Relational Databases
Database Design
Distributed Systems
Apache Hadoop
MapReduce
Apache Hive
Python (Programming Language)
Machine Learning
MongoDB
Neo4j
NoSQL
Object-Oriented Software Development
Raw Data
Redis
Standard Sql
Scala (Programming Language)
Workflow Management Systems
Scripting
Google Cloud
Data Ingestion
Snowflake
Apache Spark
Cassandra
Presto
Data Pipelines
Databricks
Job description
- Design, build and maintain batch or real-time data pipelines in production.
- Maintain and optimize the data infrastructure required for accurate extraction, transformation, and loading of data from a wide variety of data sources.
- Develop ETL (extract, transform, load) processes to help extract and manipulate data from multiple sources.
- Automate data workflows such as data ingestion, aggregation, and ETL processing.
- Prepare raw data in Data Warehouses into a consumable dataset for both technical and non-technical stakeholders.
- Partner with data scientists and functional leaders in sales, marketing, and product to deploy machine learning models in production.
- Build, maintain, and deploy data products for analytics and data science teams on cloud platforms (e.g. AWS, Azure, GCP).
- Ensure data accuracy, integrity, privacy, security, and compliance through quality control procedures.
- Monitor data systems performance and implement optimization strategies.
- Leverage data controls to maintain data privacy, security, compliance, and quality for allocated areas of ownership.
Requirements
- 3+ years of SQL skills and experience with relational databases and database design.
- Experience working with cloud Data Warehouse solutions - Databricks, Apache Spark
- Experience working with data ingestion tools such as Fivetran, stitch, or Matillion.
- Working knowledge of Cloud-based solutions (e.g. AWS, Azure, GCP).
- Experience building and deploying machine learning models in production.
- Strong proficiency in object-oriented languages: Python, Java, C++, Scala.
- Strong proficiency in scripting languages like Bash.
- Strong proficiency in data pipeline and workflow management tools (e.g., Airflow).
- Strong project management and organizational skills.
- Excellent problem-solving, communication, and organizational skills.
- Proven ability to work independently and with a team.
Extra dose of awesome if you have…
- Good understanding of NoSQL databases like CrateDB, Redis, Cassandra, MongoDB, or Neo4j.
- Experience with working on large data sets and distributed computing (e.g. Hive/Hadoop/Spark/Presto/MapReduce).
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.techcareers.comGood distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
BB
Benedikt Bischof
about 4 years ago
DS
Dhannush Subramani
Top Big Data Technologies That You Need to Know
about 4 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago
EM
Eli McGarvie
Data Engineer Salary UK
about 3 years ago
CH
Chris Heilmann
Dev Digest 120 - Apple and peers
about 2 years ago
AJ
Austin Joy
What Are The Top Skills Required For Azure Developers?
over 4 years ago