Architect Associate AWS Certified Data Analytics Specialty AWS Certified Developer Associate
Newt Global
United States
2 days ago
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source
Tech stack
Airflow
Amazon Web Services
Amazon Elastic Compute Cloud
Amazon S3
Apache HTTP Server
Big Data
Cloud Computing
Continuous Integration
Data Governance
Data Transformation
Data Security
Data Stores
+40 more
Data Systems
Data Warehousing
Database Queries
Distributed Systems
Github
Apache Hive
Identity and Access Management
Subnetting
Python (Programming Language)
Routing
Performance Tuning
Query Optimization
Shell Script
SQL Databases
Workflow Management Systems
Data Logging
S3 Bucket
Data Processing
Data Storage Technologies
Data Ingestion
Delivery Pipeline
Apache Spark
Boto3
Amazon Virtual Private Cloud (VPC)
AWS ECS
Git
Containerization
Data Lakes
Pyspark
Gitlab-ci
Kubernetes
Apache Flink
AWS Glue
Data Analytics
Apache Kafka
Presto
Software Version Control
Data Pipelines
Amazon Elastic Mapreduce (EMR)
Programming Languages
Job description
- AWS Solution Design Implementation Design develop and deploy scalable and costeffective data solutions on AWS leveraging services such as S3 for data lakes EC2 EMR Glue Athena Lambda Redshift and Kinesis
- Data Pipeline Development Build and maintain robust ETLELT data pipelines using PySpark for data ingestion transformation and loading into various data stores including those utilizing open table formats like Iceberg
- Big Data Processing Develop and optimize big data processing jobs using PySpark on AWS EMR or AWS Glue handling large datasets efficiently and integrating with Iceberg table formats
- Data Warehousing Design implement and manage data warehousing solutions including schema design data modeling and query optimization with a focus on Hive and modern data lake table formats like Iceberg for historical data and analytical queries
- Cloud Infrastructure Networking Implement secure and robust cloud infrastructure components including VPCs subnets routing and security groups to ensure proper connectivity and isolation for data solutions
- Containerized Workloads Design deploy and manage containerized data processing applications on Amazon Elastic Kubernetes Service EKSPerformance Tuning Optimization Optimize AWS resources and big data applications Spark Hive Iceberg for performance cost and efficiency
- Data Governance Security Implement best practices for data security access control and compliance within AWS including IAM policies S3 bucket policies and encryption
- Monitoring Troubleshooting Set up monitoring ing and logging for data pipelines and AWS infrastructure troubleshoot and resolve issues promptly
- Automation Develop and maintain automation scripts using Python and shell scripting for infrastructure provisioning deployment and operational tasks
- Collaboration Work closely with data scientists analysts and other engineering teams to understand data requirements and deliver reliable data solutions
Requirements
- AWS Certification Hold at least one AWS certification eg AWS Certified Solutions Architect Associate AWS Certified Data Analytics Specialty AWS Certified Developer Associate
- AWS Services Expertise Handson experience with key AWS services for data processing and storage including
- Storage S3 for data lakes EC2
- Data Processing EMR Glue Athena Lambda
- Networking VPC Subnets Routing Security Groups
- Containerization EKS
- Big Data Processing Strong proficiency in PySpark for developing complex data transformations and analytics
- Data Lake Table Formats Practical experience with Apache Iceberg for managing and querying data lakes
- Data Warehousing Indepth knowledge and practical experience with Apache Hive for data storage querying and schema management
- Programming Languages
- Python Expert level proficiency in Python for scripting data manipulation and AWS automation Boto3
- Shell Scripting Proficient in shell scripting for automation and operational tasks
- Database SQL Strong SQL skills for data querying and manipulation
- Data Concepts Solid understanding of ETLELT processes data modeling distributed computing and data governance
Good to Have Skills
- Containerization Orchestration Experience with Kubernetes for deploying and managing containerized applications
- CICD Experience with CICD tools and practices eg AWS CodePipeline GitHub Actions GitLab CI for automating deployment of data solutions
- Orchestration Experience with workflow orchestration tools like Apache Airflow
- Version Control Proficient in using Git for source code management
- Other Big Data Technologies Exposure to other big data technologies like Apache Kafka Flink or Presto
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Apply on www.dice.com
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
DS
Dhannush Subramani
about 4 years ago
AJ
Austin Joy
What Are The Top Skills Required For Azure Developers?
over 4 years ago
BB
Benedikt Bischof
Making Data Warehouses Fast: A Developer’s Story
about 4 years ago
ER
Erin Rifkin
Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud
over 1 year ago
LM
Luis Minvielle
7 Cloud Computing Trends Coming in 2025 for Developers
over 2 years ago
EM
Eli McGarvie
Highest Paying Tech Companies for Developers
over 3 years ago