Data Scientist II (Adbl175)

Amazon.com, Inc.
Newark, NJ, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$169,550.0 - $207,500.0
Working hours
Regular working hours
Job source

Tech stack

Amazon S3 Data Analysis Artificial Neural Networks Big Data Databases Computer Engineering Extract Transform Load (ETL) Genetic Algorithm Python (Programming Language) Machine Learning Natural Language Processing Network Model
+9 more
Recommender Systems SQL Databases Data Processing Large Language Models Deep Learning Model Validation Information Technology Data Analytics Xgboost

Job description

Duties: Independently own, design, and implement scalable and reliable solutions to support or automate decision making throughout the business. Apply a range of data science techniques and tools combined with subject matter expertise to solve difficult business problems and cases in which the approach is unclear. Acquire data by building the necessary SQL/ETL queries. Import processes through various company specific interfaces for accessing RedShift, and S3/edX storage systems. Deliver artifacts on medium size projects that affect important business decisions. Build relationships with stakeholders and counterparts, and communicate model outputs, observations, and key performance indicators (KPIs) to the management to develop sustainable and consumable products and product features. Explore and analyze data by inspecting univariate distributions and multivariate interactions, constructing appropriate transformations, and tracking down the source and meaning of anomalies.

Requirements

Build production-ready models using statistical modeling, mathematical modeling, econometric modeling, machine learning algorithms, network modeling, social network modeling, natural language processing, large language models and/or genetic algorithms. Validate models against alternative approaches, expected and observed outcome, and other business defined key performance indicators. Implement models that comply with evaluations of the computational demands, accuracy, and reliability of the relevant ETL processes at various stages of production. Position reports to Newark, NJ office; however, telecommuting from a home office may be allowed.

Requirements: Requires a Master’s degree in Statistics, Computer Science, Computer Engineering, Data Science, Machine Learning, Applied Math, Operations Research, or a related field plus two (2) years of experience as a Data Scientist or other occupation involving data processing and predictive Machine Learning modeling at scale. Experience may be gained concurrently and must include:

Two (2) years in each of the following:

  • Utilizing specialized modelling software including Python or R

  • Building statistical models and machine learning models using large datasets from multiple resources

  • Building non-linear models including Neural Nets, Deep Learning, or Gradient Boosting.

One (1) year in each of the following:

  • Building production-ready solutions or applications relying on Large Language Models (LLM), accessed programmatically and beyond just prompting

  • Evaluating LLM results at scale or fine-tuning LLMs

  • Building production-ready recommendation systems

  • Using database technologies including SQL or ETL.

Alternatively, will accept a Bachelor’s degree and five (5) years of experience., Requirements: Requires a Master’s degree in Statistics, Computer Science, Computer Engineering, Data Science, Machine Learning, Applied Math, Operations Research, or a related field plus two (2) years of experience as a Data Scientist or other occupation involving data processing and predictive Machine Learning modeling at scale. Experience may be gained concurrently and must include:

Two (2) years in each of the following:

  • Utilizing specialized modelling software including Python or R

  • Building statistical models and machine learning models using large datasets from multiple resources

  • Building non-linear models including Neural Nets, Deep Learning, or Gradient Boosting.

One (1) year in each of the following:

  • Building production-ready solutions or applications relying on Large Language Models (LLM), accessed programmatically and beyond just prompting

  • Evaluating LLM results at scale or fine-tuning LLMs

  • Building production-ready recommendation systems

  • Using database technologies including SQL or ETL.

Alternatively, will accept a Bachelor’s degree and five (5) years of experience.

Benefits & conditions

Salary: $169,550 - 207,500 /year. Multiple positions. Apply online: www.amazon.jobs Job Code: ADBL175.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on juju.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

1:09 min

Configuring synthetic data for safe interactive programming

Mingshen Sun Mingshen Sun · WWC 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:01 min

Managing application isolation via pluggable database models

Wei Hu Wei Hu · WWC 2022

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

Videos

See all

Related articles

See all