Data Scientist

LOCALHOST LLC
United States
10 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Compensation
$33,600.0
Working hours
Regular working hours
Job source

Tech stack

Amazon Web Services Artificial Neural Networks Big Data Cloud Computing Cluster Analysis Data Architecture Data Governance Extract Transform Load (ETL) Data Mining Data Visualization Data Warehousing Relational Databases
+18 more
Information Lifecycle Management Python (Programming Language) Logistic Regression Machine Learning Metadata Meta-Data Management Natural Language Processing NumPy Operational Data Store SciPy SQL Databases Data Processing Snowflake Pandas Data Lakes Data Management Text Analysis Data Pipelines

Job description

As a Data Science Engineer focusing on data pipelines, you will play a critical role in developing, deploying, and optimizing cloud based solutions. This position requires experience and proficiency with Python and AWS to drive business value and enhance user experiences across our platforms., * Directing the data mining, data gathering, and data processing in large volume; creating appropriate data models.

  • Exploring, promoting, and implementing semantic data capabilities through Natural Language Processing, text analysis and machine learning techniques.
  • Defining requirements and scope of data analyses; presenting and reporting possible business insights to colleagues/management using data visualization technologies.
  • Evaluating and conducting research on data model optimization and algorithms to improve effectiveness and accuracy on data analyses.
  • Determining the root cause of organizational problems and creating alternative solutions that resolve these problems.

Employee perks and benefits

  • 6 extra days off: 3x localhost days and 3x sick days
  • Referral bonus
  • Benefit plus budget
  • Financial contribution to Pension plan
  • Multisport Card
  • Education support (certificates, courses, trainings)
  • Physiotherapist sessions once a week in the office
  • Bonuses at every smashing life events
  • Flexible working arrangements
  • Transparent approach and communication
  • Supporting your ideas

Requirements

  • Proven experience with Python (NumPy, SciPy, Pandas, etc.).
  • Experience with AWS services, Cloud ELT/ETL, Snowflake, SQL.
  • Working knowledge and experience in practical applications of Machine Learning techniques such as Clustering, Logistic Regression, Random Forests, SVM or Neural Networks.
  • Strong knowledge of end-to-end data lifecycle across traditional data warehouses, relational databases, operational data stores, business intelligence reporting, as well as new concepts such as data fabrics, data mesh, and data lake house.
  • Strong understanding of data governance, metadata management, data quality, data modeling, and data architecture concepts.
  • Experience utilizing quantitative analysis/ data management (data design, data quality, metadata, governance, etc.).
  • Knowledge of data technology products and components for Big Data and Cloud (AWS, Data Lakes, and similar)
  • Ability to clearly communicate complex technical ideas, regardless of the technical capacity of the audience.
  • Knowledge of data management systems; ability to use, support and access facilities for searching, extracting and formatting data for further use.

Benefits & conditions

From 2800 Eur/month

Final agreement depends on your skill-set and experience. This is the subject of a partnership agreement during the selection process.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

4:54 min

Development history of scientific computation libraries and PyViz tools

Radovan Kavický · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

3:33 min

Refactoring data science workflows using Rapids QDF and Pandas

Paul Graham Paul Graham · LIVE

2:01 min

Executing remote data exploration and model training

Mingshen Sun Mingshen Sun · WWC 2024

Videos

See all

Related articles

See all