Data Scientist Engineer

International Technologies & Systems Corporation
Baltimore, MD, United States
2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$120,000.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Data Analysis Computer Vision Microsoft Azure Big Data Cloud Computing Data as a Services Python (Programming Language) Machine Learning Natural Language Processing NumPy
+14 more
Tensorflow SQL Databases Data Processing Pytorch Large Language Models Apache Spark Deep Learning Generative AI Pandas Pyspark Scikit Learn Information Technology Machine Learning Operations Databricks

Job description

This role involves leveraging advanced machine learning models and AI-driven solutions to address complex business problems., * AI/ML Model Development: Design and train machine learning models using various algorithms, including deep learning, NLP, and computer vision.

  • Databricks Orchestration: Build and optimize end-to-end AI/ML pipelines on Databricks, utilizing Unity Catalog for governance and MLflow for experiment tracking.
  • Generative AI & LLMs: Implement advanced AI patterns such as Retrieval-Augmented Generation (RAG) and fine-tune pre-trained models for specific enterprise tasks.
  • Python Expertise: Write production-quality, idiomatic PySpark and Python code that leverages Spark’s distributed nature.
  • Collaboration: Partner with Engineering and Product teams to translate business problems into scalable analytical solutions.
  • Insight Extraction: Perform exploratory data analysis (EDA) and extract meaningful insights from massive, complex datasets to drive strategic decisions.

Requirements

Do you have experience in Technical Proficiency?, Do you have a Master’s degree?, * Education: MS or PhD in a quantitative field such as Computer Science, Statistics, or Math.

  • Experience: 5+ years of hands-on experience in data science or AI engineering in high-growth environments., * Technical Proficiency: Expert-level Python (pandas, NumPy, scikit-learn, PySpark).
  • Extensive experience with Apache Spark for large-scale data processing.
  • Proficiency in SQL for data manipulation and querying in Lakehouse environments.
  • AI Foundations: Strong understanding of statistics, probability, and advanced ML lifecycle management (MLOps).

Preferred Skills

  • Experience with Deep Learning frameworks (TensorFlow, PyTorch).
  • Familiarity with Cloud Platforms (AWS, Azure, or GCP) and their native data services.
  • Databricks certifications, such as Databricks Certified Machine Learning Professional.

Benefits & conditions

Pulled from the full job description

  • Health insurance
  • Paid time off
  • Employee assistance program

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · WWC 2024

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

6:58 min

Analyzing production code coverage data using pandas

Markus Harrer Markus Harrer · WWC 2021

Videos

See all

Related articles

See all