Associate Data Scientist

Birlasoft Inc
Noida, United States
11 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Data Analysis ARM Architecture Microsoft Azure Big Data Cloud Computing Cluster Analysis Computer Programming Continuous Integration Information Engineering Distributed Computing Environment
+26 more
Statistical Hypothesis Testing Python (Programming Language) Machine Learning NumPy Tensorflow Standard Sql Workflow Management Systems Data Processing Feature Engineering Data Ingestion Pytorch Large Language Models Snowflake Deep Learning Generative AI Pandas Containerization Data Lakes Scikit Learn Kubernetes Data Analytics Machine Learning Operations Software Version Control Data Pipelines Docker Databricks

Job description

We are seeking a highly skilled Data Scientist with strong expertise in Artificial Intelligence (AI), Machine Learning (ML), and MLOps. The ideal candidate will be responsible for building scalable predictive models, driving advanced analytics, and operationalizing ML models in production environments. This role requires a deep understanding of statistical modeling, predictive analytics, and Python-based data ecosystems, with exposure to modern platforms such as Databricks Mosaic AI and Snowflake Cortex being an added advantage., * Design, develop, and deploy machine learning and AI models for real-world business problems.

  • Perform advanced statistical analysis and build predictive models to derive actionable insights.
  • Develop and implement end-to-end ML pipelines, including data ingestion, feature engineering, model training, validation, and deployment.
  • Build and manage MLOps frameworks for continuous integration, delivery, monitoring, and model governance.
  • Work closely with data engineering teams to ensure robust and scalable data pipelines.
  • Conduct exploratory data analysis (EDA) and hypothesis testing to support data-driven decision-making.
  • Optimize model performance through hyperparameter tuning and advanced techniques.
  • Deploy and monitor models in production environments ensuring performance, reliability, and scalability.
  • Collaborate with cross-functional teams including business stakeholders, architects, and product owners.
  • Stay updated with the latest advancements in AI/ML, GenAI, and data science tools and frameworks.

Requirements

  • Strong programming expertise in Python (NumPy, Pandas, Scikit-learn, TensorFlow/PyTorch).
  • Hands-on experience in Machine Learning & AI algorithms (supervised, unsupervised, deep learning).
  • Expertise in statistical analysis, hypothesis testing, regression, classification, clustering, and forecasting models.
  • Experience with predictive modeling and advanced analytics techniques.
  • Solid understanding of MLOps practices including CI/CD pipelines, model versioning, monitoring, and deployment.
  • Experience working with large-scale datasets and distributed computing frameworks.
  • Strong knowledge of SQL and data manipulation techniques.
  • Familiarity with cloud platforms such as AWS, Azure, or GCP.

Good to Have

  • Exposure to Databricks (Mosaic AI, MLflow, Delta Lake).
  • Experience with Snowflake Cortex / Snowflake ML capabilities.
  • Understanding of Generative AI / LLM-based applications.
  • Experience in model explainability, fairness, and governance frameworks.
  • Knowledge of containerization tools like Docker and orchestration tools like Kubernetes.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:34 min

Maximizing execution memory effectively via python numpy broadcasting

Jodie Burchell · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:03 min

Accelerating pandas dataframes using cudf module plugins

Ankit Patel Ankit Patel · World Congress 2024

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

1:25 min

Replacing NumPy with cuPy for straightforward GPU acceleration

Paul Graham Paul Graham · World Congress 2025

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

Videos

See all

Related articles

See all