Data Scientist

Tekshapers Inc
Raritan, NJ, United States
11 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure Big Data BigQuery Cloud Computing Cloud Computing Security Cloud Storage Code Review Continuous Integration Data Cleansing
+32 more
Data Governance Extract Transform Load (ETL) Data Visualization Distributed Computing Environment Data Flow Control Identity and Access Management Python (Programming Language) Machine Learning Performance Tuning Systems Development Life Cycle Tensorflow Software Engineering SQL Databases Data Processing Google Cloud Pytorch Large Language Models Generative AI Containerization Scikit Learn Kubernetes Information Technology Data Lineage Google Cloud Functions Graphql Machine Learning Operations Api Design Software Version Control Data Pipelines GXP Docker Microservices

Job description

Summary: We are seeking a seasoned Senior Data Scientist with overall 10-12 yrs and at least 5-7 years of hands-on experience in developing GenAI/machine learning models and deploying them in a cloud environment, preferably on Google Cloud Platform (Google Cloud Platform). The ideal candidate will design microservice-based solutions, containerize deployments (e.g., GKE), and drive end-to-end SDLC practices. Experience in the pharma domain is a strong advantage., * Lead end-to-end development of GenAI/ML models: problem framing, data preparation, model selection, training, evaluation, and iteration.

  • Architect and implement microservice-based AI solutions and deploy them in containerized environments (preferably GKE); define APIs and data contracts.
  • Incorporate and operationalize defined ML pipelines with MLOps practices: model versioning, feature stores, experiment tracking, CI/CD for ML, monitoring, and rollback strategies.
  • Leverage Google Cloud Platform offerings (Vertex AI, BigQuery, Dataflow, Cloud Storage, Pub/Sub, Cloud Run, GKE, etc.) to design scalable AI solutions and efficient data workflows.
  • Knowledge of Retrieval-Augmented Generation (RAG) concepts and processes
  • Proficiency with Google Cloud Platform (Google Cloud Platform) and its AI/ML offerings (e.g., Vertex AI, BigQuery, Dataflow, Cloud Storage, GKE).
  • Deploy, monitor, and maintain models in production; implement observability (logs, metrics, tracing), cost optimization, and performance tuning.
  • Collaborate with cross-functional teams (data engineers, software engineers, product, regulatory/compliance, analytics) to translate business needs into robust ML solutions.
  • Uphold SDLC standards: requirements gathering, design, development, testing, deployment, maintenance, and documentation; promote reusable patterns and best practices.
  • Mentor and guide junior scientists; contribute to code reviews, standards, and knowledge sharing.
  • Stay current with GenAI advancements and evaluate new tools/approaches; produce reproducible experiments and artifacts.

Requirements

  • Overall 10-12yrs and Minimum 5-7 years of hands-on experience developing GenAI/ML models and deploying them in a cloud environment.
  • Proficiency with Google Cloud Platform (Google Cloud Platform) and its AI/ML offerings (e.g., Vertex AI, BigQuery, Dataflow, Cloud Storage, Pub/Sub, Cloud Run, GKE).
  • Must have experience working with any agentic framework
  • Knowledge of Retrieval-Augmented Generation (RAG) concepts and processes
  • Strong software engineering skills: Python (primary), experience with ML frameworks (TensorFlow, PyTorch, scikit-learn), and API development (REST/GraphQL).
  • Experience designing and deploying microservices architectures and containerized solutions (Docker, Kubernetes; preference for GKE).
  • Solid experience in MLOps: model versioning, experiments, automated training, feature stores, model registries, monitoring, and governance.
  • Data processing and analytics expertise: SQL, data pipelines, ETL/ELT concepts, data quality, and data visualization support.
  • Excellent problem-solving, communication, and collaboration skills; ability to work with cross-disciplinary teams.
  • Understanding of cloud security concepts, IAM, and basic principles of data privacy and compliance.
  • Demonstrated ability to translate business problems into scalable ML solutions and to communicate technical concepts to non-technical stakeholders., * Experience in the pharmaceutical/pharma domain or regulated industries; familiarity with GxP, or similar data governance requirements.
  • Exposure to other cloud providers (AWS/Azure) is a plus, but a strong preference for Google Cloud Platform.
  • Experience with distributed training, large-scale data processing, and fine-tuning of large language models.
  • Knowledge of privacy-preserving ML methods (differential privacy, synthetic data) and data lineage tools.

Education:

  • Minimum qualification: Graduate degree in Information Technology.
  • Preferred: Higher education (e.g., Master’s degree in Computer Science, Information Technology, Data Science, or a related field) or relevant professional degrees/certifications.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:52 min

Generating APIs with the Neo4j GraphQL library

William Lyon · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

3:13 min

Core components of the internal Optimize ecosystem

Dominik Schneider Dominik Schneider · World Congress 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

4:33 min

Overview of the GraphQL API query language

William Lyon · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all