Data Scientist (R-19646)

Dun & Bradstreet
Short Hills, United States
3 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Experienced
Experience required
4 years minimum
Working hours
Regular working hours
Job source

Tech stack

Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Data Analysis Software Applications Automated Storage and Retrieval Systems Microsoft Azure Program Optimization Computer Programming Continuous Integration Data Validation Distributed Data Store
+16 more
Distributed Systems Iterative and Incremental Development Python (Programming Language) Machine Learning Search Technologies Software Engineering Delivery Pipeline Large Language Models Model Validation Generative AI Containerization Pyspark Machine Learning Operations Api Design GPT Docker

Job description

The Role: We are looking for an experienced AI Engineer to design, build, and operationalize AI driven solutions for our global Analytics organization. The ideal candidate will have strong hands on expertise in Python, PySpark, agentic workflow development, and modern GenAI frameworks, with experience building scalable applications using LLMs, retrieval systems, and automation pipelines. You will work closely with data scientists, MLOps engineers, and business stakeholders to build intelligent, production grade systems that power, 2. Agent Development & Architecture

  • Build agentic workflows using LangChain/LangGraph and similar frameworks.
  • Develop autonomous agents for data validation, reporting, document processing, and domain workflows.
  • Deploy scalable, resilient agent pipelines with monitoring and evaluation. 3. GenAI Application Engineering* Develop GenAI applications using models like GPT, Gemini, and LLaMA.
  • Implement RAG, vector search, prompt orchestration, and model evaluation.
  • Partner with data scientists to productionize POCs. 4. Data & Platform Engineering* Build distributed data pipelines (Python, PySpark).
  • Develop APIs, SDKs, and integration layers for AI-powered applications.
  • Optimize systems for performance and scalability across cloud/hybrid environments. 5. MLOps / LLMOps
  • Contribute to CI/CD workflows for AI models-deployment, testing, monitoring.
  • Implement governance, guardrails, and reusable GenAI frameworks. 6. Collaboration & Stakeholder Engagement
  • Work with analytics, product, and engineering teams to define and deliver AI solutions.
  • Participate in architecture reviews and iterative development cycles.
  • Support knowledge sharing and internal GenAI capability building.

Requirements

  • 5-8 years of experience in AI/ML engineering, data science, or software engineering, with at least 4 years focused on GenAI.
  • Strong programming expertise in Python, distributed computing using PySpark, and API development.
  • Hands on experience with LLM frameworks (LangChain, LangGraph, Transformers, OpenAI/Vertex/Bedrock SDKs).
  • Experience developing AI agents, retrieval pipelines, tool calling structures, or autonomous task orchestration.
  • Solid understanding of GenAI concepts: prompting, embeddings, RAG, evaluation metrics, hallucination identification, model selection, fine tuning, context engineering.
  • Experience with cloud platforms (Azure/AWS/GCP), containerization (Docker), and CI/CD pipelines for ML/AI.
  • Strong problem solving, system design thinking, and ability to translate business needs into scalable AI solutions.
  • Excellent verbal, written communication and presentation skills.

Good to Have

  • Experience in workflow automation and building reusable AI components.
  • Background in analytics, statistical models, or enterprise data products.
  • Experience with MLOps / LLMOps tooling

About the company

Shape the Future with Dun & Bradstreet

At Dun & Bradstreet, we believe data has the power to create a better tomorrow. As a global leader in business decisioning data and analytics, we help companies worldwide grow, manage risk, and innovate. For over 180 years, businesses have trusted us to turn uncertainty into opportunity. We’re a diverse, global team that values creativity, collaboration, and bold ideas. Are you ready to make an impact and help shape what’s next? Join us! Explore opportunities at dnb.com/careers.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

Videos

See all

Related articles

See all