AI/LLM Engineer

NTT
Pittsburgh, PA, United States
6 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
1 year minimum
Compensation
$114,400.0 - $128,960.0
Working hours
Regular working hours
Job source

Tech stack

JavaScript (Programming Language) Application Programming Interfaces (APIs) Artificial Intelligence Amazon Web Services Microsoft Azure Cloud Computing Cloud Computing Security Program Optimization DevOps Python (Programming Language) Machine Learning Search Technologies
+21 more
Software Engineering Systems Integration Enterprise Software Applications Delivery Pipeline Large Language Models Multi-Agent Systems Prompt Engineering IT Architecture Model Validation Software Application Programming Generative AI Backend Event Driven Architecture Containerization AI Platforms Information Technology Machine Learning Operations Virtual Agents Restful APIs Docker Microservices

Job description

We are seeking an experienced Senior AI/LLM Engineer to design, develop, and deploy enterprise-grade Generative AI solutions that leverage Large Language Models (LLMs) to solve complex business challenges. The ideal candidate will have strong expertise in Python development, Retrieval-Augmented Generation (RAG), AI orchestration frameworks, and cloud-native AI architectures. This role will work closely with product owners, solution architects, data engineers, and business stakeholders to build scalable, secure, and production-ready AI applications powered by OpenAI and other leading foundation models.

Key Responsibilities AI Solution Design & Development

  • Design, develop, and deploy enterprise-scale Generative AI applications using modern LLM technologies.
  • Build production-grade backend services using Python and modern software engineering practices.
  • Develop scalable AI architectures utilizing OpenAI, Anthropic Claude, Gemini, Llama, or similar foundation models.
  • Design and implement Retrieval-Augmented Generation (RAG) pipelines using vector databases and semantic search capabilities.
  • Develop intelligent multi-agent AI systems capable of orchestrating complex business workflows.

AI Architecture & Integration

  • Design AI solution architectures that are scalable, secure, maintainable, and aligned with enterprise standards.
  • Integrate AI capabilities into existing enterprise applications, APIs, and business workflows.
  • Develop and consume REST APIs, microservices, and event-driven services for AI applications.
  • Implement AI orchestration frameworks such as LangChain and related agent frameworks.

Prompt Engineering & Model Optimization

  • Develop and optimize prompts for improved accuracy, reasoning, and business outcomes.
  • Evaluate LLM performance and implement techniques to improve response quality.
  • Establish AI governance, model evaluation, and responsible AI best practices.
  • Monitor AI application performance and continuously optimize latency, cost, and quality.

Cloud & Enterprise AI

  • Build cloud-native AI solutions using Azure or AWS AI services.
  • Implement Azure AI Search and vector search capabilities.
  • Design secure enterprise AI applications following cloud security and governance standards.
  • Collaborate with DevOps teams to deploy AI solutions using CI/CD pipelines.
  • Stakeholder Collaboration
  • Partner with business stakeholders to understand AI use cases and translate them into scalable technical solutions.
  • Present architecture decisions, solution approaches, and AI strategies to both technical and non-technical audiences.
  • Mentor junior engineers and contribute to AI engineering best practices across the organization.

Requirements

  • Bachelor’s or Master’s degree in Computer Science, Artificial Intelligence, Software Engineering, or a related field.
  • 6+ years of software engineering or machine learning engineering experience.
  • Minimum 3+ years of hands-on experience building Artificial Intelligence and Generative AI solutions.
  • 3 to 5 years of strong hands-on experience developing production applications using Python.
  • 1 to 3 years of experience building applications using OpenAI (preferred), Anthropic Claude, Gemini, Llama, or similar LLM platforms.
  • 3+ years of strong experience implementing Retrieval-Augmented Generation (RAG) architectures.
  • 1 to 3 years of experience integrating vector databases and semantic search solutions.
  • 1 to 3 years of strong understanding of LangChain and AI orchestration frameworks.
  • Experience designing multi-agent AI architectures.
  • Strong knowledge of prompt engineering, model evaluation, AI governance, and Responsible AI principles.
  • Experience building REST APIs, microservices, and event-driven architectures.
  • Experience with Azure or AWS cloud platforms.
  • Strong understanding of scalable enterprise application architecture.
  • Excellent analytical, problem-solving, and communication skills.

Required Technical Skills

  • Python
  • JavaScript
  • Generative AI
  • Large Language Models (LLMs)
  • OpenAI (Preferred)
  • Retrieval-Augmented Generation (RAG)
  • Multi-Agent AI Orchestration
  • LangChain
  • Langfuse
  • AI Solution Architecture
  • Azure AI Search
  • Vector Databases
  • Prompt Engineering
  • REST APIs
  • Microservices
  • Event-Driven Architecture
  • Azure or AWS Cloud Services

Preferred Qualifications

  • Experience with LangGraph, AutoGen, CrewAI, Semantic Kernel, or similar AI agent frameworks.
  • Experience with vector databases such as Pinecone, Weaviate, ChromaDB, Qdrant, Milvus, or Azure AI Search.
  • Experience with containerization technologies such as Docker and Kubernetes.
  • Knowledge of CI/CD pipelines and MLOps practices.
  • Experience with AI observability and monitoring platforms.
  • Familiarity with enterprise security, compliance, and Responsible AI frameworks.
  • Experience working in Agile/Scrum environments.

Nice to Have

  • Experience developing enterprise copilots or AI assistants.
  • Experience integrating AI into enterprise SaaS platforms.
  • Knowledge of AI governance, security, and compliance standards.
  • Experience optimizing LLM inference performance and AI operational costs.

Benefits & conditions

NTT DATA provides a reasonable range of compensation for U.S.-based positions. The starting pay range for this remote role is $55 to $62/hour. This range reflects the minimum and maximum target compensation for the position across all US locations. Actual compensation will depend on a number of factors, including the candidate’s actual work location, relevant experience, technical skills, and other qualifications. This position is eligible for company benefits including participation in medical, dental, and vision insurance, flexible spending or health savings account, and AD&D insurance, employee assistance, participation in a 401k program, and additional voluntary or legally-required benefits #indist #li-northamerica

About the company

NTT DATA is a $30 billion trusted global innovator of business and technology services. We serve 75% of the Fortune Global 100 and are committed to helping clients innovate, optimize and transform for long term success. As a Global Top Employer, we have diverse experts in more than 50 countries and a robust partner ecosystem of established and start-up companies. Our services include business and technology consulting, data and artificial intelligence, industry solutions, as well as the development, implementation and management of applications, infrastructure and connectivity. We are one of the leading providers of digital and AI infrastructure in the world. NTT DATA is a part of NTT Group, which invests over $3.6 billion each year in R&D to help organizations and society move confidently and sustainably into the digital future. Visit us at us.nttdata.com

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:08 min

Essential engineering roles in the generative AI space

Mary Grygleski Mary Grygleski · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

1:52 min

Structuring and scaling the backend engineering team

Stefan Lingler Stefan Lingler +1 · Coffee With Developers

5:30 min

Building components of a real-world LLM lifecycle

Maxim Salnikov Maxim Salnikov · LIVE

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · World Congress 2026 Europe

Videos

See all

Related articles

See all