Data Scientist

Sidram Technologies
Fort Wayne, IN, United States
1 day ago
Apply on www.careerjet.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Compensation
$50,000.0
Working hours
Regular working hours

Tech stack

Artificial Intelligence Amazon Web Services Microsoft Azure Supervisory Control and Data Acquisition (SCADA) Python (Programming Language) Machine Learning Operational Data Store Tensorflow Search Technologies SQL Databases Pinecone Cloud Platform System
+11 more
Pytorch LangChain Retrieval-Augmented Generation Large Language Models Prompt Engineering Vector Embeddings Llamaindex Generative AI Weaviate Milvus GPT

Job description

We are seeking a highly skilled and motivated Data Scientist with a strong background in Large Language Models (LLMs) and Retrieval-Augmented Generation (RAG) systems. The ideal candidate has direct experience applying AI within the utility, energy, or industrial sector. You will lead the design, development, and deployment of intelligent systems that translate unstructured operational data (technical manuals, maintenance logs, field reports) into actionable insights. Key Responsibilities Develop RAG Pipelines :: Design and optimize Retrieval-Augmented Generation (RAG) models to enhance LLM-driven troubleshooting and field operations. AI Model Development :: Fine-tune and evaluate LLMs (e.g., GPT-4, Llama 3) for task-specific reasoning and accuracy. Knowledge Integration :: Build high-quality Python code to integrate vector databases (e.g., Pinecone, Milvus, Weaviate) to store and retrieve technical utility documents efficiently. Operational Intelligence :: Apply LLM solutions to analyze unstructured utility data, such as asset management logs, SCADA alerts, and maintenance records. Deployment :: Work with software engineers to deploy and scale models in cloud environments (Azure/AWS). Collaboration: Work cross-functionally with field operations, engineering teams, and stakeholders to understand data needs and deliver solutions.

Requirements

Experience: Professional experience in Data Science/Machine Learning, with at least 1 year focusing specifically on LLMs and Generative AI. Technical Skills: Advanced proficiency in Python (PyTorch/TensorFlow, LangChain, LlamaIndex) and SQL. Domain Knowledge: Proven experience in the utility, energy, oil & gas, or heavy industrial sectors. RAG Expertise: Hands-on experience with vector embeddings, semantic search, and prompt engineering

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.careerjet.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · World Congress 2023

4:32 min

Evaluating open-source and cloud-based vector database vendors

Erik Bamberg · LIVE

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

1:06 min

Compiling PyTorch environments for advanced time forecasting

Christoph Lohrmann Christoph Lohrmann +1 · World Congress 2026 Europe

3:04 min

Database evolution and the funding behind vector databases

Erik Bamberg · LIVE

Videos

See all

Related articles

See all