> Markdown version of [/jobs/ext/2866326-senior-data-scientist](https://www.wearedevelopers.com/jobs/ext/2866326-senior-data-scientist). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior Data Scientist - **Company:** Managed Markets Insight & Technology, LLC - **Location:** Topeka, KS, United States - **Experience:** Expert - **Salary:** $170,000.0 - $215,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Algorithm Design, Amazon Web Services, Amazon S3, Data Retrieval, Distributed Systems, Graph Database, Apache Hadoop, Python (Programming Language), Machine Learning, Natural Language Processing, NumPy, Open Source Technology, Data Logging, Pytorch, Large Language Models, Snowflake, Multi-Agent Systems, Prompt Engineering, Apache Spark, Deep Learning, Generative AI, Backend, Fastapi, Pandas, Core Data, Kubernetes, Information Technology, Low Latency, Data Management, Api Gateway, Restful APIs, Serverless Computing, Docker, Databricks, Microservices - **Published:** September 12, 2026 - **Apply:** https://www.techcareers.com/job.asp?id=3387058616&tx=JT10193UTI&pt=1&aff=0B19D771-A501-4A5E-8338-2A822B784D54&utm_source=Job%20Feed&utm_medium=textkernel&utm_campaign=DE&utm_term=0B19D771-A501-4A5E-8338-2A822B784D54 ## About the Role * 5+ years of experience developing AI / ML applications and data driven solutions, preferably in regulated industries (pharma, legal, financial services, or energy) * Graduate degree in Computer Science, Engineering, Statistics or a related quantitative discipline, or equivalent work experience * Substantial depth and breadth in NLP, Deep Learning, Generative AI, LLMs, and other state of the art AI / ML techniques * Deep experience with LLM orchestration frameworks such as LangChain, LlamaIndex, or similar libraries * Expert-level knowledge of LLM APIs (OpenAI, Anthropic Claude) and open-source models (Llama, Mistral) * Deep understanding of CS fundamentals, computational complexity and algorithm design * Experience with building large-scale distributed systems in an agile environment and the ability to build quick prototypes * Excellent knowledge of Python and core data science and AI libraries including Pandas, NumPy, PyTorch, and similar * Experience building or utilizing Model Context Protocol (MCP) servers to bridge models with data tools * Strong background in scalable backend environments (Docker, Kubernetes, AWS/GCP) * Experience moving AI from prototype to production-grade services with monitoring, logging, and rate-limiting * Ability to independently conduct research and develop appropriate algorithmic solutions to complex business problems * Experience mentoring junior team members * Excellent problem solving and communication skills, * Knowledge of the healthcare/pharma domain and experience with applying AI to healthcare data * Experience with AWS, especially ECS, Bedrock, API Gateway, SageMaker, serverless compute and storage such as S3 and Snowflake * Proficiency with vector databases such as Pinecone, Qdrant, or similar for high-performance retrieval * Experience with RAG patterns, prompt engineering, model fine tuning, and knowledge graphs * Experience with unstructured document processing (legal document analysis, contract management, data retrieval) * Experience with Big Data tools like Apache Spark, Hadoop, or Databricks ## Description Our dedicated Data Science team is at the forefront of revolutionizing pharma intelligence and how patients gain access to life-saving therapies. Armed with cutting-edge technology and a passion for innovation, we leverage the vast landscape of data to extract actionable insights that drive informed decision making. Our unique collaborative approach fosters a dynamic synergy between data science and product development. Our deep expertise in machine learning, artificial intelligence, large language models, and generative AI, combined with our domain knowledge, enables us to deliver comprehensive, production-grade AI solutions that empower our clients to stay ahead in a rapidly evolving industry., In this role as a Senior Data Scientist, you will: * Design and deploy production-ready AI systems that leverage LLMs and advanced ML techniques to solve complex business problems across pharma intelligence * Build and maintain multi-agent systems and agentic orchestration workflows using frameworks like LangChain, LangGraph, or AutoGen to execute autonomous tasks * Develop and optimize Retrieval-Augmented Generation (RAG) pipelines, ensuring high-fidelity context retrieval and vector database management * Implement and extend MCP (Model Context Protocol) servers to allow LLMs to interact safely and efficiently with local and remote data sources * Architect robust, scalable APIs and microservices to serve AI features to end-users with low latency (FastAPI or similar) * Collaborate with product partners and other scientists to identify new opportunities to apply AI / ML to our content and products * Conduct research and identify AI / ML algorithms and methods to solve specific business problems, and deliver these algorithms as microservices in collaboration with content and product engineering teams * Implement rigorous testing and evaluation frameworks for LLM outputs to ensure prompt stability, prevent regressions, and manage hallucination risks * Contribute towards the common data science platform * Stay up-to-date, constantly learning about advances in the field, and deliver periodic presentations to internal teams on these developments * All other duties as assigned ## Related Videos - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [Advanced Typing in TypeScript](https://www.wearedevelopers.com/videos/496-advanced-typing-in-typescript) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [MLOps – What’s the deal behind it?](https://www.wearedevelopers.com/magazine/125-mlops-what-s-the-deal-behind-it) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development)