> Markdown version of [/jobs/ext/516292-artificial-intelligence-engineer](https://www.wearedevelopers.com/jobs/ext/516292-artificial-intelligence-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Artificial Intelligence Engineer - **Company:** The Perfect Child LLC - **Location:** United States (Remote available) - **Salary:** $100,000.0 - $160,000.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Amazon Elastic Compute Cloud, Databases, Extract Transform Load (ETL), Django Web Framework, Python (Programming Language), Language Modeling, Open Source Technology, Performance Tuning, Search Technologies, Speech Recognition, Data Logging, Application Enhancement Tool, Flask (Web Framework), Large Language Models, Indexer, Fastapi, Kubernetes, Machine Learning Operations, Speech Synthesis, Data Pipelines, Docker - **Published:** June 13, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=afd305fc8641cc0e ## About the Role Do you have experience in System performance optimization?, * AI / ML / LLM * Hands-on deployment of open-source models (fine-tuned or instruct-tuned models a plus). * Strong understanding of vector search, embeddings, context window strategies, and RAG best practices. * Experience building agent architectures with structured function/tool calling. * Familiarity with MCP, LangGraph, LlamaIndex, LangChain, or similar orchestration frameworks. Software & Systems * Strong Python experience; familiar with Django / FastAPI / Flask or similar frameworks. * Experience building data pipelines (ETL/ELT, semantic chunking, scheduled indexing). * Experience deploying AI systems: Docker, AWS EC2 / ECS / Lambda, GPU instances, or local inference stacks. * Comfort with monitoring, logging, and performance optimization. Communication & Business Understanding * Ability to understand business workflows, not just code. * Can explain complex technical ideas simply and clearly. * Works directly with end-users and adapts tools based on feedback. Nice to Have * Healthcare operations / scheduling / staffing workflow familiarity. * Speech-to-Text, Text-to-Speech, Voice Processing Experience * Knowledge of HIPAA and security practices around PHI/PII. * Experience with Apple Silicon GPU/ML workloads (e.g., Mac Studio-based compute clusters) ## Description We are seeking an AI Engineer with hands-on experience in open-source language models, local inference, RAG systems, agent architectures, function/tool calling, MCP, and end-to-end data pipelines. This role requires someone who can: * Understand business processes deeply, * Communicate effectively with non-technical team members, * Architect AI solutions that are stable, scalable, and actually useful in day-to-day operations. * This is a design * build * deploy * iterate role where your work will directly impact core business workflows. Key Responsibilities * Design, build, and deploy AI-powered tools and assistants that support clinical, staffing, * scheduling, analytics, and administrative workflows. * Work with open-source LLMs (LLaMA, Mistral, Gemma, etc.) and local inference runtimes * (Ollama, vLLM, Text Generation Inference). * Implement RAG pipelines using embeddings, vector databases (Chroma, Qdrant, Weaviate, pgvector), and retrieval heuristics tailored to business context. * Build multi-tool / function-calling agents, including execution planning, state management, and iterative reasoning flows. * Architect and integrate MCP-based agents with internal systems, CRMs, databases, analytics dashboards, and forms/workflows. * Develop and maintain data pipelines for ingestion, cleaning, semantic indexing, embeddings, storage, and scheduled refresh. * Deploy models and pipelines both locally and in the cloud (AWS, containerized GPU servers, macOS AI compute environments). * Optimize inference performance, caching, batching, routing, and cost vs. latency trade-offs. * Collaborate directly with non-technical staff to gather requirements and translate real operational needs into practical AI tools. * Document workflows, maintain best practices, and train internal users on effective tool usage. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Optimizing Discovery: PostgreSQL's Role in Transforming GetYourGuide's Search](https://www.wearedevelopers.com/videos/1647-optimizing-discovery-postgresql-s-role-in-transforming-getyourguide-s-search) - [Intro to FastAPI](https://www.wearedevelopers.com/videos/462-intro-to-fastapi) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Multilingual NLP pipeline up and running from scratch](https://www.wearedevelopers.com/videos/901-multilingual-nlp-pipeline-up-and-running-from-scratch) - [Dynamic Entities in .NET: Building Low-Code Systems on Top of Entity Framework Core](https://www.wearedevelopers.com/videos/100218-dynamic-entities-in-net-building-low-code-systems-on-top-of-entity-framework-core) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path](https://www.wearedevelopers.com/magazine/655-from-prototype-to-production-build-ai-agents-with-this-free-4-course-learning-path) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)