Artificial Intelligence Engineer
- Discuss this with your agent
- Open in Claude
- Open in ChatGPT
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Role details
Tech stack
+11 more
Job description
We are seeking an AI Engineer with hands-on experience in open-source language models, local inference, RAG systems, agent architectures, function/tool calling, MCP, and end-to-end data pipelines.
This role requires someone who can:
- Understand business processes deeply,
- Communicate effectively with non-technical team members,
- Architect AI solutions that are stable, scalable, and actually useful in day-to-day operations.
- This is a design * build * deploy * iterate role where your work will directly impact core business workflows.
Key Responsibilities
- Design, build, and deploy AI-powered tools and assistants that support clinical, staffing,
- scheduling, analytics, and administrative workflows.
- Work with open-source LLMs (LLaMA, Mistral, Gemma, etc.) and local inference runtimes
- (Ollama, vLLM, Text Generation Inference).
- Implement RAG pipelines using embeddings, vector databases (Chroma, Qdrant, Weaviate, pgvector), and retrieval heuristics tailored to business context.
- Build multi-tool / function-calling agents, including execution planning, state management, and iterative reasoning flows.
- Architect and integrate MCP-based agents with internal systems, CRMs, databases, analytics dashboards, and forms/workflows.
- Develop and maintain data pipelines for ingestion, cleaning, semantic indexing, embeddings, storage, and scheduled refresh.
- Deploy models and pipelines both locally and in the cloud (AWS, containerized GPU servers, macOS AI compute environments).
- Optimize inference performance, caching, batching, routing, and cost vs. latency trade-offs.
- Collaborate directly with non-technical staff to gather requirements and translate real operational needs into practical AI tools.
- Document workflows, maintain best practices, and train internal users on effective tool usage.
Requirements
Do you have experience in System performance optimization?, * AI / ML / LLM
- Hands-on deployment of open-source models (fine-tuned or instruct-tuned models a plus).
- Strong understanding of vector search, embeddings, context window strategies, and RAG best practices.
- Experience building agent architectures with structured function/tool calling.
- Familiarity with MCP, LangGraph, LlamaIndex, LangChain, or similar orchestration frameworks.
Software & Systems
- Strong Python experience; familiar with Django / FastAPI / Flask or similar frameworks.
- Experience building data pipelines (ETL/ELT, semantic chunking, scheduled indexing).
- Experience deploying AI systems: Docker, AWS EC2 / ECS / Lambda, GPU instances, or local inference stacks.
- Comfort with monitoring, logging, and performance optimization.
Communication & Business Understanding
- Ability to understand business workflows, not just code.
- Can explain complex technical ideas simply and clearly.
- Works directly with end-users and adapts tools based on feedback.
Nice to Have
- Healthcare operations / scheduling / staffing workflow familiarity.
- Speech-to-Text, Text-to-Speech, Voice Processing Experience
- Knowledge of HIPAA and security practices around PHI/PII.
- Experience with Apple Silicon GPU/ML workloads (e.g., Mac Studio-based compute clusters)
Benefits & conditions
3.33.3 out of 5 stars Remote $100,000 - $160,000 a year - Full-time, Pulled from the full job description
- Health insurance
- 401(k) matching
- Paid time off
- Vision insurance
- Dental insurance, * 401(k) matching
- Dental insurance
- Health insurance
- Paid time off
- Vision insurance
About the company
Are you passionate about making a meaningful difference in children’s lives? Do you value a supportive team environment and opportunities to grow your career? If so, The Perfect Child is the place for you! We are dedicated to providing top-tier ABA services! We are a data-driven organization building secure, modern systems that power clinical, operational, and administrative workflows. Our next stage of growth involves integrating practical, production-grade AI into daily business operations.
Apply for this position
This job is hosted externally. Click below to view the full posting and apply.
Prepare application
- Draft this with your agent
- Open in Claude
- Open in ChatGPT
Good distractions
Talks and stories from around this role — technically off-topic, practically not.
Moments
Explore playlistsVideos
See allRelated articles
See all
How to Become an AI Engineer
What Are Large Language Models?
From Prototype to Production: Build AI Agents with This Free 4-Course Learning Path
Dev Digest 137 - AI'm not sure about this