> Markdown version of [/jobs/ext/2116965-ai-software-engineer](https://www.wearedevelopers.com/jobs/ext/2116965-ai-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Software Engineer - **Company:** Oteemo, Inc - **Location:** United States - **Contract:** Permanent contract - **Skills:** JavaScript (Programming Language), Application Programming Interfaces (APIs), Artificial Intelligence, Amazon Web Services, Microsoft Azure, Code Review, Encodings, Information Technology Consulting, Databases, Software Debugging, Identity and Access Management, Python (Programming Language), Key Management, PostgreSQL, Machine Learning, MongoDB, MySQL, NoSQL, OAuth, Redis, E2e Testing, Prometheus, Next.js, Search Technologies, Data Streaming, TypeScript, Web Application Frameworks, WebSocket, Datadog, Data Logging, Pulumi, Cloud Platform System, ReactJS, Retrieval-Augmented Generation, Large Language Models, Grafana, Prompt Engineering, Backend, Cloudformation, Fastapi, Event Driven Architecture, AI Platforms, Kubernetes, Infrastructure Automation Frameworks, Graphql, Machine Learning Operations, Restful APIs, Terraform, GPT, Docker - **Published:** August 19, 2026 - **Apply:** https://jobs.smartrecruiters.com/OteemoInc/744000144378909-ai-software-engineer ## About the Role * Expert-level proficiency in Python with modern frameworks (FastAPI, Flask). * Strong TypeScript/JavaScript skills with deep React and Next.js experience; proven track record designing and building RESTful and GraphQL APIs. * Solid understanding of relational (PostgreSQL, MySQL) and NoSQL (MongoDB) databases. * Experience with authentication systems (OAuth2, JWT, SSO) and security best practices. * Proven track record of shipping high-quality, scalable software to production. * Hands-on experience building and deploying AI/ML applications in production environments. * Deep understanding of LLM integration, prompt engineering, and context managemen. * Proven expertise with RAG systems, including document processing, chunking, embedding, retrieval, and generation. * Experience working with vector databases (Pinecone, Weaviate, Chroma, FAISS, or Qdrant). * Strong grasp of semantic search, similarity algorithms, and hybrid search techniques. * Knowledge of evaluation frameworks for assessing AI system quality and performance. * Production experience with Docker containerization and Kubernetes orchestration. * Strong knowledge of at least one major cloud platform (AWS, Azure, or GCP) and its AI services. * Experience building CI/CD pipelines for ML/AI application. * Proficiency with infrastructure as code tools (Terraform, CloudFormation, Pulumi). * Understanding of monitoring, logging, and alerting best practices; cost optimization experience for cloud and AI workloads. * Strong computer science fundamentals and algorithmic thinking. * Proficiency with Git workflows, code review practices, and collaborative development. * Excellent debugging and problem-solving skills. * Clear technical communication and documentation abilities. ## Description We're seeking an exceptional AI Software Engineer to build and scale enterprise AI applications end to end from database to UI. In this role, you'll work with cutting-edge LLM technology, RAG systems, and production ML infrastructure, combining full-stack development expertise with hands-on AI/ML engineering to ship intelligent systems that deliver real business value at scale. You'll be a key technical contributor, shipping production-ready AI features that users love while ensuring reliability, performance, and cost-effectiveness., * Design and implement end-to-end RAG (Retrieval-Augmented Generation) pipelines for intelligent document search and question-answering across enterprise knowledge bases. * Build production-ready integrations with leading LLMs (GPT-4, Claude, Gemini) for accurate, contextual responses to user queries. * Develop prompt engineering strategies and evaluation frameworks to ensure consistent, high-quality AI outputs. * Create agent systems with tool integration capabilities that can autonomously complete complex tasks. * Implement vector search solutions using Pinecone, Weaviate, or similar technologies for semantic similarity and knowledge retrieval. * Build scalable backend services using Python/FastAPI with type-safe APIs, authentication, and robust error handling. * Develop responsive, performant frontend applications using React/Next.js with real-time streaming for LLM responses. * Design and optimize database schemas across PostgreSQL, MongoDB, and Redis to support high-throughput AI workloads. * Implement WebSocket servers and event-driven architectures for real-time user experiences. * Create comprehensive testing strategies covering unit, integration, and end-to-end tests. * Deploy and manage ML/AI services using Docker containers and Kubernetes orchestration. * Build and maintain CI/CD pipelines for rapid, safe deployment of AI features. * Implement infrastructure as code using Terraform to manage cloud resources (AWS, Azure, or GCP). * Set up monitoring and observability using Datadog, Prometheus/Grafana, and LLM-specific tools (LangSmith, Weights & Biases). * Optimize costs through intelligent caching, batching strategies, and model selection algorithms. * Ensure enterprise-grade security through authentication, authorization, secrets management, and compliance measures. ## Related Videos - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Keeping applications secure by evolving OAuth 2.0 and OpenID Connect](https://www.wearedevelopers.com/videos/100152-keeping-applications-secure-by-evolving-oauth-2-0-and-openid-connect) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) - [AI Driven Development](https://www.wearedevelopers.com/videos/100340-ai-driven-development) ## Related Articles - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline)