> Markdown version of [/jobs/ext/2737621-senior-ai-back-end-engineer](https://www.wearedevelopers.com/jobs/ext/2737621-senior-ai-back-end-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI Back-End Engineer - **Company:** The Job Network - **Location:** United States (Remote available) - **Experience:** Expert - **Salary:** $90,000.0 - $150,000.0 - **Contract:** Permanent contract - **Skills:** Adobe Flash, Application Programming Interfaces (APIs), Artificial Intelligence, Data Analysis, Microsoft Azure, Continuous Integration, Cursor (Graphical User Interface Elements), JSON, Python (Programming Language), PostgreSQL, Queueing Systems, RabbitMQ, Search Technologies, SQLAlchemy, WebSocket, Azure Service Bus, GitHub Copilot, Retrieval-Augmented Generation, Large Language Models, Grafana, Multi-Agent Systems, Backend, Fastapi, Apache Kafka, Bitbucket, GPT, Docker, Jenkins - **Published:** September 5, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/phpy7e5xiw ## About the Role * 5 + yrs building production APIs in Python; 2 + yrs with FastAPI (or similar async stack). * Deep knowledge of async I/O, Pydantic v2, DI, and observability. * Hands-on with Semantic Kernel or comparable agent frameworks. * Practical RAG implementations using Azure AI Search, pgvector, or Chroma. * Strong Postgres skills, including SQLModel/SQLAlchemy 2 and Alembic migrations. * Proven integrations or Side Projects with LLM APIs (OpenAI, Gemini) and structured-output design. * Dependency management via Poetry and virtual-env isolation. * End-to-end CI/CD ownership (build * scan * test * deploy). * Excellent analytical and problem-solving ability. * Remote work readiness with daily overlap of at least 09:00 - 13:00 EST. Nice to have * Event/message queues (RabbitMQ, Azure Service Bus, Kafka). * Observability stacks (Grafana, LangFuse) for LLM cost governance. ## Description * Design & scale async REST/WebSocket APIs with Python 3.11+ + FastAPI, using dependency-injection, type hints, and clean vertical-slice architecture. * Implement multi-agent workflows with Semantic Kernel (handoff, sequential, concurrent) to route traffic among specialised LLM agents. * Integrate LLM providers (OpenAI GPT-4.1/mini, Google Gemini 2.5 Flash) behind a provider-agnostic layer for A/B and cost-aware routing. * Deliver Retrieval-Augmented Generation with vector stores such as Azure AI Search, pgvector, or Chroma. * Expose tool-using agents via OpenAI Assistants (Code-Interpreter) for data-analysis / file-manipulation tasks. * Evolve schemas with SQLModel / SQLAlchemy 2 & Alembic; tune Postgres for high-concurrency async access. * Maintain robust CI/CD (Bitbucket Jenkins) that lint, type-check, test, package (Docker), and deploy. * Instrument services with structlog JSON logs, OpenTelemetry traces, and cost/latency metrics; hold p95 < 100 ms. * Champion AI-assisted development (GitHub Copilot, Cursor) and share pragmatic problem-solving practices with the team. ## Related Videos - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [On a Secret Mission: Developing AI Agents](https://www.wearedevelopers.com/videos/1510-on-a-secret-mission-developing-ai-agents) - [Streaming AI Responses in Real-Time with SSE in Next.js & NestJS](https://www.wearedevelopers.com/videos/1630-streaming-ai-responses-in-real-time-with-sse-in-next-js-nestjs) - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) ## Related Articles - [Dev Digest 137 - AI'm not sure about this](https://www.wearedevelopers.com/magazine/485-dev-digest-137-ai-m-not-sure-about-this) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Graph and AI Trends 2026: Why Is AI Running but Not Yet Delivering?](https://www.wearedevelopers.com/magazine/680-graph-and-ai-trends-2026-why-is-ai-running-but-not-yet-delivering) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix)