Ai Engineer Backend (Remote - Es/Uk Only)

Quadrivia Ai
Madrid, Spain
7 days ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
2 years minimum
Working hours
Regular working hours

Tech stack

Artificial Intelligence ARM Architecture Cloud Computing Data Stores Python (Programming Language) PostgreSQL Octopus Deploy Redis Software Engineering TypeScript WebRTC Data Logging
+8 more
ReactJS Caching Backend Fastapi Kubernetes Api Design Terraform Docker

Job description

About Us Quadrivia is the health technology company behind Q, a comprehensive, controllable, and customizable assistant AI built by clinicians, for clinicians.Addressing the urgent shortage of healthcare professionals, Q provides real?time, personal, and reliable support for clinical tasks across the care continuum.Designed for providers, payers, and pharmaceutical companies, Q is easy to customize and integrates seamlessly into workflows, delivering precise assistance across the care spectrum.The RoleYou’ll build and run Cortex, the core AI architecture behind Q, and the services that sit on top of it: automated AI audits, patient simulators, retrieval (RAG), and the escalation agents that take over in red?guardrail situations.This is a backend role first.The job is to make our AI systems reliable, fast, and observable in production, not to invent new ML.You own the software underneath the agents.What You’ll DoDesign and maintain robust, modular backend systems using clean architectural (SOLID) principles to ensure long?term maintainability, scalability and flexibility as the agentic stack evolves.Own Cortex end?to?end: architecture, API design, service boundaries, reliability targets, and proactively managing failure modes.Build the platform services around it: the automated audit and eval pipeline, patient simulators for testing agents at scale, and the retrieval layer.Write fast, well?tested Python services with FastAPI, asyncio, and pydantic, and get the queues, caching, and data stores right.Wire up the multi?agent orchestration: routing between agents, shared state, and clean tool interfaces.Engineer the RAG pipeline for high?signal retrieval (chunking, hybrid search, re?ranking, caching) and prove the grounding holds.Make the whole thing observable: structured logs, OTEL tracing across the agent graph, cost, latency and token visibility, dashboards, and CI gates that catch regressions before they ship.Minimum QualificationsYour core is backend and software engineering.You write clean, maintainable services and you care how they behave in production.Deep understanding of architectural design patterns (e.g., Clean/Hexagonal Architecture, Domain?Driven Design, SOLID, event?driven) to manage complex system boundaries.At least 2 years, demonstrable, building or scaling user?facing AI software that real users touched.We’ll want to see it.Expert Python, with strong FastAPI, asyncio, pydantic, and production observability.Comfortable with agent patterns and eval?driven development.You’ve worked at a startup before and know what wearing several hats actually costs.Nice to HaveReal?time and voice: WebRTC, LiveKit, SIP, VAD, barge?in, turn?taking.Useful here, not required.Programmatic prompt optimization techniques.LLM?as?judge setups and other evaluation tooling.GCP: Cloud Run or GKE, Pub/Sub, Vertex AI, GCS, Secret Manager, Cloud Logging and Trace.Healthcare data familiarity.Example Problems You’ll TackleStand up the AI audit pipeline so evals run automatically on slices of production traffic, with regression gates wired into CI.Build a patient simulator that lets us stress?test agents at scale before they ever reach a real call.Improve the RAG pipeline with hybrid retrieval and re?ranking, then prove the gains with faithfulness and context metrics.Get OTEL?first tracing across the agent graph, with automated eval triggers on live traffic.Turn EHR integrations into reliable tools the agents can call.Tech StackPython, FastAPI, pydantic, asyncio, Redis, Postgres, vector stores, Docker, Kubernetes, Terraform, ArgoCD, OTEL, TypeScript, React.Real?time stacks (WebRTC, LiveKit, SIP, STT/TTS) where the work touches voice.What Success Looks LikeQuadrivia’s backend becomes a reference for reliability, safety, and performance.Your services run above **% availability under strict regulatory constraints.Other engineers build new clinical workflows and agent capabilities quickly and safely.AI?generated code gets reviewed, corrected, and owned by you.#J-**-Ljbffr

Requirements

Your core is backend and software engineering. You write clean, maintainable services and you care how they behave in production. Deep understanding of architectural design patterns (e.g., Clean/Hexagonal Architecture, Domain?Driven Design, SOLID, event?driven) to manage complex system boundaries. At least 2 years, demonstrable, building or scaling user?facing AI software that real users touched. We’ll want to see it. Expert Python, with strong FastAPI, asyncio, pydantic, and production observability. Comfortable with agent patterns and eval?driven development. You’ve worked at a startup before and know what wearing several hats actually costs. Nice to Have Real?time and voice: WebRTC, LiveKit, SIP, VAD, barge?in, turn?taking. Useful here, not required.

About the company

Madrid, España

About Us Quadrivia is the health technology company behind Q, a comprehensive, controllable, and customizable assistant AI built by clinicians, for clinicians. Addressing the urgent shortage of healthcare professionals, Q provides real?time, personal, and reliable support for clinical tasks across the care continuum. Designed for providers, payers, and pharmaceutical companies, Q is easy to customize and integrates seamlessly into workflows, delivering precise assistance across the care spectrum.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.buscojobs.com.es

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · WWC 2025

2:07 min

Simplifying peer-to-peer connections using WebRTC abstraction libraries

André Dietrich André Dietrich · WWC 2024

2:14 min

Exploring internal AI product initiatives and global engineering roles

Maria Apazoglou · Coffee With Developers

1:23 min

Choosing the right communication protocol for real-time needs

Ahmed Megahd Ahmed Megahd · WWC 2025

2:34 min

Docker sandbox architecture and microVM environment integration

Manuel de la Peña Manuel de la Peña · WWC Europe 2026

Videos

See all

Related articles

See all