> Markdown version of [/jobs/ext/3075555-ai-engineer](https://www.wearedevelopers.com/jobs/ext/3075555-ai-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # AI Engineer - **Company:** Realty Professionals, LLC - **Location:** Austin, TX, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Microsoft Azure, Cloud Computing, Software Debugging, Monitoring of Systems, Identity and Access Management, Python (Programming Language), Machine Learning, Open Web Application Security, Software Safety, Software Technical Review, Google Cloud, Large Language Models, Backend, Low Latency, Virtual Agents, Terraform, New Relic (SaaS) - **Published:** September 25, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=4e245311bd6fbfdf ## About the Role * 4+ years of professional software/ML engineering experience, with hands-on LLM application work * Bachelor's degree or equivalent experience * Strong Python proficiency * Real classification-metrics literacy - precision/recall/F1 (macro vs weighted), confusion matrix debugging, dataset curation, and calibration to a production distribution - * Experience building or tuning LLM prompts/classifiers and evaluating them with an eval framework (DeepEval, RAGAS, Arize Phoenix, LangSmith, OpenAI Evals, or a home grown harness) * Experience turning small seed sets into robust labeled eval datasets (dataset-synthesis tooling, prompt-optimization loops) * Familiarity with prompt-injection / jailbreak defense concepts (OWASP LLM Top 10, input/output filtering, least-privilege tool access, adversarial testing) * Guardrails & Infrastructure Experience (Bonus) * Hands-on with a cloud guardrail service (Google Cloud Model Armor, AWS Bedrock Guardrails, Azure AI Content Safety) * Terraform / IaC for cloud infrastructure * Experience with Google Cloud Vertex AI / Gemini * Understanding of monitoring and observability tools (New Relic or similar) * Exposure to regulated / compliance-sensitive domains (fair housing, fair lending, healthcare, finance, trust & safety) * Red-teaming or AI-security background Expectations Technical Excellence * Independently design, implement, and tune guardrail and evaluation systems - Make sound architectural decisions that balance safety, latency, and user experience - Write clean, maintainable, well-tested code that sets the standard for the team - Keep eval datasets calibrated to real traffic and emerging attack patterns - Stay current with AI-safety research, jailbreak techniques, and new guardrail tooling ## Description We are seeking an AI Engineer to join our AI Integrations Team with deep expertise in AI safety, guardrails, and LLM evaluation to join our AI Integrations team. You will be responsible for building and tuning the classifiers that keep our AI assistant compliant, safe, and helpful - working at the intersection of AI and consumer-facing products., The AI Integrations Team is a cross-functional squad that builds the future of how Realtor.com integrates AI across its products. Our flagship is RealAssist, an LLM-powered real-estate assistant embedded across web and mobile. Because it advises consumers on housing, it operates under real Fair Housing Act compliance obligations - our AI safety layer is a launch gate, not an afterthought. This role owns that layer, including: * LLM-as-a-judge classifiers for Fair-Housing compliance and content moderation The guardrail evaluation harness that proves it all works before it ships * A cloud prompt-injection / jailbreak backstop (Google Cloud Model Armor) Technical Ownership * Build, tune, and own LLM-as-a-judge classifiers for Fair-Housing compliance and content moderation, targeting high recall on disallowed content without overblocking legitimate users * Design and run guardrail evaluation pipelines: curate labeled datasets, maintain train/test/validate splits, and run offline prod-replay overblocking evals - * Report confusion matrices and precision/recall/F1 per category to make safety decisions defensible * Integrate and operate cloud guardrail services as a runtime prompt-injection/jailbreak screening layer, with fail-open behavior and alerting * Red-team the assistant and wire safety-regression checks into CI so guardrails don't silently degrade as the product grows * Manage guardrail infrastructure as code with Terraform (templates, IAM, project shape) - Participate in technical design reviews and architecture discussions * Partner with ML, backend, product, and legal/compliance stakeholders to define risk tiering and get new AI capabilities through safety review before launch, * Own the AI safety layer end-to-end from design through deployment * Anticipate problems and proactively address them through red-teaming and regression checks * Make data-driven decisions grounded in eval metrics * Deliver high-quality work consistently on schedule Why Join Us: Cutting-Edge Technology * Work with the latest AI technologies from leading providers * Build the safety layer for LLMs running in production * Shape the future of safe, compliant AI-powered real estate search at scale - Millions of users rely on Realtor.com to find their next home * Your work directly affects major life decisions for consumers * Keep the AI that guides those decisions fair, safe, and trustworthy Growth * Work alongside Staff and Principal engineers who will challenge and support you - Exposure to full-stack technologies including mobile, backend, and AI - Clear path to career growth with opportunities for technical leadership How We Work: We balance creativity and innovation on a foundation of in-person collaboration. For most roles, our employees work four or more days in our offices, where they have the opportunity to collaborate in-person, adding richness to our culture and knitting us closer together. ## Related Videos - [Hack Me If You Can: Designing Unbreakable LLM Guardrails](https://www.wearedevelopers.com/videos/100209-hack-me-if-you-can-designing-unbreakable-llm-guardrails) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [Infrastructure as Code: The Developer's Secret Weapon](https://www.wearedevelopers.com/videos/1221-infrastructure-as-code-the-developer-s-secret-weapon) - [Swapping Low Latency Data Storage Under High Load](https://www.wearedevelopers.com/videos/746-swapping-low-latency-data-storage-under-high-load) - [Beyond the Hype: Building Trustworthy and Reliable LLM Applications with Guardrails](https://www.wearedevelopers.com/videos/1594-beyond-the-hype-building-trustworthy-and-reliable-llm-applications-with-guardrails) - [Nest.js - TypeScript in the backend can also be clean](https://www.wearedevelopers.com/videos/1033-nest-js-typescript-in-the-backend-can-also-be-clean) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [ I Gave a Video Editor More Autonomy Than a Trading Bot. On Purpose.](https://www.wearedevelopers.com/magazine/773-i-gave-a-video-editor-more-autonomy-than-a-trading-bot-on-purpose) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Dev Digest 132 - Binging WADFlix?](https://www.wearedevelopers.com/magazine/473-dev-digest-132-binging-wadflix)