Agentic AI, Senior Software Development Engineer

Zillow Home Loans, LLC
Los Angeles County, CA, United States
2 days ago
Apply on startup.jobs
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$160,900.0 - $257,100.0
Working hours
Regular working hours
Job source

Tech stack

Training Data Application Programming Interfaces (APIs) Artificial Intelligence Machine Learning Azure Machine Learning Workflow Management Systems Large Language Models Multi-Agent Systems AI Platforms Information Technology Machine Learning Operations GPT
+1 more
Databricks

Job description

We are seeking a collaborative, product-minded backend engineer with strong Agentic AI fundamentals and a passion for building reliable, scalable systems that power AI applications.

Requirements

You are a roll-up-your-sleeves technical expert who can combine state-of-the-art AI technology with large-scale backend engineering. You thrive in ambiguous environments, enjoy solving challenging problems, and care deeply about creating reliable systems that other teams can trust., * 4+ years of backend engineering experience, with a track record of designing, shipping, and operating scalable production ML services

  • Experience building platform infrastructure or developer-facing services consumed by multiple teams; you’ve thought carefully about APIs, user experience, reliability, and what it means to have internal customers
  • Hands-on experience with ML Evals & Observability frameworks (Databricks MLflow, or evolving LLM frameworks preferred - LangSmith, Braintrust, Promptfoo, or similar)
  • Experience evaluating third-party AI platforms and making principled build vs. buy decisions for platform infrastructure
  • Fluency working across applied science, ML, product, design, and engineering; you translate between disciplines without losing precision
  • Comfort with LLMs, agentic systems, and evaluation frameworks; familiarity with orchestration tools (LangChain, LangGraph, MCP, or similar)
  • A habit of reaching for AI-assisted development tools in your own workflow; you don’t just build for AI, you build with it
  • Strong ownership and a growth mindset; you’re energized by ambiguity, not slowed by it, * Advanced degree in Computer Science, Machine Learning, or a related field, or equivalent practical experience building products using frontier LLMs, multimodal models, or agent-based systems
  • Experience designing and operating evaluation pipelines: offline batch evals, online scoring, cost/latency/signal tradeoffs, and closing the loop from user signals back to training data.

Benefits & conditions

In this role, you will:

  • Design, build, and scale evaluation frameworks for Zillow’s agentic AI experiences.
  • Build tracing, observability, and quality measurement systems for production AI agents.
  • Create platform infrastructure that enables domain teams across Zillow to build on top of the Evals platform without bottlenecks.
  • Partner with applied scientists and machine learning engineers to integrate new AI evaluation capabilities into production systems.
  • Help evolve how Zillow measures AI quality, reliability, trustworthiness, and product impact.
  • Stay current with emerging agentic AI paradigms, evaluation techniques, and LLM tooling, and translate them into practical platform innovation.
  • Support scaling, reliability, performance optimization, incident response, and cost management for the evaluation layer.
  • Apply first-principles thinking to ambiguous problems and iterate quickly on novel solutions.

This role has been categorized as a Remote position. “Remote” employees do not have a permanent corporate office workplace and, instead, work from a physical location of their choice, which must be identified to the Company. U.S. employees may live in any of the 50 United States, with limited exceptions.

In California, Connecticut, Maryland, Massachusetts, New Jersey, New York, Washington state, and Washington DC the standard base pay range for this role is $160,900.00 - $257,100.00 annually. This base pay range is specific to these locations and may not be applicable to other locations.

In Colorado, Hawaii, Illinois, Maine, Minnesota, Nevada, Ohio, Rhode Island, Vermont, and Virginia the standard base pay range for this role is $152,900.00 - $244,300.00 annually. The base pay range is specific to these locations and may not be applicable to other locations.

In addition to a competitive base salary this position is also eligible for equity awards based on factors such as experience, performance and location. Actual amounts will vary depending on experience, performance and location. Employees in this role will not be paid below the salary threshold for exempt employees in the state where they reside.

About the company

The Agentic AI team at Zillow is transforming the real estate industry by helping millions of people use AI assistants to find their next home. We are building always-on AI experiences that combine personalized user insights with Zillow’s deep real estate knowledge. The Agentic Evals team focuses on one of the hardest parts of building AI experiences: measuring quality. We build evaluation frameworks, tracing systems, observability tools, and feedback loops that help teams understand what works, identify failure modes, and continuously improve AI assistant behavior. Our work ensures Zillow’s AI agents remain trustworthy, responsible, measurable, and ready to scale. As part of this lean, customer-focused team, you will partner with applied scientists, software engineers, machine learning engineers, and product leaders to evolve Zillow’s next-generation AI platform., At Zillow, we’re reimagining how people move-through the real estate market and through their careers. As the most-visited real estate platform in the U.S., we help customers navigate buying, selling, financing and renting with greater ease and confidence. Whether you’re working in tech, sales, operations, or design, you’ll be part of a company that’s reshaping an industry and helping more people make home a reality.

Zillow is honored to be recognized among the best workplaces in the country. Zillow was named one of FORTUNE 100 Best Companies to Work For® in 2025, and included on the PEOPLE Companies That Care® 2025 list, reflecting our commitment to creating an innovative, inclusive, and engaging culture where employees are empowered to grow.

No matter where you sit in the organization, your work will help drive innovation, support our customers, and move the industry-and your career-forward, together.

Zillow Group is an equal opportunity employer committed to fostering an inclusive, innovative environment with the best employees. We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender identity or Veteran status. If you have a disability or special need that requires accommodation, please contact your recruiter directly.

Qualified applicants with arrest or conviction records will be considered for employment in accordance with applicable state and local law.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on startup.jobs
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

40 sec

Generative pre-trained transformer models powering code completions

lgonta lgonta +1 · World Congress 2024

2:27 min

Managing traffic and tracking costs with Databricks Unity Catalog

Viktoria Semaan Viktoria Semaan · World Congress 2026 Europe

2:36 min

Choosing between managed AI platforms and custom governance

Péter Farkas Péter Farkas · Europe 2026 Virtual

1:31 min

Essential AI and human skills for future teams

Alexander Weißhaupt Alexander Weißhaupt +1 · World Congress 2025

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

51 sec

Assessing GPT-4o performance for pull request feedback

Merrill Lutsky Merrill Lutsky · World Congress 2025

Videos

See all

Related articles

See all