> Markdown version of [/jobs/ext/2462953-senior-ai-researcher](https://www.wearedevelopers.com/jobs/ext/2462953-senior-ai-researcher). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Senior AI Researcher - **Company:** Ivo Inc. - **Location:** San Francisco, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $200,000.0 - $325,000.0 - **Contract:** Permanent contract - **Skills:** Microsoft Word, Artificial Intelligence, Cluster Analysis, Databases, Information Extraction, Open Source Technology, Tensorflow, Web Application Frameworks, Pytorch, Large Language Models, Deep Learning, Information Technology, Virtual Agents, Software Coding - **Published:** August 4, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=68d394c4b8480d5b ## About the Role * A Ph.D. in Computer Science, Engineering, Mathematics, Physics, or a related quantitative field - or equivalent industry research experience with a comparable track record. * Evidence of exceptional ability - a paper, shipped system, open-source contribution, competition result, or hard problem you cracked that puts you meaningfully ahead of your peers. * Deep, hands-on experience in deep learning research and development, particularly with LLMs. Strong working knowledge of modern frameworks (PyTorch, JAX, or TensorFlow) and the surrounding open-source ecosystem. * Expertise in at least one of: agentic systems, reasoning, parameter-efficient fine-tuning (PEFT) methods, quantization, inference optimization (e.g., speculative decoding), hallucination mitigation, novel architectures in deep learning, or robust evaluation methodology for LLMs. * Excellent communication skills, with the ability to articulate complex research findings clearly to both technical and non-technical audiences. * A bias toward action: you ship rather than perfect, you measure rather than guess, and you'd rather have a working prototype today than a polished plan next week., * Experience with long-context modeling, retrieval, or grounded generation in high-stakes domains (legal, medical, financial). * Prior work on hallucination detection, calibration, or interpretability. * A track record of building from zero in fast-paced startup or research environments. You'll thrive at Ivo if: * You love writing code, but you love impact more. We're a team of engineers at heart, but our #1 goal is the best possible product. That means making pragmatic choices and looking for 80/20 solutions over the elegant 100% answer. * You're relentlessly resourceful. When the path isn't clear, you find one. When the tool doesn't exist, you build it. * You have a strong internal sense of urgency. You'd rather do it today than tomorrow. * You're excited by the adventure of building a company. Startup experience is preferred, but not required - the disposition matters more than the resume line. ## Description * An AI agent that lives in MS Word and edits the document for you [2023] * Ditching imprecise embeddings models in favor of agentic RAG [2023] * Large-scale LLM-based legal fact extraction [2024] * A legal assistant that can search large contract databases without sacrificing accuracy [2024] * Clustering legal documents descended from the same family [2025] * Automatic deviation analysis to locate buried risk in huge contract databases [2025] * Merging contracts with their amendments to produce a time series of "composite" contracts (a customer actually cried when we showed her this) [2025], * You'll own a research roadmap end-to-end: identifying the right problems, designing experiments, prototyping, and shipping the winners into production alongside the engineering team. Concretely, that looks like: * Advance the core AI platform. Design and implement novel approaches to the problems at the heart of Ivo's product: reasoning over long-context legal corpora, contract comparison and redlining, information extraction, and automated drafting and editing. * Make our models trustworthy. Conduct research on procedural hallucination detection and resolution, calibration, and explainability. In a domain where a single fabricated citation can sink a deal, the bar for groundedness is uncompromising - your job is to keep raising it. Push frontier techniques into production. Explore and apply advanced fine-tuning, PEFT, and distillation techniques to make our models faster, cheaper, and more accurate on legal- specific tasks. Evaluate emerging work in agentic systems, long-context modeling, and reasoning, and figure out which ideas actually move the needle for our customers. * Build the evaluation infrastructure. Design and maintain datasets, benchmarks, and evals for training and measuring model performance on complex legal text. Define the metrics that matter, and hold the team to them. * Ship. Partner closely with Engineering and Product to take prototypes from notebook to production, write internal reports that influence the technical direction of the platform, and present findings to both technical and non-technical audiences across the company. ## Related Videos - [Hiring AI Native Talents](https://www.wearedevelopers.com/videos/100268-hiring-ai-native-talents) - [Machine learning in the browser with TensorFlowjs](https://www.wearedevelopers.com/videos/155-machine-learning-in-the-browser-with-tensorflowjs) - [Kubernetes and Microservices with Multi-Model Databases](https://www.wearedevelopers.com/videos/382-kubernetes-and-microservices-with-multi-model-databases) - [Photonic Computing: Programming a New Class of AI Accelerators (incl. Live Coding)](https://www.wearedevelopers.com/videos/100196-photonic-computing-programming-a-new-class-of-ai-accelerators-incl-live-coding) - [The End of Software as we know it](https://www.wearedevelopers.com/videos/1393-the-end-of-software-as-we-know-it) - [How We Built a Machine Learning-Based Recommendation System (And Survived to Tell the Tale)](https://www.wearedevelopers.com/videos/752-how-we-built-a-machine-learning-based-recommendation-system-and-survived-to-tell-the-tale) ## Related Articles - [Navigating the AI Shift](https://www.wearedevelopers.com/magazine/629-navigating-the-ai-shift) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [How to start an AI project for a good cause and boost your career](https://www.wearedevelopers.com/magazine/15-how-to-start-an-ai-project-for-a-good-cause-and-boost-your-career) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [MLOps And AI Driven Development](https://www.wearedevelopers.com/magazine/82-mlops-and-ai-driven-development) - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models)