> Markdown version of [/jobs/ext/2435702-data-scientist-ai-research](https://www.wearedevelopers.com/jobs/ext/2435702-data-scientist-ai-research). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist - AI Research - **Company:** Lyra Health - **Location:** Burlingame, CA, United States (Remote available) - **Experience:** Expert - **Salary:** $128,000.0 - $197,000.0 - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Data Analysis, Database Queries, Statistical Hypothesis Testing, JSON, Python (Programming Language), Raw Data, Delivery Pipeline, Prompt Engineering - **Published:** August 8, 2026 - **Apply:** https://jobs.lever.co/lyrahealth/eafbaf7e-3bd5-445d-b3e2-d5927db6bd9f/apply?lever-origin=applied&lever-source%5B%5D=BuiltInNationwide ## About the Role * 5+ years of data analysis in an industry setting with a preference for individuals who have experience in product analytics in the AI/ML space and have partnered with engineering and product teams in prior roles. * Experience collaborating with cross-functional stakeholders across technical, business, and (ideally) clinical domains * Demonstrated understanding of statistical applications and methods (experimentation, probabilities, regression). Experience applying hypothesis testing in an industry or research setting. * Experience developing comprehensive analyses using irregular data from disparate sources. Willing to do whatever it takes to clean messy data prior to deriving insights. * Strong SQL proficiency. Experience writing complex queries with multiple schemas and tables. Preference for candidates who have experience creating views. * Graduate-level Statistics coursework with MS in a quantitative field is preferred (statistics, econometrics, biostatistics, quantitative social sciences)., * Interest in public health a plus * Experience with AI/ML R&D, e.g. prompt engineering and RAG * Experience with Python, including cleaning and analyzing data as well as using Python to parse JSON structured data, call APIs, create user-defined functions, and automate. ## Description Lyra is transforming mental health care by creating a frictionless experience for members, providers, and employers. We connect companies and their employees to mental health providers, therapy, and coaching programs that work. We are looking for a self-driven, technical Data Scientist who cares about impact, ownership and innovation. In this role, you will work closely with Lyra's Data, Product, Engineering, and Clinical teams on AI/ML initiatives that will enhance the experience of mental health patients and providers. This Data Scientist role can be filled remotely anywhere within the US, as well as locally in our Burlingame, CA headquarters. Please note that remote-based candidates must be physically located within the United States. Responsibilities * Work closely with Lyra's Data, Product, Engineering, and other associated teams to enhance the Lyra platform to serve the needs of our providers and clients. * Become an expert in the diverse data sources we collect and partner with data engineers to build pipelines that apply clinical logic to raw data * Extract actionable insights, develop predictive and generative models, and implement rigorous evaluation frameworks to support product iterations * Complete time-sensitive ad hoc data requests for internal and external stakeholders, By applying for this position, you acknowledge that your personal information will be processed as per the Lyra Health Workforce Privacy Notice. Through this application, to the extent permitted by law, we will collect personal information from you including, but not limited to, your name, email address, gender identity, employment information, and phone number for the purposes of recruiting and assessing suitability, aptitude, skills, qualifications, and interests for employment with Lyra. We may also collect information about your race, ethnicity, and sexual orientation, which is considered sensitive personal information under the California Privacy Rights Act (CPRA) and special category data under the UK and EU GDPR. Providing this information is optional and completely voluntary, and if you provide it you consent to Lyra processing it for the purposes as described at the point of collection, for example for diversity and inclusion initiatives. If you are a California resident and would like to limit how we use this information, please use the Limit the Use of My Sensitive Personal Information form. This information will only be retained for as long as needed to fulfill the purposes for which it was collected, as described above. Please note that Lyra does not "sell" or "share" personal information as defined by the CPRA. Outside of the United States, for example in the EU, Switzerland and the UK, you may have the right to request access to, or a copy of, your personal information, including in a portable format; request that we delete your information from our systems; object to or restrict processing of your information; or correct inaccurate or outdated personal information in our systems. These rights may be subject to legal limitations. To exercise your data privacy rights outside of the United States, please contact [email protected]. For more information about how we use and retain your information, please see our Workforce Privacy Notice." ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Data Governance in the Era of AI](https://www.wearedevelopers.com/videos/1622-data-governance-in-the-era-of-ai) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Bringing Clarity to Event Streams: Enabling Analytics and AI Through Rich Metadata](https://www.wearedevelopers.com/videos/1616-bringing-clarity-to-event-streams-enabling-analytics-and-ai-through-rich-metadata) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [The Innovation Formula: Fast Prototyping, Data Analysis, and Real User Insights](https://www.wearedevelopers.com/videos/1421-the-innovation-formula-fast-prototyping-data-analysis-and-real-user-insights) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Data Analyst Salary in the UK](https://www.wearedevelopers.com/magazine/278-data-analyst-salary-in-the-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [How to start an AI project for a good cause and boost your career](https://www.wearedevelopers.com/magazine/15-how-to-start-an-ai-project-for-a-good-cause-and-boost-your-career)