> Markdown version of [/jobs/ext/2222700-data-scientist-33394](https://www.wearedevelopers.com/jobs/ext/2222700-data-scientist-33394). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Scientist [33394] - **Company:** Stealth Startup - **Location:** New York, NY, United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Artificial Intelligence, Data Analysis, Application Integration Architecture, Information Engineering, Data Infrastructure, Information Leak Prevention, Systems Theories, Github, Information Extraction, Python (Programming Language), Machine Learning, Open Source Technology, Operational Databases, Software Engineering, SQL Databases, Data Processing, Freeform SQL, Feature Engineering, Large Language Models, Model Validation, Fastapi, Build Tools, Data Management, Machine Learning Operations, Data Pipelines - **Published:** August 25, 2026 - **Apply:** https://arc.dev/remote-jobs/j/redirect/pecg4besk7 ## About the Role * Senior Data Scientist or Analytics Engineer with a demonstrated ability to independently take business problems from initial exploration to production-ready solutions. * Extensive experience with Python and SQL. * Comfortable working across notebooks, APIs, application code, and modern business intelligence platforms. * Strong understanding of applied machine learning, including: * Feature engineering * Model evaluation * Missing data handling * Data leakage prevention * Model calibration * Interpretability * Knowing when simpler solutions outperform more complex ones * Solid data engineering experience, including building pipelines, integrating APIs, maintaining production workflows, and troubleshooting data quality issues. * Excellent communication skills with the ability to explain technical findings to non-technical audiences. * Highly self-motivated with strong ownership and entrepreneurial instincts. * Comfortable leveraging modern AI tools and large language models to improve research, analysis, automation, and productivity while applying sound judgment to model outputs. * Must be based in New York City or willing to relocate. Preferred Qualifications Strong candidates may have experience with one or more of the following: * Building production data products or machine learning systems used by real customers or internal teams. * Maintaining an active GitHub profile, contributing to open-source projects, publishing technical research, or producing high-quality technical writing. * Applying data science to finance, venture investing, marketplaces, growth, CRM, or other operational datasets. * Building AI-assisted research systems, LLM evaluation frameworks, or structured information extraction pipelines. * Working as an early technical hire or founder at a fast-growing startup. * Exceptional quantitative background demonstrated through research, competitions, Olympiads, or a highly rigorous technical education., * You prefer managing projects rather than building and shipping technical solutions yourself. * You prioritize predictable work hours over working in a fast-moving, high-performance environment. * You require frequent direction or detailed task management. * You are unable or unwilling to work onsite in New York City. ## Description We're looking for an experienced Data Scientist to join a small, high-performing engineering team. This is a highly autonomous role where you'll own projects from problem definition through production deployment. You'll work closely with business stakeholders to build data products, predictive models, internal tools, and AI-powered systems that drive strategic decision-making. This position is best suited for someone who enjoys solving ambiguous problems, shipping production-quality solutions, and having direct ownership over their work. What You'll Do * Own data science initiatives end-to-end, turning loosely defined business questions into actionable analyses, models, internal tools, and production systems. * Build and maintain production-grade Python applications for data collection, enrichment, scoring, and AI-assisted research using both internal and external data sources. * Design, train, evaluate, and deploy predictive models while ensuring strong feature engineering, robust validation, and high data quality. * Write complex SQL queries and develop dashboards, recurring reports, ad hoc analyses, and data reconciliations. * Transform large, messy datasets into practical recommendations that improve investment decisions, business operations, portfolio management, and company growth. * Partner directly with stakeholders to identify high-impact opportunities, communicate insights clearly, and continuously improve solutions based on real-world usage. * Help improve internal data infrastructure, automation, and analytical capabilities across the organization., * Solve challenging technical problems that directly influence important business decisions. * Join a small, collaborative team where individual contributions have significant impact. * Build systems that combine data science, machine learning, software engineering, and modern AI technologies. * Work closely with experienced operators, investors, founders, and technical leaders across a broad range of industries. * Gain exposure to emerging technologies, high-growth companies, and real-world business challenges. * Enjoy meaningful ownership, rapid professional growth, and the opportunity to shape the organization's data capabilities. ## Related Videos - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Intro to FastAPI](https://www.wearedevelopers.com/videos/462-intro-to-fastapi) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [Building a framework-independent component library](https://www.wearedevelopers.com/videos/1679-building-a-framework-independent-component-library) - [Building and Deploying Multi-Agent Systems with ADK and Vertex AI](https://www.wearedevelopers.com/videos/1918-building-and-deploying-multi-agent-systems-with-adk-and-vertex-ai) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [The Fastest-Growing Tech Sectors to Look Out for in 2025](https://www.wearedevelopers.com/magazine/373-the-fastest-growing-tech-sectors-to-look-out-for-in-2025) - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)