Forecasting Lead Data Scientist

Helishores Inc
United States
23 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
2 years minimum
Working hours
Regular working hours
Job source

Tech stack

Encodings Python (Programming Language) Machine Learning SQL Databases Snowflake Git

Job description

The candidate must be able to name the industry and the outcome variable for each such engagement. Retail same-store analysis is the classic form; the analogue here is comparing similar schools and events rather than following one trend line.

  • Presents to non-statisticians: business outcome first, method second; confidence stated in plain language; explicitly states what the forecast cannot do; never opens with an undefined statistical term.

  • Can teach the method to a client team, not only execute it.

  • Participate actively in stand-ups and backlog refinement, engage business stakeholders directly, understand why the business is asking a question, and challenge or refine the request when it is wrong.

  • Strategic recommendations are expected alongside hands-on delivery.

Requirements

Required:

  • Must be able to work EST hours

  • 8+ years of applied forecasting.

  • Two or more comparable forecasting engagements led start to finish.

  • Comparable-unit / same-store forecasting experience.

  • Executive communication.

  • Thought leadership.

  • Multivariable regression, plus collinearity analysis and VIF interpretation.

  • Forecast model development, tuning, selection and holdout validation.

  • Metric fluency: R², WAPE, MAPE, p-values - and why WAPE is used at event grain (many events sell zero, which breaks MAPE).

  • Sparse and zero-inflated data. Many variables populate on under 25% of events, some as low as 10%. Nulls must never be silently treated as zeros.

  • Data-leakage discipline and point-in-time correctness: every feature must exist before the event starts.

  • Python and SQL; reproducible notebooks.

  • Snowflake, including Snowflake ML Model Registry (model versions carry metrics and training-dataset references).

  • Git and pull-request workflow; all code merged to the client repository, no private forks.

Preferred:

  • Architecture Decision Records (ADRs) and written process documentation.

  • Categorical encoding at scale (~30-35 source variables expand to ~70 columns).

  • Sports, streaming, ticketing or subscription-business domain exposure.

  • Hierarchical or mixed-effects models for low-volume segments.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:17 min

Optimizing character encoding with Kim variable byte encoding

Douglas Crockford Douglas Crockford · World Congress 2024

1:33 min

Integrating internal APIs and maintaining data sovereignty

Mahran Meißner Mahran Meißner · World Congress 2026 Europe

1:41 min

Scaling role analysis and forecasting using artificial intelligence

Rudi Bauer Rudi Bauer +1 · Cappuccino with HR

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

6:07 min

Identifying public and proprietary data sources for predictive modeling

Becky Gandillon · LIVE

Videos

See all

Related articles

See all