Senior Data Scientist

Hewlett-Packard Enterprise
Houston, TX, United States
3 days ago
Apply on arc.dev
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
5 years minimum
Compensation
$153,500.0 - $310,500.0
Working hours
Regular working hours
Job source

Tech stack

Artificial Intelligence Amazon Web Services Artificial Neural Networks Microsoft Azure Big Data Computer Programming Continuous Integration Data Centers Data Mining Data Visualization DevOps Distributed Data Store
+23 more
Elasticsearch Apache Hadoop Apache Hive Network Topologies Information Retrieval Python (Programming Language) Machine Learning Natural Language Processing Open Source Technology Redis Azure Machine Learning Software Engineering Data Streaming Data Processing Apache Spark Multi-Cloud Naive Bayes Indexer Information Technology Apache Flink Data Pipelines Golang Apache Storm

Job description

Senior Data Scientist will be engaged in data science-related research and software application development and engineering duties related to our AI Datacenter technology and autonomous platform to provide an unprecedented visibility and operational efficiency and into the user experience. The Senior Data Scientist will collaborate with other engineers and to build the next generation of autonomous Datacenter networks leveraging big data and predictive models., * Design and implement machine learning solutions which require to process terabytes of streaming data to detect anomalies in DC networks of our customers, predict problems and future trends, provide Root Cause Analysis (60%), All legitimate job opportunities will come through official company channels, and candidates are responsible for verifying the credentials of any third party claiming to represent the company. Any reliance on fraudulent communication is at the individual’s own risk, and HPE disclaims legal liability for any resulting damages. If you suspect recruitment fraud, do not share personal information or make any payments and report the incident to your local authorities immediately.

Requirements

  • Solid statistics and math background, good knowledge of machine learning methods like k-Nearest Neighbors, Naive Bayes, SVM, Decision Forests.
  • Excellent Communication Skills to articulate observations and use cases with PM and network domain experts who are not experienced in AI/ML through data visualization tool.
  • Have done time series data analysis, forecasting and correlation is preferrable.
  • Have utilized latest AI/ML techniques, such as Neural Networks, Transformer, etc. for time series data or interested to explore these techniques for time series data.
  • Analyze feature requirements from product manager, collaborate with engineers and data scientists to design the solutions.
  • Require good understanding of datacenter networking topology and protocols.
  • Troubleshoot production environment and customer reported issues (20%)
  • Require the knowledge of the multi-cloud production environment
  • Require the agility to troubleshoot open-source data processing engine, such as Apache Spark, Apache Storm and Apache Flink
  • Utilize analytical and programming skills and open-source systems, such as Hadoop, Hive, Spark, Elasticsearch, Redis, etc. develop data processing pipeline required efficacy and latency (20%)
  • Require good knowledge and experience of the big data tool sets and techniques of distributed storage and computation engine
  • Require the experience to develop the reusable and highly scalable data processing component
  • Require good knowledge and experience to work with cloud based CICD tools and cloud devops teams to collect stats and create monitors for our data processing pipelines
  • Require good understanding of MCPs and Agentic frameworks., * Bachelor’s degree in Computer Science/ Engineering/Mathematics or equivalent experience
  • 5+ years of experience Search Indexing, Ranking, Information Retrieval and Querying.
  • Proficient in Python and Golang
  • Proficient in implementing NLP, Machine Learning models and algorithms into production at scale., * PhD degree in Statistics, Operations Research, Computer Science or equivalent and 5+ years of relevant experience. Or Master´s Degree in these areas and at least 8 years of relevant experience.
  • Experience with statistical data analysis, data mining, and querying.
  • Experience in deploying and leading complete ML platforms in AWS/GCP/Azure.

Benefits & conditions

United States of America: Annual Salary USD 153,500 - 310,500 in California

The listed salary range reflects base salary. Variable incentives may also be offered.”

About the company

Hewlett Packard Enterprise is the global edge-to-cloud company advancing the way people live and work. We help companies connect, protect, analyze, and act on their data and applications wherever they live, from edge to cloud, so they can turn insights into outcomes at the speed required to thrive in today’s complex world. Our culture thrives on finding new and better ways to accelerate what’s next. We know varied backgrounds are valued and succeed here. We have the flexibility to manage our work and personal needs. We make bold moves, together, and are a force for good. If you are looking to stretch and grow your career our culture will embrace you. Open up opportunities with HPE.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on arc.dev
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:55 min

Demonstrating semantic routing thresholds with the Redis vector library

1:08 min

Building solutions with open source GoLang infrastructure tools

Jad Wahab · LIVE

2:17 min

Mapping the maturity roadmap for scaled devops adoption

Dominik Krichbaum Dominik Krichbaum · World Congress 2026 Europe

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

3:42 min

Comparing in-memory and Redis storage for cache scalability

Simone Sanfratello · World Congress 2022

Videos

See all

Related articles

See all