Data Scientist

Allruva Technology Services Incorporated
Irving, TX, United States
3 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Working hours
Regular working hours
Job source

Tech stack

A/B Testing Algorithm Design Data Analysis Big Data Data Cleansing Data Transformation Data Presentation Data Visualization Apache Hadoop Statistical Hypothesis Testing Python (Programming Language) Machine Learning
+16 more
Raw Data Recommender Systems Power BI Tensorflow SQL Databases Tableau (Software) Data Processing Feature Engineering Pytorch Apache Spark Scikit Learn Information Technology Data Analytics Feature Selection Data Pipelines Software Library

Job description

  • Data Analysis and Modeling:
  • Analyze large datasets to extract actionable insights and identify trends.
  • Develop and implement advanced statistical models and machine learning algorithms.

  • Predictive Analytics:
  • Build predictive models for forecasting, classification, and recommendation systems.
  • Evaluate model performance and refine algorithms for continuous improvement.

  • Data Cleaning and Preprocessing:
  • Clean and preprocess raw data for analysis, ensuring data quality and integrity.
  • Collaborate with data engineers to develop and maintain efficient data pipelines.

  • Feature Engineering:
  • Identify relevant features and variables for model development.
  • Conduct exploratory data analysis to inform feature selection and extraction.

  • Collaboration:
  • Collaborate with cross-functional teams to understand business requirements and goals.
  • Communicate findings and insights to both technical and non-technical stakeholders.

  • Algorithm Development:
  • Develop and deploy machine learning models into production environments.
  • Stay updated on the latest advancements in machine learning and data science.

  • Visualization:
  • Create data visualizations and dashboards to present findings and insights.
  • Use tools such as Tableau, Power BI, or similar for effective data storytelling.

  • Testing and Validation:
  • Conduct rigorous testing and validation of models to ensure accuracy and reliability.
  • Perform A/B testing and other experiments to evaluate model effectiveness.

Requirements

Do you have experience in SQL databases?, Do you have a Master’s degree?, + Master’s or Ph.D. degree in Computer Science, Statistics, Mathematics, or a related field.

  • Experience:
  • Proven experience as a Data Scientist with [X] years of relevant experience.
  • Demonstrated success in developing and deploying machine learning models.

  • Technical Skills:
  • Proficiency in programming languages such as Python or R.
  • Experience with machine learning libraries/frameworks (e.g., scikit-learn, TensorFlow, PyTorch).

  • Statistical Analysis:
  • Strong background in statistical analysis and hypothesis testing.
  • Familiarity with advanced statistical techniques.

  • Data Manipulation:
  • Proficient in data manipulation and analysis using SQL.
  • Experience with data preprocessing tools and techniques.

  • Communication Skills:
  • Excellent communication and presentation skills.
  • Ability to convey complex technical concepts to a non-technical audience.

  • Problem-Solving Skills:
  • Strong analytical and problem-solving abilities.
  • Ability to approach business challenges with a data-driven mindset.

Additional Preferred Skills:

  • Experience with big data technologies (e.g., Hadoop, Spark).
  • Knowledge of natural language processing (NLP) for text data analysis.
  • Industry-specific expertise (e.g., finance, healthcare, e-commerce).

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on indeed.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou · Coffee With Developers

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

1:24 min

Moving the semantic layer upstream to avoid vendor lock-in

Piotr Menclewicz Piotr Menclewicz · Europe 2026 Virtual

2:35 min

Preventing remote code execution in PyTorch models

Balázs Kiss · WWC 2023

1:34 min

Bringing diverse skills to industrial data science roles

Katja Träumner

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

Videos

See all

Related articles

See all