> Markdown version of [/jobs/ext/1498019-data-scientist-machine-learning-aws-and-data-bricks](https://www.wearedevelopers.com/jobs/ext/1498019-data-scientist-machine-learning-aws-and-data-bricks). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data scientist machine, learning, AWS and data bricks - **Company:** Elevate IQ Inc - **Location:** Plano, TX, United States - **Salary:** $76,736.0 - $92,413.0 - **Contract:** Permanent contract - **Skills:** Agile Methodology, Amazon Web Services, Amazon S3, Data Analysis, Big Data, Cloud Database, Cluster Analysis, Computer Programming, Databases, Continuous Integration, Data Cleansing, Data Infrastructure, Database Queries, Revision Control Systems, Identity and Access Management, Python (Programming Language), Linear Regression, Machine Learning, Natural Language Processing, NumPy, Scrum Methodology, SQL Databases, Unstructured Data, Data Processing, Cloud Platform System, Feature Engineering, Random Forest, Apache Spark, State Machines, Deep Learning, Model Validation, AWS Lambda, Git, Pandas, Pyspark, Scikit Learn, Infrastructure Automation Frameworks, Information Technology, Xgboost, Feature Selection, Machine Learning Operations, Functional Programming, Cloudwatch, Api Gateway, Restful APIs, Software Version Control, Serverless Computing, Docker, Unsupervised Learning, Databricks - **Published:** July 30, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=aba5f50b13f9e876 ## About the Role The candidate should have hands-on experience with algorithms such as XGBoost and Random Forest, along with strong proficiency in Python libraries including Pandas and NumPy., Bachelor's or Master's degree in Computer Science, Data Science, Statistics, Mathematics, Engineering, or a related field. - Strong programming experience with Python. - Strong experience with Pandas, NumPy, and Scikit-learn. - Hands-on experience building machine-learning models using XGBoost and Random Forest. - Strong understanding of supervised and unsupervised machine-learning techniques. - Experience with data preprocessing, feature engineering, model validation, and hyperparameter optimization. - Hands-on experience with Databricks, Spark, or PySpark. - Experience working with AWS services, particularly AWS Lambda and Amazon S3. - Strong SQL skills and experience working with relational or cloud-based databases. - Understanding of model deployment, monitoring, scalability, and production support. - Experience using Git or another version-control system. - Strong analytical, problem-solving, and communication skills. Preferred Qualifications - Experience with Amazon SageMaker and MLflow. - Experience developing end-to-end machine-learning pipelines. - Knowledge of MLOps, CI/CD, Docker, and Infrastructure as Code. - Experience building REST APIs for machine-learning model inference. - Knowledge of time-series forecasting, natural language processing, or deep learning. - Experience working in Agile or Scrum development environments. - AWS or Databricks certifications are a plus. Key Technical Skills Programming: Python, SQL Data Processing: Pandas, NumPy, PySpark Machine Learning: XGBoost, Random Forest, Scikit-learn, Gradient Boosting Cloud: AWS, Lambda, S3, SageMaker, Glue, Step Functions, CloudWatch Big Data Platform: Databricks, Apache Spark ## Description We are seeking an experienced Data Scientist with strong expertise in machine learning, Python, AWS, and Databricks. The ideal candidate will be responsible for analyzing complex datasets, developing predictive models, and deploying scalable machine-learning solutions in cloud environments., Collect, clean, transform, and analyze structured and unstructured datasets. - Perform exploratory data analysis to identify patterns, trends, correlations, and business insights. - Develop, train, test, and optimize machine-learning models using algorithms such as: - XGBoost - Random Forest - Decision Trees - Logistic and Linear Regression - Gradient Boosting - Clustering and other supervised or unsupervised learning techniques - Perform feature engineering, feature selection, and hyperparameter tuning. - Evaluate models using appropriate metrics such as accuracy, precision, recall, F1-score, ROC-AUC, RMSE, and MAE. - Build reusable data-processing and machine-learning pipelines using Python, Pandas, and NumPy. - Work with large-scale datasets using Databricks, Apache Spark, and PySpark. - Develop and deploy cloud-based data science solutions on AWS. - Build serverless workflows and model-processing components using AWS Lambda. - Work with AWS services such as S3, SageMaker, Glue, Step Functions, CloudWatch, IAM, and API Gateway. - Deploy, monitor, maintain, and retrain machine-learning models in production environments. - Collaborate with data engineers, cloud engineers, software developers, analysts, and business stakeholders. - Translate business requirements into analytical and machine-learning solutions. - Document model assumptions, methodologies, performance, limitations, and results. - Follow machine-learning development, testing, version control, security, and governance best practices. ## Related Videos - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Vectorize all the things! Using linear algebra and NumPy to make your Python code lightning fast.](https://www.wearedevelopers.com/videos/562-vectorize-all-the-things-using-linear-algebra-and-numpy-to-make-your-python-code-lightning-fast) - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Docker build without Docker](https://www.wearedevelopers.com/videos/100114-docker-build-without-docker) ## Related Articles - [Top Big Data Technologies That You Need to Know](https://www.wearedevelopers.com/magazine/108-top-big-data-technologies-that-you-need-to-know) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Got AI ideas but no money? Here are 10 free ways to level up your AI skills with Google Cloud](https://www.wearedevelopers.com/magazine/600-got-ai-ideas-but-no-money-here-are-10-free-ways-to-level-up-your-ai-skills-with-google-cloud) - [Best Coding Boot Camps in Germany](https://www.wearedevelopers.com/magazine/237-best-coding-boot-camps-in-germany) - [MLops – Deploying, Maintaining And Evolving Machine Learning Models in Production](https://www.wearedevelopers.com/magazine/115-mlops-deploying-maintaining-and-evolving-machine-learning-models-in-production)