Data Scientist in Arlington

Energy Jobline
Arlington, VA, United States
1 day ago
Apply on www.energyjobline.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience required
1 year minimum
Working hours
Regular working hours

Tech stack

JavaScript (Programming Language) Application Programming Interfaces (APIs) Agile Methodology Artificial Intelligence Amazon Web Services Business Analytics Applications Microsoft Azure Big Data Continuous Integration Web Scraping Extract Transform Load (ETL) Data Security
+20 more
Relational Databases Github R (Programming Language) Java Database Connectivity Python (Programming Language) Machine Learning Natural Language Processing Open Database Connectivity Azure Data Lake SQL Databases Systems Integration Text Mining Unstructured Data Azure Data Factory Data Lakes Pyspark Restful APIs Software Version Control Data Pipelines Databricks

Job description

  • Develop, manage, and optimize automated data pipelines to support reliable analytics and data quality.
  • Extract, clean, transform, normalize, and validate structured and unstructured data.
  • Integrate data from flat files, relational databases, APIs, external systems, and other sources using JDBC/ODBC, REST APIs, and web scraping.
  • Develop data and analytics solutions using AWS and/or Azure, including Databricks, Azure Data Factory, and Azure Data Lake.
  • Use Azure DevOps and/or GitHub to support source control, CI/CD pipelines, and Agile development.
  • Apply Python, SQL, R, and/or JavaScript to develop datasets, models, dashboards, visualizations, and reports.
  • Apply advanced analytics including machine learning, AI, NLP, predictive analytics, statistical modeling, and data/text mining.
  • Develop, train, evaluate, deploy, and maintain machine learning and AI models.
  • Use PySpark/PySpark SQL to process and analyze large-scale datasets.
  • Identify trends, anomalies, patterns, and risks to support audit, investigative, oversight, and fraud, waste, and abuse activities.
  • Translate technical findings into clear narratives, recommendations, visualizations, and presentations.
  • Develop technical documentation, requirements, test plans, methodologies, and training materials.
  • Collaborate with technical teams, stakeholders, and management and provide guidance on data access, quality, storage, and analytics.

Requirements

Required

  • US Citizenship required
  • 1+ year of experience working with AWS and/or Azure services, such as Databricks, Azure Data Factory, and Azure Data Lake.
  • Experience with Azure DevOps and/or GitHub, CI/CD pipelines, and Agile methodologies.
  • Experience developing and optimizing automated data pipelines and ETL/ELT processes.
  • Proficiency with Python and SQL; experience with R and/or JavaScript is a plus.
  • Experience working with structured and unstructured data and integrating multiple data sources.
  • Knowledge of machine learning, artificial intelligence, NLP, predictive analytics, statistical analysis, or data/text mining.
  • Strong analytical, problem-solving, technical writing, and communication skills.
  • Ability to work independently and effectively within a collaborative team environment.

  • Experience with contract, procurement, and accounts payable (AP) data to identify trends, anomalies, and potential risks.
  • Experience using PySpark/PySpark SQL for large-scale data processing.
  • Experience with data lakes, data lakehouses, relational databases, and business intelligence tools.
  • Experience supporting federal audit, investigation, oversight, fraud, waste, or abuse-related activities.
  • Experience developing predictive models, NLP solutions, dashboards, and investigative or analytical visualizations.
  • Experience presenting technical findings, training materials, or conference presentations to large audiences.

About the company

The Midtown Group is an Equal Opportunity Employer. All qualified applicants will receive consideration for employment without regard to , , , , , , , , or status as a protected veteran. We are a small, woman-owned business certified by the Women’s Business Enterprise Council (WBENC). Operating from our headquarters in Washington, DC, we provide trusted staffing services nationwide. Our clients include thousands of the most prestigious Fortune 500 companies, law firms, financial organizations, tech innovators, non-profits, and lobbying firms, as well as federal, state and local government agencies.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.energyjobline.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

6:36 min

Funding open source through GitHub Accelerator and Sponsors

Stormy Peters · World Congress 2023

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy · LIVE

3:14 min

Structuring career paths and localized data architectures

Ulrich Wurstbauer +1 · LIVE

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph · LIVE

Videos

See all

Related articles

See all