Data Engineering - Research Assistant

The University of Maryland
College Park, United States
6 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Temporary contract
Employment type
Full-time (> 32 hours)
Compensation
$47,840.0 - $104,000.0
Working hours
Regular working hours
Job source

Tech stack

Automation of Tests Basic Access Authentication Cloud Computing Information Systems Databases Continuous Integration Information Engineering Extract Transform Load (ETL) Database Queries DevOps Information Retrieval Intelligence Analysis
+19 more
JSON Python (Programming Language) PostgreSQL Linux Commands Machine Learning Microsoft Office MongoDB Data Classification Retrieval-Augmented Generation Flask (Web Framework) Delivery Pipeline Git Fastapi Pytest Containerization Information Technology Restful APIs Data Pipelines Docker

Job description

The Applied Research Laboratory for Intelligence & Security (ARLIS) at the University of Maryland is a University-Affiliated Research Center (UARC) dedicated to advancing research, innovation, and technology transition to improve decision making for U.S. national security. ARLIS combines deep scientific expertise with operational insight to address challenges in intelligence analysis, cybersecurity, artificial intelligence / machine learning, quantum science, and human-machine teaming. Researchers, scientists, engineers, and analysts at ARLIS collaborate with government agencies, industry partners, and academic institutions to deliver actionable insights and transformative solutions through research and development. Employees at ARLIS work on projects of critical importance, contribute directly to the nation’s security, and are supported by a culture that values integrity, collaboration, and professional growth.

The Data Engineering Research Assistant [TW2.1]works directly with project teams to build and run Python-based data pipelines, prepare, load, and validate data in databases, and help implement search and retrieval components under the guidance of senior technical staff.

Physical Demands: Sedentary work performed in a normal office environment; exerts up to 10 pounds of force occasionally and/or negligible amount of force frequently or constantly to lift, carry, push, pull or otherwise move objects, including the human body. Ability to attend meetings both on and off campus. Spending long hours in front of a computer screen.

Requirements

  • Bachelor’s degree in Computer Science, Data Science, Information Systems, Engineering, Mathematics, or a related field from an accredited college or university., * Professional or substantial project-based experience programming in Python.
  • Experience designing and implementing data pipelines or ETL/ELT jobs that read from one or more sources, transform data, and load into a database.
  • Hands-on experience with at least one database (document or relational), including schema design, writing queries, and validating loaded data.
  • Experience using Linux command-line tools and working in a Git-based workflow (branches and pull requests).
  • Experience writing and executing tests in Python (e.g., using Pytest).
  • Experience using Docker or similar containerization tools to run or troubleshoot applications.

Knowledge, Skills, and Abilities:

  • Ability to quickly internalize and work within modular codebases and design re-runnable, configurable jobs.
  • Conceptual understanding of embeddings and vector similarity search, or prior experience with information retrieval / NLP projects.
  • Familiarity with tools and technologies such as MongoDB, PostgreSQL, JSON Schema, Qdrant, pgvector, FAISS, or sentence-transformers (experience with some subset is acceptable).
  • Ability to build or confidently learn to build REST APIs with FastAPI or Flask, including basic authentication.
  • Strong analytical and problem-solving skills, including the ability to profile datasets, identify anomalies, and communicate implications clearly.
  • Strong written communication skills, including concise technical documentation and status reporting.
  • Demonstrated discipline and care when working with sensitive or controlled data; familiarity with data classification, provenance, and access-control concepts is a plus.
  • Ability to work both independently and collaboratively in a multidisciplinary research environment.
  • Skill in the use of Microsoft Office products.
  • Skill in troubleshooting system errors.
  • Ability to multi-task and prioritize assignments.
  • Ability to analyze situations and determine the best recourse for response.

Must be able to obtain a US security clearance. If selected, must meet the requirements for access to classified information and will be subject to a government security clearance investigation that includes criminal and credit history checks, as well as verification of U.S. citizenship, birth, education, employment, and military history. Final offer is contingent upon the candidate’s ability to successfully obtain the necessary interim Secret security clearance, as determined by ARLIS, prior to commencing employment., * Relevant technical certifications (e.g., cloud, database, or DevOps) are welcome but not required.

  • Experience contributing to a retrieval-augmented generation, search system, or similar information-retrieval project (academic, internship, or research).
  • Experience with Pytest or similar tools for testing Python code.
  • Experience using Docker or similar tools to run or troubleshoot applications.
  • Exposure to DevOps or CI/CD concepts (e.g., automated testing, basic deployment workflows).
  • Exposure to data classification, provenance, or access-control concepts in a research or operational context.
  • Comfort working independently and collaboratively in a fast-paced, evolving environment
  • Strong organizational skills and attention to detail
  • Active or eligible for a security clearance.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:09 min

Balancing data science skillings alongside systems engineering rigor

Nico Schmidt · LIVE

6:21 min

Investigating push inefficiencies with upstream Git experts

Jonathan Creamer · Coffee With Developers

3:47 min

Exploring JSON, CBOR, and JOSE for data serialization

Aaron Russell · LIVE

2:07 min

Inspecting default bridge architectures and custom Docker networks

Oliver Seitz Oliver Seitz · World Congress 2025

2:37 min

Optimizing technical profiles for AI sourcing and recruitment

Mina Golesorkhi Mina Golesorkhi · World Congress 2026 Europe

56 sec

Favorite git commands and the importance of patch commits

Eileen Uchitelle Eileen Uchitelle +1 · Coffee With Developers

Videos

See all

Related articles

See all