> Markdown version of [/jobs/ext/2313150-data-engineering-research-assistant](https://www.wearedevelopers.com/jobs/ext/2313150-data-engineering-research-assistant). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineering - Research Assistant - **Company:** The University of Maryland - **Location:** College Park, United States - **Salary:** $47,840.0 - $104,000.0 - **Contract:** Temporary contract - **Skills:** Automation of Tests, Basic Access Authentication, Cloud Computing, Information Systems, Databases, Continuous Integration, Information Engineering, Extract Transform Load (ETL), Database Queries, DevOps, Information Retrieval, Intelligence Analysis, JSON, Python (Programming Language), PostgreSQL, Linux Commands, Machine Learning, Microsoft Office, MongoDB, Data Classification, Retrieval-Augmented Generation, Flask (Web Framework), Delivery Pipeline, Git, Fastapi, Pytest, Containerization, Information Technology, Restful APIs, Data Pipelines, Docker - **Published:** August 30, 2026 - **Apply:** https://www.dice.com/job-detail/de777286-1b0c-43be-bea1-9ca4641a21a3 ## About the Role * Bachelor's degree in Computer Science, Data Science, Information Systems, Engineering, Mathematics, or a related field from an accredited college or university., * Professional or substantial project-based experience programming in Python. * Experience designing and implementing data pipelines or ETL/ELT jobs that read from one or more sources, transform data, and load into a database. * Hands-on experience with at least one database (document or relational), including schema design, writing queries, and validating loaded data. * Experience using Linux command-line tools and working in a Git-based workflow (branches and pull requests). * Experience writing and executing tests in Python (e.g., using Pytest). * Experience using Docker or similar containerization tools to run or troubleshoot applications. Knowledge, Skills, and Abilities: * Ability to quickly internalize and work within modular codebases and design re-runnable, configurable jobs. * Conceptual understanding of embeddings and vector similarity search, or prior experience with information retrieval / NLP projects. * Familiarity with tools and technologies such as MongoDB, PostgreSQL, JSON Schema, Qdrant, pgvector, FAISS, or sentence-transformers (experience with some subset is acceptable). * Ability to build or confidently learn to build REST APIs with FastAPI or Flask, including basic authentication. * Strong analytical and problem-solving skills, including the ability to profile datasets, identify anomalies, and communicate implications clearly. * Strong written communication skills, including concise technical documentation and status reporting. * Demonstrated discipline and care when working with sensitive or controlled data; familiarity with data classification, provenance, and access-control concepts is a plus. * Ability to work both independently and collaboratively in a multidisciplinary research environment. * Skill in the use of Microsoft Office products. * Skill in troubleshooting system errors. * Ability to multi-task and prioritize assignments. * Ability to analyze situations and determine the best recourse for response. Must be able to obtain a US security clearance. If selected, must meet the requirements for access to classified information and will be subject to a government security clearance investigation that includes criminal and credit history checks, as well as verification of U.S. citizenship, birth, education, employment, and military history. Final offer is contingent upon the candidate's ability to successfully obtain the necessary interim Secret security clearance, as determined by ARLIS, prior to commencing employment., * Relevant technical certifications (e.g., cloud, database, or DevOps) are welcome but not required. * Experience contributing to a retrieval-augmented generation, search system, or similar information-retrieval project (academic, internship, or research). * Experience with Pytest or similar tools for testing Python code. * Experience using Docker or similar tools to run or troubleshoot applications. * Exposure to DevOps or CI/CD concepts (e.g., automated testing, basic deployment workflows). * Exposure to data classification, provenance, or access-control concepts in a research or operational context. * Comfort working independently and collaboratively in a fast-paced, evolving environment * Strong organizational skills and attention to detail * Active or eligible for a security clearance. ## Description The Applied Research Laboratory for Intelligence & Security (ARLIS) at the University of Maryland is a University-Affiliated Research Center (UARC) dedicated to advancing research, innovation, and technology transition to improve decision making for U.S. national security. ARLIS combines deep scientific expertise with operational insight to address challenges in intelligence analysis, cybersecurity, artificial intelligence / machine learning, quantum science, and human-machine teaming. Researchers, scientists, engineers, and analysts at ARLIS collaborate with government agencies, industry partners, and academic institutions to deliver actionable insights and transformative solutions through research and development. Employees at ARLIS work on projects of critical importance, contribute directly to the nation's security, and are supported by a culture that values integrity, collaboration, and professional growth. The Data Engineering Research Assistant [TW2.1]works directly with project teams to build and run Python-based data pipelines, prepare, load, and validate data in databases, and help implement search and retrieval components under the guidance of senior technical staff. Physical Demands: Sedentary work performed in a normal office environment; exerts up to 10 pounds of force occasionally and/or negligible amount of force frequently or constantly to lift, carry, push, pull or otherwise move objects, including the human body. Ability to attend meetings both on and off campus. Spending long hours in front of a computer screen. ## Related Videos - [How a Small Team Shrank a Microsoft Monorepo by 94%](https://www.wearedevelopers.com/videos/1236-how-a-small-team-shrank-a-microsoft-monorepo-by-94) - [Tips and Tricks for Working with JSON](https://www.wearedevelopers.com/videos/1229-tips-and-tricks-for-working-with-json) - [Docker Compose: Rediscovered](https://www.wearedevelopers.com/videos/1978-docker-compose-rediscovered) - [Data Science, ML & AI in the Oil and Gas Industry at NDT Global - Dr. Katja Träumner](https://www.wearedevelopers.com/videos/1308-data-science-ml-ai-in-the-oil-and-gas-industry-at-ndt-global-dr-katja-traumner) - [Git for Code Reviews](https://www.wearedevelopers.com/videos/429-git-for-code-reviews) - [Coffee with Developers - Maria Apazoglou](https://www.wearedevelopers.com/videos/1209-coffee-with-developers-maria-apazoglou) ## Related Articles - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Dev Digest 121 - AI goes offline](https://www.wearedevelopers.com/magazine/456-dev-digest-121-ai-goes-offline) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [Coffee with Developers - Maria Apazoglou - Making AI understandable for all in production](https://www.wearedevelopers.com/magazine/475-coffee-with-developers-maria-apazoglou-making-ai-understandable-for-all-in-production) - [Résumé-Driven Development: How IT trends affect the job market for software developers](https://www.wearedevelopers.com/magazine/59-resume-driven-development-how-it-trends-affect-the-job-market-for-software-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers)