> Markdown version of [/jobs/ext/3475675-database-engineer](https://www.wearedevelopers.com/jobs/ext/3475675-database-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Database Engineer - **Company:** Initiate Government Solutions - **Location:** Washington, DC, United States (Remote available) - **Contract:** Permanent contract - **Skills:** Amazon Web Services, Microsoft Azure, Bash Shell, Big Data, Cloud Computing, Cloud Database, Cloud Engineering, Databases, Information Engineering, Data Governance, Data Infrastructure, Data Integrity, Extract Transform Load (ETL), Data Transformation, Data Migration, Data Systems, Relational Databases, Python (Programming Language), Microsoft SQL Server, Operational Databases, Query Optimization, Remote Access Technology, SQL Databases, Data Streaming, Test Data, Backup and Restore, Parquet, Apache Spark, Database Performance, HybridCloud, Cloudformation, Amazon Relational Database Service, Integration Tests, Infrastructure Automation Frameworks, Information Technology, Performance Monitor, Health Level Seven International, Api Design, Terraform, Data Pipelines, Api Management, Databricks, Data Generation - **Published:** September 22, 2026 - **Apply:** https://www.indeed.com/viewjob?jk=fc365d8cfcbda713 ## About the Role * Bachelor's degree in Information Technology, Computer Science, Public Administration, or a related field (Master's preferred). * Hands-on experience with relational databases including SQL Server and distributed SQL platforms such as CockroachDB; familiarity with Amazon RDS or equivalent cloud-managed database services. * Strong ability to query, design, and architect databases for performance and scalability. * Experience with cloud platforms (e.g., AWS, Azure) and cloud-native database services. * Familiarity with distributed SQL or NewSQL database platforms; experience with columnar and open table formats such as Parquet and Delta Parquet. * Proficiency in scripting and automation tools (e.g., Terraform, CloudFormation, Bash, Python). * Proficiency in Python for data engineering tasks including pipeline development, data transformation, and test data generation. * Experience with big data processing frameworks, particularly Apache Spark, for large-scale data transformation and pipeline execution. * Experience developing and maintaining ETL workflows using Databricks or a comparable notebook-based data platform. * Strong understanding of security, compliance, and data governance standards. * Excellent communication skills and attention to detail. * Analytical mind and problem-solving aptitude * Ability to obtain and maintain a Public Trust security clearance * Ability to work in the United States without sponsorship and US Citizenship because of clearance requirement * Strong organizational skills, * Active VA Public Trust * Prior experience with VA data systems or federal health data infrastructure. * Experience with hybrid cloud architectures spanning multiple providers (e.g., Azure and AWS within a single data pipeline). * Databricks certification (e.g., Databricks Certified Associate Developer for Apache Spark) or equivalent data platform certification. * Prior, successful experience working in a remote environment, * Integrity, Honesty, and Ethics: We conduct our business with the highest level of ethics. Doing things like being accountable for mistakes, accepting helpful criticism, and following through on commitments to ourselves, each other, and our customers. * Empathy, Emotional Intelligence: How we interact with others including peers, colleagues, stakeholders, and customers' matters. We take collective responsibility to create an environment where colleagues and customers feel valued, included, and respected. We work within a diverse, integrated, and collaborative team to drive towards accomplishing the larger mission. We conscientiously and meticulously learn about our customers' and end-users' business drivers and challenges to ensure solutions meet not only technical needs but also support their mission. * Strong Work Ethic (Reliability, Dedication, Productivity): We are driven by a strong, self-motivated, and results-driven work ethic. We are reliable, accountable, proactive, and tenacious and will do what it takes to get the job done. * Life-Long Learner (Curious, Perspective, Goal Oriented): We challenge ourselves to continually learn and improve ourselves. We strive to be an expert in our field, continuously honing our craft, and finding solutions where others see problems. ## Description This is a remote access assignment. The Candidate will work remotely daily and will remotely access IGS and customer systems and therein use approved IGS or customer provided communications systems. Travel is not required; however, the candidate may be required to attend onsite client meetings as requested. In this role, you will collaborate with cross-functional teams-including product managers, engineers, security specialists, and business analysts-to deliver high-quality API products in alignment with federal standards and objectives. You will contribute to the program's mission of improving veteran experiences by enabling digital innovation, ensuring compliance, and promoting best practices in API development and platform management. Responsibilities and Duties (Included but not limited to): * Database Architecture & Design: Design scalable, secure, and highly available cloud database solutions with emphasis on relational data modeling and schema design for large-scale migration initiatives. Evaluate and implement appropriate database technologies based on application requirements, including distributed SQL databases in hybrid cloud environments. * Cloud Database Provisioning: Execute structured data migration pipelines across multi-environment workflows (dev-test, pre-production, production). Migrate on-premises relational database sources to cloud-based target platforms, ensuring data integrity, minimal downtime, and performance parity. Manage data staging through cloud object storage and coordinate pipeline progression through each environment. * Performance Monitoring & Optimization: Monitor database performance using cloud-native and third-party tools. Identify and resolve bottlenecks in high-volume environments, including large-scale datasets with sub-second latency requirements. Implement indexing, partitioning, and query optimization strategies to meet performance SLAs. * Security & Compliance: Implement access controls, encryption, and auditing. Ensure compliance with data protection regulations (e.g., HIPAA, HL7). Regularly review and update security policies. * Backup & Disaster Recovery: Validate migrated data across pipeline stages to ensure accuracy, completeness, and structural integrity relative to source systems. Develop and execute test data strategies, including synthetic data generation that mirrors production structures without real PII/PHI. * Automation & Scripting: Develop and maintain ETL notebooks and workflows for initial and incremental data loads. Automate data transformation, format conversion (e.g., raw to Delta Parquet), and pipeline orchestration. Develop scripts for synthetic data generation, routine maintenance, and monitoring using Python and infrastructure-as-code tools such as Terraform. * ETL Pipeline Development: Build and maintain data pipelines using Databricks notebooks and Apache Spark for large-scale data processing. Manage data flow from source systems through cloud object storage staging into target database environments across dev-test, pre-production, and production tiers. * Synthetic Data Generation: Generate structured test datasets that replicate production data schemas without containing real PII or PHI. Load synthetic data into containerized local database instances for initial pipeline validation prior to integration testing. * Cost Optimization: Monitor and optimize database resource usage to control cloud costs. Recommend right-sizing and cost-saving measures. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Dirty Tests And How To Clean Them](https://www.wearedevelopers.com/videos/515-dirty-tests-and-how-to-clean-them) - [Alibaba Big Data and Machine Learning Technology](https://www.wearedevelopers.com/videos/37-alibaba-big-data-and-machine-learning-technology) - [Parquet, Delta, Iceberg & Ducklake - An introduction for developers](https://www.wearedevelopers.com/videos/100075-parquet-delta-iceberg-ducklake-an-introduction-for-developers) - [Modern Data Architectures need Software Engineering](https://www.wearedevelopers.com/videos/1030-modern-data-architectures-need-software-engineering) - [How building an industry DBMS differs from building a research one](https://www.wearedevelopers.com/videos/768-how-building-an-industry-dbms-differs-from-building-a-research-one) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [What Are The Top Skills Required For Azure Developers?](https://www.wearedevelopers.com/magazine/77-what-are-the-top-skills-required-for-azure-developers) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [The Most Popular IT Jobs on the Market](https://www.wearedevelopers.com/magazine/376-the-most-popular-it-jobs-on-the-market) - [Top-Paying Tech Jobs (with Salaries)](https://www.wearedevelopers.com/magazine/372-top-paying-tech-jobs-with-salaries)