> Markdown version of [/jobs/ext/2845869-mid-level-data-engineer](https://www.wearedevelopers.com/jobs/ext/2845869-mid-level-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Mid-Level Data Engineer - **Company:** Crisis Text Line - **Location:** Atlanta, GA, United States (Remote available) - **Experience:** Experienced - **Salary:** $99,704.0 - $116,990.0 - **Contract:** Permanent contract - **Skills:** Artificial Intelligence, Amazon Web Services, Automation of Tests, Cloud Computing, Code Review, Continuous Integration, Information Engineering, Data Infrastructure, Extract Transform Load (ETL), Data Stores, Data Systems, Data Warehousing, Software Debugging, Distributed Computing Environment, Monitoring of Systems, Python (Programming Language), PostgreSQL, Machine Learning, MySQL, Operational Databases, Redis, Software Engineering, Software Systems, SQL Databases, Datadog, Apache Spark, Pyspark, Infrastructure Automation Frameworks, Deployment Automation, AWS Data Analytics, Terraform, Data Pipelines, Databricks - **Published:** September 11, 2026 - **Apply:** https://www.careerjet.com/jobad/us9b941d1b1155bee1156ced509d9a5c41 ## About the Role * Approximately three to five years of professional data engineering, software engineering, or related experience, including meaningful responsibility for production data pipelines. * Professional experience using SQL and Python to build, transform, and troubleshoot data. * Experience building, maintaining, and supporting production data pipelines or ETL/ELT workflows. * Experience with distributed data processing technologies such as Spark or PySpark. * Familiarity with cloud-based environments such as AWS and modern data engineering practices. * Experience testing, debugging, monitoring, and improving the reliability and performance of production data systems. * The ability to independently complete well-defined engineering work, investigate problems, and clearly communicate technical decisions, risks, and tradeoffs. * A collaborative approach to working with engineers, data scientists, and cross-functional partners, with a commitment to building secure and reliable data systems. Preferred Qualifications The following experience would be useful, but we encourage you to apply even if you do not meet every item: * Databricks or a comparable cloud-based data platform. * AWS services and cloud-based data architecture. * Terraform or another infrastructure-as-code tool. * Lakehouse modeling such as medallion architecture and data warehousing. * Datadog or comparable monitoring and observability tools. * Data quality, automated testing, CI/CD, or deployment automation. * Designing data systems for reliability, performance, scalability, and security. * Supporting data used for analytics, reporting, or machine learning workloads. * Claude or other AI-assistance code tools. * Working in a mission-driven, regulated, safety-sensitive, or high-trust environment. Reliable High-Speed Internet Required: Must have a stable high-speed internet connection to support seamless remote collaboration, virtual meetings, online job tasks, etc. ## Description We're looking for a Data Engineer to build, improve, and support the data pipelines and infrastructure that power our products and organization. You'll work with SQL, Python, Spark/PySpark, Databricks, AWS, and Infrastructure as Code to develop reliable, scalable data solutions and help ensure the performance and quality of our production data environment. You'll collaborate with engineers, data scientists, and teams across Crisis Text Line to translate data needs into thoughtful, maintainable solutions. We're looking for someone with strong data engineering fundamentals and hands-on experience building and supporting production data pipelines. You don't need experience with every technology in our environment-we value transferable experience, curiosity, and the ability to learn and solve problems collaboratively. Our Stack You'll work across a modern, cloud-based environment, including: * Languages & Processing: SQL, Python, Spark/PySpark * Data Platform: Databricks * Cloud: AWS * Infrastructure as Code: Terraform * Observability: Datadog * Data Stores: MySQL, PostgreSQL, and Redis * Delivery & Reliability: Automated CI/CD, monitoring, and shared ownership of production data pipelines, * Design, build, maintain, and improve reliable production data pipelines as part of our broader data and engineering environment. * Develop and optimize data processing workflows using SQL, Python, Spark/PySpark, and Databricks. * Partner with engineers, data scientists, and business teams to understand data needs and translate them into practical, maintainable solutions. * Build and maintain integrations between data sources and our data infrastructure. * Use Infrastructure as Code practices to support and improve our AWS data environment. * Monitor pipeline health, data quality, and performance; troubleshoot issues and contribute to solutions when something isn't working as expected. * Automate repeatable processes that improve the reliability, efficiency, and maintainability of our data pipelines. * Participate in shared production support and develop confidence diagnosing and resolving pipeline issues independently. * Contribute to code reviews, technical documentation, and engineering standards that help the team build reliable and maintainable systems. * Bring ideas, ask questions, and identify opportunities to improve our data tools, processes, and engineering practices over time., The Lead Software Engineer is a hands-on technical leader responsible for driving the full-stack design, development, and delivery of software solutions across one or more small, h… + 8 hours ago, The Lead Software Engineer is a hands-on technical leader responsible for driving the full-stack design, development, and delivery of software solutions across one or more small, h… + 8 hours ago + ## Related Videos - [PySpark - Combining Machine Learning & Big Data](https://www.wearedevelopers.com/videos/44-pyspark-combining-machine-learning-big-data) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [MySQL Protocol Features You Should Be Aware Of](https://www.wearedevelopers.com/videos/100267-mysql-protocol-features-you-should-be-aware-of) - [Reducing LLM Calls with Vector Search Patterns - Raphael De Lio (Redis)](https://www.wearedevelopers.com/videos/1714-reducing-llm-calls-with-vector-search-patterns-raphael-de-lio-redis) - [Data Science in Retail](https://www.wearedevelopers.com/videos/586-data-science-in-retail) - [Empowering Retail Through Applied Machine Learning](https://www.wearedevelopers.com/videos/976-empowering-retail-through-applied-machine-learning) ## Related Articles - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [How to Become an AI Engineer](https://www.wearedevelopers.com/magazine/331-how-to-become-an-ai-engineer) - [Making Data Warehouses Fast: A Developer’s Story](https://www.wearedevelopers.com/magazine/107-making-data-warehouses-fast-a-developer-s-story)