> Markdown version of [/jobs/ext/3017053-data-engineer](https://www.wearedevelopers.com/jobs/ext/3017053-data-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Data Engineer - **Company:** Greenhouse Software, Inc. - **Location:** San Francisco, CA, United States - **Experience:** Experienced - **Contract:** Permanent contract - **Skills:** Application Programming Interfaces (APIs), Airflow, Amazon Web Services, Databases, Information Engineering, Extract Transform Load (ETL), Data Warehousing, Cursor (Graphical User Interface Elements), Django Web Framework, Elasticsearch, Python (Programming Language), Ruby on Rails, Search Technologies, Service-Oriented Architecture, Data Streaming, Chatbots, Large Language Models, Prompt Engineering, Backend, GPT, Data Pipelines - **Published:** September 20, 2026 - **Apply:** https://job-boards.greenhouse.io/homelight/jobs/8211546 ## About the Role * 2+ years of experience in data, software, or backend engineering roles * Experience applying LLMs and prompt engineering to real-world workflows * Strong proficiency in Python for building services, data pipelines, and backend systems * Experience designing, building, and operating service-oriented architectures and APIs * Solid understanding of data modeling, system design, and scalable architecture patterns * Hands-on experience with data pipelines and ETL workflows (Airflow or similar) * Experience with AI-assisted development tools (e.g., Cursor, Codex, Claude Code) * Strong analytical thinking with a bias toward automation and building scalable systems Nice to Have * Experience building and scaling product features using frameworks like Ruby on Rails or Django * Familiarity with AWS to support production applications * Familiarity with Elasticsearch or other search technologies for user-facing features * Exposure to LLM-powered product use cases (ChatGPT, conversational interfaces, automation) * Experience working in a startup or fast-paced environment with high ownership ## Description We are building our Data Engineering team to tackle HomeLight's diverse data challenges. This position is an excellent opportunity for an engineer who wants to own the development, optimization, and operation of our data pipeline, which collects, processes, and distributes data to a suite of HomeLight products and teams. You will provide mission-critical data to both our algorithms and internal users, refining our product and identifying new markets. Some projects you will work on: * Optimize and execute on requests to pull, analyze, interpret, and visualize data * Partner with team leaders across the organization to build out and iterate on team and individual performance metrics * Optimize our data release processes, and partner with team leads to iterate on and improve existing data pipelines. * Design and develop systems that ingest and transform our data streams using the latest tools. * Design, build, and integrate new cutting-edge databases and data warehouses, develop new data schemas and figure out new innovative ways of storing and representing our data. * Research, architect, build, and test robust, highly available, and massively scalable systems, software, and services. ## Related Videos - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Coffee with Developers - Maria Apazoglou](https://www.wearedevelopers.com/videos/1209-coffee-with-developers-maria-apazoglou) - [Developing the Backend with Stefan Lingler, CTO at Shpock](https://www.wearedevelopers.com/videos/100360-developing-the-backend-with-stefan-lingler-cto-at-shpock) - [ Evaluating AI models for code comprehension](https://www.wearedevelopers.com/videos/1462-evaluating-ai-models-for-code-comprehension) - [Enjoying SQL data pipelines with dbt](https://www.wearedevelopers.com/videos/823-enjoying-sql-data-pipelines-with-dbt) - [AI Model Management Life Circles: ML Ops For Generative AI Models From Research to Deployment](https://www.wearedevelopers.com/videos/1152-ai-model-management-life-circles-ml-ops-for-generative-ai-models-from-research-to-deployment) ## Related Articles - [What Are Large Language Models?](https://www.wearedevelopers.com/magazine/304-what-are-large-language-models) - [The Prompt Engineer ✍️](https://www.wearedevelopers.com/magazine/216-the-prompt-engineer) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [Data Engineer Salary UK](https://www.wearedevelopers.com/magazine/253-data-engineer-salary-uk) - [The Top ChatGPT Plugins for Developers in 2023](https://www.wearedevelopers.com/magazine/259-the-top-chatgpt-plugins-for-developers-in-2023)