Software Engineer

Parafin Inc
San Francisco, CA, United States
about 2 months ago

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
4 years minimum
Compensation
$230,000.0 - $265,000.0
Working hours
Regular working hours
Job source

Tech stack

Airflow Amazon Web Services Amazon S3 Data Infrastructure Dataspaces Distributed Systems Apache Hadoop Apache Hive Python (Programming Language) Machine Learning Azure Machine Learning Software Engineering
+21 more
SQL Databases Management of Software Versions Datadog Cloud Platform System Data Ingestion System Availability Grafana Apache Spark Backend Build Management Data Lakes Pyspark Core Data Data Analytics Performance Monitor Apache Kafka Machine Learning Operations Presto Terraform Data Pipelines Databricks

Job description

We’re looking for a seasoned software engineer to join Parafin’s Infrastructure team and lead the development of our next-generation Data Platform. This role is critical to ensuring that our data infrastructure is reliable, scalable, and developer-friendly as we continue to power financial services for small businesses.

As a Senior Software Engineer, you’ll be responsible for designing, building, and maintaining the systems that ingest, transform, and serve data across the company. You’ll partner closely with Data Science, Platform Engineering, and Product Engineering teams to support data-driven product development and decision-making.

What You’ll Do:

  • Design and build robust, highly scalable data pipelines and lakehouse infrastructure with PySpark, Databricks, and Airflow on AWS.
  • Improve the data platform development experience for Engineering, Data Science, and Product by creating intuitive abstractions, self-service tooling, and clear documentation.
  • Own and maintain core data pipelines and models that power internal dashboards, ML models, and customer-facing products.
  • Own the Data & ML platform infrastructure using Terraform, including end-to-end administration of Databricks workspaces: manage user access, monitor performance, optimize configurations (e.g., clusters, lakehouse settings), and ensure high availability of data pipelines.
  • Lead projects to improve data quality, testing, observability, and cost efficiency across existing pipelines and backend systems (e.g., migrating Databricks SQL pipelines to dbt, scaling data ingestion, improving data-lineage tracking, and enhancing monitoring).
  • Act as the primary engineering partner for the Data Science team-embedded closely to gather requirements, design scalable solutions, and provide end-to-end support on all engineering aspects of their work.
  • Work closely with backend engineers and data scientists to design performant data models and support new product development initiatives.
  • Share best practices and mentor other engineers working on data-centric systems.

Requirements

  • 4+ years of experience in software engineering with a strong background in data infrastructure, pipelines, and distributed systems.
  • Advanced proficiency in Python and SQL.
  • Hands-on Spark development experience.
  • Expertise with modern cloud data stacks-AWS (S3, RDS), Databricks, and Airflow-and lakehouse architectures.
  • Hands-on experience with foundational data-infrastructure technologies such as Hadoop, Hive, Kafka (or similar streaming platforms), Delta Lake/Iceberg, and distributed query engines like Trino/Presto.
  • Familiarity with ingestion frameworks, developer-experience tooling, and best practices for data versioning, lineage, partitioning, and clustering.
  • Strong problem-solving skills and a proactive attitude toward ownership and platform health.
  • Excellent communication and collaboration skills, especially in cross-functional settings.

Bonus Points

  • Experience with AWS infrastructure using Terraform.
  • Familiarity with observability tools (e.g., Datadog) and cost tracking in cloud environments.
  • Experience with financial systems or building platforms in a fintech setting.
  • Prior work on ML infrastructure: Feature stores (e.g., Tecton), ML model lifecycle (training, deployment, monitoring, retraining), real-time inference.
  • Contributions to internal tooling or open-source projects in the data ecosystem.

Benefits & conditions

  • Salary Range: $230k-$265k
  • Equity grant
  • Medical, dental & vision insurance
  • Work from home flexibility
  • Unlimited PTO
  • Commuter benefits
  • Free lunches
  • Paid parental leave
  • 401(k)
  • Employee assistance program

About the company

At Parafin, we’re on a mission to grow small businesses.

Small businesses are the backbone of our economy, but traditional banks often don’t have their backs. We build tech that makes it simple for small businesses to access the financial tools they need through the platforms they already sell on.

We partner with companies like DoorDash, Amazon, Worldpay, and Mindbody to offer fast and flexible funding, spend management, and savings tools to their small business users via a simple integration. Parafin takes on all the complexity of capital markets, underwriting, servicing, compliance, and customer service for our partners.

We’re a tight-knit team of innovators hailing from Stripe, Square, Plaid, Coinbase, Robinhood, CERN, and more - all united by a passion for building tools that help small businesses succeed. Parafin is backed by prominent venture capitalists including GIC, Notable Capital, Redpoint Ventures, Ribbit Capital, and Thrive Capital. Parafin is a Series C company, and we have raised more than $194M in equity and $340M in debt facilities.

Join us in creating a future where every small business has the financial tools they need.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on dice.com

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

2:15 min

Empowering domain teams with an open data platform

Sandhya Menon Sandhya Menon Ā· WWC Europe 2026

1:34 min

Pivoting careers into specialized platform engineering roles

Xavier Portilla Edo Ā· LIVE

4:32 min

Harnessing Spark with Python using PySpark and Py4J

Ayon Roy Ā· LIVE

3:05 min

Audience questions on AI agents and pipeline vectorization

Joy Joy Ā· WWC 2024

3:37 min

Scaling machine learning pipelines from prototypes to petabytes

Julian Joseph Ā· LIVE

3:03 min

Career evolution in data engineering and AI platforms

Maria Apazoglou Ā· Coffee With Developers

Videos

See all

Related articles

See all