Lead Software Engineer

THERON & COMPANY LLC
Columbus, OH, United States
4 days ago
Apply on www.dice.com
Prepare application

Role details

Contract type
Permanent contract
Employment type
Full-time (> 32 hours)
Experience level
Expert
Experience required
7 years minimum
Working hours
Regular working hours
Job source

Tech stack

Java (Programming Language) Application Programming Interfaces (APIs) Agile Methodology Bash Shell Batch Processing Big Data Code Generation Profiling Data Governance Cursor (Graphical User Interface Elements) Data Flow Control Apache Hadoop
+10 more
Job Scheduling Python (Programming Language) Software Engineering Software Systems Data Streaming GitHub Copilot Database Optimization Apache Spark Information Technology Cassandra

Job description

As a Lead Software Engineer, you will be responsible for leading software development initiatives. You will independently design, develop, and test complex software programs and systems. You will also collaborate with team members, mentor junior engineers, and provide technical guidance to ensure the delivery of high-quality software solutions. You will also collaborate with product managers, designers, and other engineers to define, refine, and implement features and enhancements., * Own and evolve large-scale data flows end-to-end batch processing, transformation, clustering, and publication of datasets and indexes ensuring they are reliable, documented, and ready for modernization.

  • Drive modernization of data-flow architecture: evaluate and introduce new techniques, patterns, and tooling; make build-vs-buy and technology choices; set direction others can implement against.

  • Lead solution design for batch and distributed processing: data quality checks, indexing strategies, replay/idempotency patterns, and integration with existing Java/Spring services and APIs.

  • Develop and tune Apache Spark workloads: implement batch jobs, diagnose failures, and improve performance through profiling, troubleshooting, and optimization.

  • Build and maintain scripting and orchestration at scale using bash (required) and related automation; coordinate long-running jobs through enterprise scheduling and operational runbooks.

  • Establish technical standards for batch and data-flow engineering scripting patterns, monitoring, failure handling, and operational handoff and influence practices beyond your immediate team.

  • Collaborate with product managers, leadership, and engineering teams to align roadmaps with product needs and translate organizational goals into executable technical strategy.

  • Troubleshoot and resolve complex production issues in data flows and batch systems; implement preventive, systemic improvements not one-off fixes.

  • Leverage and explore AI-assisted development tools (e.g., GitHub Copilot, Cursor, code generation, smart testing) where appropriate; help assess effectiveness and support adoption.

  • Champion agile methodologies, lead technical and design reviews, and foster cross-team collaboration.

  • Maintain awareness of security, data governance, and quality standards in an enterprise context.

Requirements

Candidates must have technical lead experience, strong batch processing experience with Java, and big data experience with Spark, Hadoop, or Cassandra. Any expertise with Python, Claude, or Cursor is a plus., * Bachelor’s degree in computer science or related discipline, or equivalent work experience.

  • 7+ years of software development experience.
  • Demonstrated professional strength in Java: shipping, maintaining, and evolving production services not occasional or peripheral use.
  • Strong bash scripting at scale in production dataflow and batch environments; comfort maintaining and extending large script-based systems.
  • Hands-on experience with Apache Spark (or Hadoop or Cassanda): developing batch workloads and troubleshooting/tuning performance and reliability (depth in patterns matters; specific cluster or vendor context can be learned on the job).
  • Proven experience owning large-scale batch processing and data flows: design, operation, troubleshooting, and evolution including publishing datasets or indexes at scale.
  • Experience with batch orchestration and job scheduling in production (enterprise schedulers, dependency chains, failure recovery, and operational runbooks).
  • Ability to work with messy, evolving data: inconsistent schemas, multiple sources, and changing requirements; design for robustness and incremental improvement.

Apply for this position

This job is hosted externally. Click below to view the full posting and apply.

Apply on www.dice.com
Prepare application

Good distractions

Talks and stories from around this role — technically off-topic, practically not.

3:28 min

Defining big data and machine learning fundamentals

Ayon Roy · LIVE

3:08 min

Scaling semantic search with Astra DB and Apache Cassandra

David Leconte David Leconte +1 · World Congress 2024

47 sec

Profiling native execution calls with async-profiler

Gonzalo Ortiz Jaureguizar Gonzalo Ortiz Jaureguizar · World Congress 2026 Europe

4:36 min

Hiring passionate software engineers to tackle unprecedented scaling challenges

Dana Lawson Dana Lawson +1 · World Congress 2026 Europe

2:10 min

Why organizations combine big data and machine learning

Ayon Roy · LIVE

4:18 min

Prioritizing communication and structural awareness over strict tool mastery

Liam Hurrel +1 · World Congress 2021

Videos

See all

Related articles

See all