> Markdown version of [/jobs/ext/2365219-software-engineer](https://www.wearedevelopers.com/jobs/ext/2365219-software-engineer). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Software Engineer - **Company:** Grid Dynamics (nasdaq: Gdyn) - **Location:** United States (Remote available) - **Experience:** Expert - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Cloud Computing, Continuous Integration, Distributed Systems, Elasticsearch, Github, Gradle, HP Systems Insight Manager, Queueing Systems, Apache Solr, Spinnaker, Systems Integration, Datadog, Data Logging, System Availability, Indexer, Backend, Build Management, Kubernetes, Information Technology, Apache Kafka, Restful APIs, Pagerduty, Jenkins - **Published:** August 28, 2026 - **Apply:** https://www.dice.com/job-detail/bfc8fa18-bfcb-43bd-8b22-02cec14c2d1c ## About the Role * Strong proficiency in Java, including backend development. * Experience building RESTful services and resilient, high-performance backend components. * Experience designing and implementing CI/CD automation using Gradle, Jenkins, Spinnaker, and GitHub. * Solid working knowledge of AWS services, including deploying and managing applications in the cloud. * Proficiency with Kubernetes for container orchestration, deployment, and application monitoring. * Experience building or operating observability tooling (metrics, dashboards, alerting, logging, or tracing) for distributed systems. * Experience designing systems for high availability, scalability, and performance. * Comfortable working independently on ambiguous problems and collaborating across multiple engineering teams. * Bachelor's degree in Computer Science, a related technical field, or equivalent practical experience Would be a plus * Experience with Apache Solr, Elasticsearch, or other distributed search and indexing technologies. * Experience integrating with Apache Kafka and other message queuing systems. * Experience building internal developer platforms or self-service infrastructure tooling. * Experience with PagerDuty or similar alerting and incident management tooling. * Familiarity with large-scale multi-tenant infrastructure and the operational challenges of onboarding new consumers safely. ## Description We are looking for a Senior Software Engineer who can design and build automation that streamlines cluster provisioning, configuration, and lifecycle management, extend observability so that on-call engineers and platform users can quickly understand cluster health and performance across a growing fleet, and build self-service tooling that reduces the need for manual, one-off engineering support during onboarding. You'll work across the CI/CD pipeline, cloud infrastructure, and Solr cluster management stack, and collaborate closely with both the platform team (managed service, deployment, and reliability) and the SRE team (alarming, capacity, on-boarding)., * Design and build automation for onboarding new Solr use cases and tenants onto the managed platform, reducing manual setup and configuration steps. * Extend and improve observability for the Solr platform, including metrics, dashboards, alerting, and tracing, so cluster health and performance are visible at scale. * Design and develop tooling that gives platform consumers self-service visibility into their clusters, capacity, and configuration. * Collaborate with the Platform and SRE sub-teams to identify gaps in tooling, automation, and monitoring as the platform scales, and prioritize the highest impact improvements. * Troubleshoot and resolve system issues, improving system availability and performance as the number of onboarded use cases grows. * Contribute to runbooks, documentation, and internal tooling that make the platform easier for the broader team to operate and support. ## Related Videos - [Remote Driving on Plant Grounds with State-of-the-Art Cloud Technologies](https://www.wearedevelopers.com/videos/251-remote-driving-on-plant-grounds-with-state-of-the-art-cloud-technologies) - [Modular Secrets to Lightning-Fast Android Builds](https://www.wearedevelopers.com/videos/1428-modular-secrets-to-lightning-fast-android-builds) - [The Road to MLOps: How Verivox Transitioned to AWS](https://www.wearedevelopers.com/videos/1050-the-road-to-mlops-how-verivox-transitioned-to-aws) - [Innovating Developer Tools with AI: Insights from GitHub Next](https://www.wearedevelopers.com/videos/1268-innovating-developer-tools-with-ai-insights-from-github-next) - [From Messy Queries to Scalable Systems - How Data Engineering actually works](https://www.wearedevelopers.com/videos/100203-from-messy-queries-to-scalable-systems-how-data-engineering-actually-works) - [Give your build some love, it will give it back!](https://www.wearedevelopers.com/videos/514-give-your-build-some-love-it-will-give-it-back) ## Related Articles - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Dev Digest 120 - Apple and peers](https://www.wearedevelopers.com/magazine/455-dev-digest-120-apple-and-peers) - [The Best X (Twitter) Accounts for Developers](https://www.wearedevelopers.com/magazine/294-the-best-x-twitter-accounts-for-developers) - [Fully Remote Software Engineer Jobs](https://www.wearedevelopers.com/magazine/447-fully-remote-software-engineer-jobs) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated)