> Markdown version of [/jobs/ext/2170252-staff-software-engineer-observability](https://www.wearedevelopers.com/jobs/ext/2170252-staff-software-engineer-observability). Every page supports `.md` or `Accept: text/markdown`. Links point to the HTML versions so they work for humans too. Agent guide: [/agents.md](https://www.wearedevelopers.com/agents.md). --- # Staff Software Engineer - Observability - **Company:** Intuit Inc. - **Location:** Mountain View, CA, United States - **Experience:** Expert - **Salary:** $188,500.0 - $255,000.0 - **Contract:** Permanent contract - **Skills:** Java (Programming Language), Artificial Intelligence, Amazon Elastic Compute Cloud, Amazon S3, Computer Programming, Distributed Systems, Python (Programming Language), Routing, Software Engineering, Data Streaming, Data Logging, Network Routers, Grafana, Containerization, Kubernetes, Information Technology, Cloudwatch, Splunk, Data Pipelines, Golang - **Published:** August 21, 2026 - **Apply:** https://dejobs.org/x/x/E9A9E68602204DF2B0CC3204D45D71C9/job/ ## About the Role * 8-10+ years of experience in software engineering, with significant experience designing and operating large-scale distributed systems, logging/data pipelines, or observability platforms * Bachelor's degree (BE/BTech/MS/MTech) in Computer Science or related field required; * Deep expertise in building and operating high-throughput data pipelines (log/event ingestion, streaming, routing) at multi-TB/PB daily scale. * Strong hands-on experience with Kubernetes, containerized workloads, and sidecar/daemonset architectures (e.g., Fluent Bit) * Proficiency with public cloud platforms (AWS - S3, Kinesis, CloudWatch, EC2; GCP - logging/monitoring services) * Experience with Splunk or equivalent log management/observability platforms at scale * Strong programming skills in Go, Java, Python, or similar languages used in infrastructure/platform engineering * Demonstrated ability to lead architecture and design for mission-critical, high-availability systems * Experience driving cost-optimization initiatives for large-scale infrastructure * Strong track record of technical leadership, mentorship, and cross-team influence without direct reporting authority * Excellent communication skills - able to translate complex technical tradeoffs for both engineering and leadership audiences ## Description * Architect, build, and evolve the One Intuit Logging system end-to-end from log generation at the edge through ingestion, routing, and storage. * Own and drive pipeline and cost optimization initiatives across the logging stack, reducing ingestion volume and infrastructure spend without losing signal fidelity. * Lead design and development of core logging components: Front End Logging Service (FELS), S3 Log Writer, Kinesis/CloudWatch Log Writer, Log Router, Asterias Splunk, GCP Logs Processor, Index Controller, and Asset-to-Log DB. * Drive the Automation Revamp/Rewrite initiative, modernizing legacy tooling into scalable, maintainable services. * Design and maintain edge/collection agents - Fluent Bit DaemonSet, OIL sidecar (Fluent Bit), EC2 Logger Agent - and integrate Kubernetes metadata enrichment into the pipeline. * Build and extend the FELS Onboarding Plugin to streamline developer onboarding to the logging platform. * Leverage MCP Server capabilities to enable AI-assisted authoring, automation, and operational tooling across the observability platform. * Build observability into the platform itself - Grafana dashboards, metrics, and alerting for pipeline health, throughput, and cost. * Partner with platform governance efforts (e.g., SplunkCraft) to enforce ingestion quality, policy, and guardrails upstream in the pipeline. * Provide technical leadership and mentorship, setting engineering standards and design direction across the team. * Collaborate cross-functionally with SRE, platform, and product engineering teams to align logging platform capabilities with organization-wide needs. ## Related Videos - [Shipping Faster with Less: Render on Cloud Hosting, AI Workloads, and the Future of DevOps](https://www.wearedevelopers.com/videos/1894-shipping-faster-with-less-render-on-cloud-hosting-ai-workloads-and-the-future-of-devops) - [5 steps for running a Kubernetes environment at scale](https://www.wearedevelopers.com/videos/88-5-steps-for-running-a-kubernetes-environment-at-scale) - [Our journey with Spring Boot in a microservice architecture](https://www.wearedevelopers.com/videos/511-our-journey-with-spring-boot-in-a-microservice-architecture) - [Go with the Flow: Stop the Leaks Before Your Memory's a Waterfall!](https://www.wearedevelopers.com/videos/100073-go-with-the-flow-stop-the-leaks-before-your-memory-s-a-waterfall) - [Building the Next Generation of Software](https://www.wearedevelopers.com/videos/100186-building-the-next-generation-of-software) - [All your telemetry data from any source in one place](https://www.wearedevelopers.com/videos/57-all-your-telemetry-data-from-any-source-in-one-place) ## Related Articles - [Highest Paying Tech Companies for Developers](https://www.wearedevelopers.com/magazine/220-highest-paying-tech-companies-for-developers) - [Is Software Engineering Over-Saturated?](https://www.wearedevelopers.com/magazine/418-is-software-engineering-over-saturated) - [How to Answer the Interview Question: “Why Do You Want to Be a Software Engineer?”](https://www.wearedevelopers.com/magazine/392-how-to-answer-the-interview-question-why-do-you-want-to-be-a-software-engineer) - [How Much Does a Software Engineer Make? Realistic Software Engineering Salaries](https://www.wearedevelopers.com/magazine/425-how-much-does-a-software-engineer-make-realistic-software-engineering-salaries) - [How We Built a Worry-Free System That Runs for 10+ Years – And What We’d Do Again](https://www.wearedevelopers.com/magazine/751-how-we-built-a-worry-free-system-that-runs-for-10-years-and-what-we-d-do-again) - [How Much FAANG Companies Actually Pay Software Engineers in 2025](https://www.wearedevelopers.com/magazine/230-how-much-faang-companies-actually-pay-software-engineers-in-2025)