Software Engineer - Observability
Listed on 2026-07-26
-
Software Development
Cloud Engineer - Software, Software Architect, Backend Developer, DevOps
Intuit is a leading software provider of business and financial management solutions for small and mid‑size businesses, consumers, financial institutions and accounting professionals. You probably know us by our flagship products, Quick Books®, Quicken® and Turbo Tax®, but that's just the start. we're taking on exciting challenges, such as SaaS and mobile applications. Over 50 million users, seven million small businesses and 1,600 financial institutions depend on Intuit because we innovate at the crossroads of real customer problems and breakthrough technology.
Join us and let your ingenious ideas be heard.
Interested in creating and leading the platforms that are high scale and mission critical? Want to solve large scale and highly availability platform challenges for on premise and public cloud deployments? Intuit is seeking Staff Software Engineer, who is characterized by progressive technical experience and has demonstrated progression in technical prowess, to join PDX Observability Engineering team.
The Staff Software Engineer will join the Core PDX Observability Engineering team at Intuit to design and deliver the next‑generation logging and observability platform. This role focuses on creating and leading high‑scale, mission‑critical pipelines and solving large‑scale, high‑availability challenges across on‑premise and public cloud (AWS, GCP) deployments. You will own the architecture and evolution of Intuit's One Logging system — spanning ingestion, routing, cost optimization, and MCP Server‑driven automation — building platform capabilities that maximize velocity for thousands of Intuit developers.
Responsibilities- Architect, build, and evolve the One Intuit Logging system end‑to‑end from log generation at the edge through ingestion, routing, and storage.
- Own and drive pipeline and cost optimization initiatives across the logging stack, reducing ingestion volume and infrastructure spend without losing signal fidelity.
- Lead design and development of core logging components:
Front End Logging Service (FELS), S3 Log Writer, Kinesis/Cloud Watch Log Writer, Log Router, Asterias Splunk, GCP Logs Processor, Index Controller, and Asset‑to‑Log DB. - Drive the Automation Revamp/Rewrite initiative, modernizing legacy tooling into scalable, maintainable services.
- Design and maintain edge/collection agents — Fluent Bit Daemon Set, OIL sidecar (Fluent Bit), EC2 Logger Agent — and integrate Kubernetes metadata enrichment into the pipeline.
- Build and extend the FELS Onboarding Plugin to streamline developer onboarding to the logging platform.
- Leverage MCP Server capabilities to enable AI‑assisted authoring, automation, and operational tooling across the observability platform.
- Build observability into the platform itself — Grafana dashboards, metrics, and alerting for pipeline health, throughput, and cost.
- Partner with platform governance efforts (e.g., Splunk Craft) to enforce ingestion quality, policy, and guardrails upstream in the pipeline.
- Provide technical leadership and mentorship, setting engineering standards and design direction across the team.
- Collaborate cross‑functionally with SRE, platform, and product engineering teams to align logging platform capabilities with organization‑wide needs.
- 8-10+ years of experience in software engineering, with significant experience designing and operating large‑scale distributed systems, logging/data pipelines, or observability platforms
- Bachelor's degree (BE/BTech/MS/MTech) in Computer Science or related field required;
- Deep expertise in building and operating high‑throughput data pipelines (log/event ingestion, streaming, routing) at multi‑TB/PB daily scale.
- Strong hands‑on experience with Kubernetes, containerized workloads, and sidecar/daemonset architectures (e.g., Fluent Bit)
- Proficiency with public cloud platforms (AWS — S3, Kinesis, Cloud Watch, EC2; GCP — logging/monitoring services)
- Experience with Splunk or equivalent log management/observability platforms at scale
- Strong programming skills in Go, Java, Python, or similar languages used in infrastructure/platform engineering
- Demonstrated ability to lead architecture and design for…
(If this job is in fact in your jurisdiction, then you may be using a Proxy or VPN to access this site, and to progress further, you should change your connectivity to another mobile device or PC).