About the Role
Join Splunk and help shape the future of machine data. Our team builds and operates large-scale services that orchestrate updates and changes across customer environments in Splunk Cloud, working at significant scale across AWS, GCP, and Kubernetes. This role offers an opportunity to apply deep cloud and distributed systems expertise while driving the next generation of automation that powers Splunk Cloud, directly influencing reliability, scalability, and customer trust.
Responsibilities
- Own features end-to-end across the complete software development lifecycle including AI, spanning requirements gathering, technical design, implementation, testing, automation, deployment, and ongoing operational ownership in production.
- Translate customer and business needs into clear technical solutions, making thoughtful trade-offs that balance speed, quality, and reliability.
- Design, build, and maintain highly scalable, distributed systems that handle large volumes of data.
- Identify and resolve pre-production bottlenecks and production issues, improving system resilience and performance.
- Participate in design and code reviews, contributing to shared engineering standards and long-term system health.
- Partner closely with product managers, SREs, and cross-functional teams to deliver meaningful customer outcomes.
- Participate in a rotating on-call schedule, diagnosing and resolving production issues to ensure reliability and customer trust.
- Mentor junior engineers and contribute to a culture of learning, ownership, and continuous improvement.
Requirements
- Bachelor's degree in computer science, Engineering, or a related field, or equivalent practical experience.
- 8+ years of professional experience in software engineering and proficiency in at least one programming language, such as Go, Python, or Java.
- 5+ years of experience in the design and implementation of distributed systems, including databases, distributed file systems, concurrency control, consistency models with API communication paradigms, including REST.
- 2+ years of experience with CI/CD pipelines, Kubernetes, container ecosystems, and microservices-based architectures.
- 3+ years of experience working with public cloud providers, such as AWS, GCP, or Azure.
- 5+ years of experience with debugging, troubleshooting, and the use of observability and diagnostic tools.
- Experience developing or deploying applications using LLMs, RAG pipelines, MCP, and multi-agent frameworks.
Skills
- Go
- Python
- Java
- Distributed systems
- Databases
- Distributed file systems
- Concurrency control
- Consistency models
- API communication
- REST
- CI/CD pipelines
- Kubernetes
- Container ecosystems
- Microservices-based architectures
- AWS
- GCP
- Azure
- Debugging
- Troubleshooting
- Observability tools
- Diagnostic tools
- LLMs
- RAG pipelines
- MCP
- Multi-agent frameworks
Experience Level
- 8+ years of professional experience in software engineering
- 5+ years of experience in the design and implementation of distributed systems
- 2+ years of experience with CI/CD pipelines, Kubernetes, container ecosystems, and microservices-based architectures
- 3+ years of experience working with public cloud providers
- 5+ years of experience with debugging, troubleshooting, and the use of observability and diagnostic tools
Education Level
- Bachelor's degree in computer science, Engineering, or a related field, or equivalent practical experience.
Salary/Compensations
- $166,300.00 - $238,300.00
- $186,900.00 - $307,800.00
- $166,300.00 - $274,100.00
Benefits
- Medical insurance
- Dental insurance
- Vision insurance
- 401(k) plan with a Cisco matching contribution
- Paid parental leave
- Short-term disability coverage
- Long-term disability coverage
- Basic life insurance
- 10 paid holidays per full calendar year
- 1 floating holiday for non-exempt employees
- 1 paid day off for employee’s birthday
- Paid year-end holiday shutdown
- 4 paid days off for personal wellness
- 16 days of paid vacation time per full calendar year (non-exempt)
- Flexible vacation time off program (exempt)
- 80 hours of sick time off provided on hire date and each January 1st thereafter
- Up to 80 hours of unused sick time carried forward
- Additional paid time away may be requested to deal with critical or emergency issues for family members
- Optional 10 paid days per full calendar year to volunteer
- Annual bonuses (non-sales roles)
- Performance-based incentive pay (sales roles)
About the Company
- At Splunk, we’re driven by a bold vision: to make machine data accessible, usable, and valuable to everyone. We’re a team of passionate builders who care deeply about our customers, our technology, and each other’s success. We work hard, move fast, and believe that great outcomes come from collaboration, trust, and curiosity.
- At Cisco, we’re revolutionizing how data and infrastructure connect and protect organizations in the AI era – and beyond. We’ve been innovating fearlessly for 40 years to create solutions that power how humans and technology work together across the physical and digital worlds. These solutions provide customers with unparalleled security, visibility, and insights across the entire digital footprint.
- Fueled by the depth and breadth of our technology, we experiment and create meaningful solutions. Add to that our worldwide network of doers and experts, and you’ll see that the opportunities to grow and build are limitless. We work as a team, collaborating with empathy to make really big things happen on a global scale. Because our solutions are everywhere, our impact is everywhere.
- We are Cisco, and our power starts with you.
