Jobs / Software Development Engineer (SDE) / Cisco Hiring Senior Site Reliability Engineer in Bangalore | Hybrid
Cisco Systems logo

Cisco Systems

Cisco Hiring Senior Site Reliability Engineer in Bangalore | Hybrid

location_on Bangalore | Hybrid
work 8 - 10 years Experience
payments Competitive Salary (Not disclosed)
schedule Full Time

Role Overview

Job Overview

Cisco is hiring a Senior Site Reliability Engineer for its Splunk Agent Resilience team in Bangalore. This role focuses on building and operating reliable infrastructure for AI-powered applications and deployment platforms. The engineer will work across Kubernetes, cloud infrastructure, automation, observability, deployment systems, and distributed production environments.

The position has a strong operational engineering focus. You will help maintain production platforms, support customer deployments, improve system resilience, and reduce operational effort through automation. The role also covers cloud and air-gapped customer environments, making it suitable for an experienced SRE, DevOps, Platform Engineer, or Infrastructure Engineer who enjoys solving complex reliability problems.


Key Responsibilities

  • Operate and continuously improve Kubernetes-based production infrastructure and deployment platforms.
  • Manage customer deployments across cloud and air-gapped environments, including installation, upgrades, troubleshooting, and lifecycle activities.
  • Build and improve monitoring, logging, alerting, and observability capabilities for deployment systems.
  • Automate repetitive operational activities and improve platform efficiency, reliability, scalability, and performance.
  • Participate in production incident response, investigate root causes, and implement improvements that reduce recurring reliability issues.
  • Analyze and tune infrastructure components, databases, and services to improve performance and resilience.
  • Develop internal engineering and operational tools using Python and/or Go.
  • Investigate complex production problems involving Kubernetes, networking, storage, and application layers.
  • Manage infrastructure through Infrastructure as Code practices using Terraform or similar tools.
  • Work with software engineering, infrastructure, security, and customer-facing teams to create secure and dependable deployment architectures.
  • Contribute to continuous improvement of deployment processes and operational standards.


Required Skills

Strong experience in Site Reliability Engineering, DevOps, Platform Engineering, Infrastructure Engineering, or a closely related discipline is required. Candidates should have at least 8 years of professional experience and substantial hands-on experience operating Kubernetes in production.

A strong understanding of Kubernetes operations and Helm is important for this position. Candidates should also understand production deployment practices, CI/CD platforms, automation, troubleshooting, and cloud infrastructure.

Programming or scripting ability in Python and/or Go is valuable for developing internal tooling and automating operational workflows. The role also requires strong debugging skills and the ability to investigate issues across multiple layers of distributed systems.

Experience with AWS, GCP, or comparable cloud platforms is required. Candidates should understand how cloud infrastructure supports scalable and reliable applications.


Preferred Skills

Experience improving production availability, reliability, scalability, and operational efficiency is preferred. Familiarity with monitoring, logging, alerting, and broader observability platforms will strengthen a candidate's profile.

Experience with Infrastructure as Code, particularly Terraform, is desirable. Exposure to MLOps is also preferred because the team supports infrastructure and resilience for AI-oriented applications.

Candidates should have a solid understanding of networking concepts such as VPCs, DNS, routing, and load balancing. Experience operating and tuning data infrastructure is useful, including OLTP and OLAP databases, message queues, and object storage systems.

Experience supporting both cloud and air-gapped or on-premises environments is valuable. Strong troubleshooting ability across distributed systems and the ability to collaborate with engineering, infrastructure, security, and customer teams are also important.


Education

The provided job description does not specify a required educational degree or qualification. Candidates should therefore be evaluated primarily against the stated professional experience and technical requirements.


Experience

This is a senior-level opportunity requiring 8 to 10 years of relevant experience. The minimum requirement is 8 years in Site Reliability Engineering, DevOps, Platform Engineering, Infrastructure Engineering, or a related field. At least 3 years of production Kubernetes experience is specified, along with experience using Helm.

The ideal candidate will have a track record of operating production systems, improving reliability, supporting deployment platforms, automating infrastructure operations, and resolving complex incidents.


Required Technologies

Key technologies and technical areas for this role include Kubernetes, Helm, Python, Go, AWS, GCP, Terraform, CI/CD, Infrastructure as Code, observability, monitoring, logging, alerting, cloud infrastructure, networking, storage, databases, message queues, and distributed systems.

Relevant networking concepts include VPCs, DNS, routing, and load balancing. The role also involves OLTP and OLAP databases, object storage, and production services. MLOps experience is preferred.


Soft Skills

The position requires clear communication and effective collaboration across multiple technical and customer-facing groups. Senior SREs should be comfortable explaining operational risks, coordinating incident response, discussing architecture decisions, and working with engineers to improve reliability.

Strong analytical thinking, structured troubleshooting, ownership, and attention to operational detail are important. The ability to work effectively across engineering, infrastructure, security, and customer-facing teams is especially relevant to this role.


Benefits of Working in this Role

This role provides an opportunity to work on production infrastructure supporting AI-oriented applications and deployment environments. It combines software development, cloud infrastructure, Kubernetes operations, automation, observability, and reliability engineering.

Working across cloud and air-gapped environments can also provide broad exposure to different deployment models and complex infrastructure challenges. The position allows experienced engineers to contribute directly to reliability, scalability, security, and operational excellence.


Work Mode

The job is listed as Hybrid. Candidates should be prepared to follow the team's applicable hybrid working expectations in Bangalore.


Location

This Senior Site Reliability Engineer position is based in Bangalore, India.


Who Should Apply

This opportunity is suitable for experienced Site Reliability Engineers, DevOps Engineers, Platform Engineers, Infrastructure Engineers, and cloud operations professionals with 8 to 10 years of relevant experience.

Candidates should have strong production Kubernetes experience, including Helm, and practical knowledge of cloud platforms, CI/CD, automation, Infrastructure as Code, and production troubleshooting. Engineers who can write Python or Go tooling and who have experience with observability, distributed systems, networking, and data infrastructure may be particularly well suited.


Career Growth

Experience in this role can strengthen career progression toward Staff Site Reliability Engineer, Principal SRE, Platform Architect, Infrastructure Architect, DevOps Architect, or engineering leadership positions. Exposure to AI resilience, cloud infrastructure, Kubernetes, automation, and large-scale production operations can broaden an engineer's platform and reliability expertise.


Application Advice

Before applying, make sure your resume clearly highlights your years of SRE, DevOps, Platform Engineering, or Infrastructure Engineering experience. Mention hands-on Kubernetes production experience and specify your experience with Helm, AWS or GCP, CI/CD, Terraform, Python or Go, observability, and incident response.

If you have worked with air-gapped or on-premises deployments, MLOps, distributed systems, databases, message queues, object storage, or advanced networking, include those details. Use concise examples that demonstrate how you improved reliability, scalability, performance, automation, or operational efficiency.

Technical Ecosystem

Eligibility Criteria

school

Education

B.Tech or equivalent degree in Computer Science, Information Technology, or a related field.

work_history

Experience

8 - 10 years

Top Picks for You

smart_toy

SoftoBot

Beta

Your job search assistant

Ask about roles, locations, remote work, or fresher jobs.