Cisco
Cisco
Cisco is an enterprise technology company focused on networking, security, and digital infrastructure. Its portfolio includes routers, switches, optical transceivers, programmable silicon, edge computing platforms, Webex collaboration tools, observability products, AI-enabled software, and cloud-managed services. Cisco serves enterprises, service providers, and governments, supporting the design, operation, and security of large-scale networks and data centers through technology, professional services, training, and ongoing support.

Senior Site Reliability Engineer, FedRAMP – Cisco

Join Cisco as a Senior Site Reliability Engineer supporting Webex for Government cloud services. Lead AWS and Kubernetes reliability, incident response, automation, security, and FedRAMP operations.

Description

  • Run resilient AWS and Kubernetes microservices for the Webex for Government environment
  • Monitor production health and investigate complex issues across a 24/7 operating environment
  • Restore services promptly during production disruptions
  • Automate recurring operational tasks and service provisioning
  • Strengthen deployment safety, efficiency, and consistency
  • Support access management, logging, system hardening, vulnerability remediation, and FedRAMP monitoring
  • Partner with engineering, security, and compliance teams to reduce operational risk and improve reliability
  • Develop documentation, runbooks, operating procedures, and incident reports
  • Contribute to incident response, root-cause analysis, capacity planning, and ongoing production improvements
  • Take part in blameless post-incident reviews, corrective actions, and engineering changes

Requirements

  • Bachelor’s degree in computer science, engineering, or a related discipline with 7+ years of relevant experience, equivalent practical experience, or a master’s degree with 4 years of experience
  • At least 4 years of programming experience in Go, Java, Python, or a comparable language
  • Background building, testing, troubleshooting, and maintaining production services and operational automation
  • Experience with Git, CI/CD, Linux, distributed systems, observability, production debugging, and SRE methods
  • Hands-on experience with monitoring, incident response, reliability improvements, and safe deployment and rollback practices
  • Ability to diagnose and resolve highly complex production issues beyond Tier 1 and Tier 2 support
  • Experience developing and maintaining Docker-based containerized applications
  • Experience creating multi-stage Dockerfiles
  • Understanding of horizontally scalable microservice architectures
  • Knowledge of Kubernetes for production container workloads
  • Experience in FedRAMP, government cloud, or another regulated environment is preferred
  • Experience with Prometheus, Grafana, CloudWatch, CloudTrail, Elastic Stack, or Splunk is preferred
  • Experience with GitLab, Jenkins, or comparable CI/CD platforms is preferred
  • Experience with highly available, multi-region distributed systems and disaster recovery is preferred
  • Ability to produce runbooks, operational procedures, and audit-ready evidence
  • Experience conducting blameless post-incident reviews, root-cause analysis, corrective actions, and preventive engineering changes is preferred

Benefits

  • Medical, dental, and vision insurance
  • 401(k) plan with Cisco matching contributions
  • Paid parental leave
  • Short- and long-term disability coverage
  • Basic life insurance
  • Cisco restricted stock unit grants may be available and vest with continued employment
  • Ten paid holidays per full calendar year
  • One floating holiday for non-exempt employees
  • Paid employee birthday off
  • Paid year-end holiday shutdown
  • Four paid personal wellness days
  • Sixteen paid vacation days per full calendar year for non-exempt employees
  • Flexible vacation program with no defined limit for eligible exempt employees
  • Eighty hours of sick leave provided on hire and each January 1
  • Up to 80 hours of unused sick leave may be carried forward
  • Additional paid leave for critical or emergency family matters
  • Optional ten paid volunteer days per full calendar year
  • Annual bonuses for non-sales roles, subject to Cisco policies
  • Performance-based incentive compensation for employees on sales plans

Related Jobs

The Partner Companies

Vice President of Information Technology — San Jose, CA (Hybrid)

The Partner Companies
501 – 1,000 Employees
AerospaceDefenseManufacturing

Lead enterprise IT and AI strategy across multi-site manufacturing operations, spanning infrastructure, cybersecurity, ERP, and digital transformation. Drive compliance, systems integration, governance, and operational technology performance.

Open