
10,000+ employees
Founded 1984
🔧 Hardware
🔐 Security
🏢 Enterprise
Hardware • Security • Enterprise
Cisco is a multinational technology company that provides networking hardware, software, and services to enterprises, service providers, and governments. It builds routers, switches, optical transceivers, programmable silicon, and edge computing platforms, and offers security, collaboration (Webex), observability, and AI-enabled software and support services to help organizations design, operate, and secure large-scale networks and data centers. Cisco also delivers professional services, training, and cloud-managed solutions to support digital transformation and AI-ready infrastructure.
🔥 32 minutes ago
🏄 California, Illinois, +3 more states – Remote
💵 $186.9k - $267.7k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
👻 Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

10,000+ employees
Founded 1984
🔧 Hardware
🔐 Security
🏢 Enterprise
Hardware • Security • Enterprise
Cisco is a multinational technology company that provides networking hardware, software, and services to enterprises, service providers, and governments. It builds routers, switches, optical transceivers, programmable silicon, and edge computing platforms, and offers security, collaboration (Webex), observability, and AI-enabled software and support services to help organizations design, operate, and secure large-scale networks and data centers. Cisco also delivers professional services, training, and cloud-managed solutions to support digital transformation and AI-ready infrastructure.
• Lead feasibility assessment, technical planning, and phased migration of eligible workloads from AWS-hosted Kubernetes clusters to the internal Kubernetes platform • Establish technical direction and staged delivery plans • Support the operation and reliability of specialized Kubernetes workloads with non-standard load-balancing, networking, and traffic requirements • Support the application after migration, including production readiness, incident response, performance improvements, and operational improvements • Solve complex issues across applications, Kubernetes, Linux, networking, containers, and infrastructure • Improve scalability, reliability, security, performance, and operability of Kubernetes-hosted services • Partner with Node Connectivity, firmware, cloud infrastructure, security, SRE, and product teams • Prioritize support and coordinate cross-system changes • Contribute to broader Kubernetes SRE and platform reliability initiatives • Design, implement, and maintain production-quality Go software for distributed, concurrent, and networked systems • Drive operational readiness through SLIs/SLOs, comprehensive testing, on-call support, root-cause analysis, and long-term corrective actions
• 10+ years of professional software, site reliability, or infrastructure engineering experience, including technical leadership of substantial production systems • Experience designing, deploying, and operating large distributed services on Kubernetes • 5+ years of programming experience in Go or a similar systems programming language • Experience supporting production services through incident response, performance analysis, Kubernetes reliability practices, observability, automation, and operational readiness • Knowledge of Linux and networking concepts, including IPv4/IPv6, TCP, routing, DNS, and TLS • Sound judgment in architecture, incident response, prioritization, and technical tradeoffs • Ability to align stakeholders across teams without relying on formal authority • Preferred: Experience operating Kubernetes across AWS, on-premises, or hybrid environments using AWS/EKS, Docker, Kustomize, GitLab CI/CD, or similar deployment systems • Preferred: Experience with Kubernetes networking, ingress, load balancing, service discovery, and traffic management for high-throughput or geographically distributed services • Preferred: Experience with gRPC, Protocol Buffers, mutual TLS, PKI, VPNs, tunneling, or network security • Preferred: Experience with OpenTelemetry, Prometheus, Datadog, or similar observability tools, distributed routing, packet processing, performance optimization, or failure testing
• Medical, dental and vision insurance • 401(k) plan with a Cisco matching contribution • Paid parental leave • Short- and long-term disability coverage • Basic life insurance • Cisco restricted stock unit grants may be available, subject to eligibility and continued employment • 10 paid holidays per full calendar year • 1 floating holiday for non-exempt employees • Paid birthday day off • Paid year-end holiday shutdown • 4 paid personal wellness days • 16 days of paid vacation per full calendar year for non-exempt employees • Flexible vacation time off program with no defined limit for eligible exempt employees • 80 hours of sick time on hire date and each January 1st thereafter • Up to 80 hours of unused sick time carried forward • Additional paid time away for critical or emergency family issues • Optional 10 paid volunteer days per full calendar year • Annual bonuses for non-sales roles, subject to Cisco policies • Incentive compensation for sales-plan employees, subject to applicable Cisco plan
Apply Now🔥 33 minutes ago
Site Reliability Engineer maintaining reliability, performance, and observability for Databento’s next-generation financial market-data platform. Improving deployments, incident response, and backend infrastructure.
🇺🇸 United States – Remote
💰 $24.3M Series A on 2021-11
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🦅 H1B Visa Sponsor
🔥 1 hour ago
Senior SRE building Python, AWS, and Terraform reliability solutions for Peraton’s national security missions. Improving cloud platform reliability through automation, observability, and incident management.
🔥 9 hours ago
DevOps Process Analyst optimizing development and operations workflows for SSI Group, a healthcare software company. Improving processes, reporting, risk management, and cross-functional delivery.
🔥 10 hours ago
Senior SRE improving reliability, observability, and cloud infrastructure for CentralReach’s autism and IDD care software platforms. Driving SLO adoption, incident response, automation, and production performance.
🇺🇸 United States – Remote
💵 $160k - $180k / year
💰 Private equity on 2018-03
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 10 hours ago
Site Reliability Engineer operating AWS and Snowflake infrastructure for CentralReach’s autism and IDD care software. Improving reliability, connectivity, deployments, and incident response across data platforms.
🇺🇸 United States – Remote
💵 $135k - $160k / year
💰 Private equity on 2018-03
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)