Customer Reliability Engineer – Hypershield

Job not on LinkedIn

🕒 6 days ago

🏄 California – Remote

info

💵 $160.8k - $214.1k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Cisco

Cisco

10,000+ employees

Founded 1984

🔧 Hardware

🔐 Security

🏢 Enterprise

Hardware • Security • Enterprise

Cisco is a multinational technology company that provides networking hardware, software, and services to enterprises, service providers, and governments. It builds routers, switches, optical transceivers, programmable silicon, and edge computing platforms, and offers security, collaboration (Webex), observability, and AI-enabled software and support services to help organizations design, operate, and secure large-scale networks and data centers. Cisco also delivers professional services, training, and cloud-managed solutions to support digital transformation and AI-ready infrastructure.

📋 Description

• Own Hypershield cases escalated from Cisco TAC through resolution, engaging customers directly as needed • Diagnose complex production failures across the N9300 Smart Switch fabric and on-premises Kubernetes controller • Localize faults across switching and forwarding, security services and enforcement, and control-plane layers • Understand customer architectures and configurations to diagnose failures in unfamiliar production environments • Reproduce customer failures, partner with engineering on fixes, and own fixes back to customers • Turn individual cases into systemic improvements including runbooks, diagnostics, knowledge-base content, and engineering feedback • Develop proactive customer-health monitoring, tooling, and reliability practices as the installed base grows

🎯 Requirements

• Bachelor's degree plus 8 years of experience, Master's degree plus 6 years, or equivalent industry experience • Experience supporting enterprise customers in an escalation capacity • Experience diagnosing and resolving complex production incidents under SLA pressure in unfamiliar environments • Production experience operating and troubleshooting Cisco Nexus / NX-OS, or equivalent depth on another major vendor • Experience localizing failures across layered data-center architectures spanning switching/forwarding, services/enforcement, and control-plane domains • Linux command-line operations and production troubleshooting experience • Working exposure to containers or Kubernetes • Network troubleshooting using packet capture and flow-telemetry analysis such as NetFlow/IPFIX • Enterprise virtualization knowledge sufficient to troubleshoot VM-based appliance deployments; vSphere experience • Operational Kubernetes and Helm proficiency, including TLS certificates, service-account authentication, API-server connectivity, service exposure, persistent storage, custom resources, and operators • Working knowledge of VXLAN EVPN fabrics • Network segmentation and firewall policy design, including zone-based or microsegmentation approaches • Familiarity with NetOps/NetSecOps operating models • Experience driving diagnosis and remediation through customer teams in environments without direct access • Experience with NX-OS automation and APIs such as NX-API, NETCONF/RESTCONF, gNMI, or Ansible • Ability to communicate incident status, root cause, and remediation to technical and executive audiences • Cisco Nexus Dashboard familiarity is a plus • CCNP Data Center, CCNP Enterprise, DevNet Professional, CCIE Data Center, CCIE Enterprise, or DevNet Expert is a plus

🏖️ Benefits

• Medical, dental and vision insurance • 401(k) plan with Cisco matching contribution • Paid parental leave • Short- and long-term disability coverage • Basic life insurance • Potential Cisco restricted stock unit grants • 10 paid holidays per full calendar year • 1 floating holiday for non-exempt employees • Paid birthday day off • Paid year-end holiday shutdown • 4 paid personal wellness days • 16 days of paid vacation per full calendar year for non-exempt employees • Flexible vacation time off with no defined limit for eligible exempt employees • 80 hours of sick time off provided on hire date and each January 1st • Up to 80 hours of unused sick time carried forward • Additional paid time away for critical or emergency family issues • Optional 10 paid volunteer days per full calendar year • Annual bonuses for non-sales roles

Apply Now

Similar Jobs

🕒 6 days ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior Site Reliability Engineer improving Akamai's distributed network platform through performance analytics, reliability tuning, monitoring, automation, and troubleshooting.

🇺🇸 United States – Remote

💵 $146.4k - $263.6k / year

💰 Post-IPO Equity on 2001-07

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 6 days ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior SRE ensuring reliability and uptime for Akamai’s dedicated AI hardware infrastructure. Automating provisioning, observability, and incident response across its distributed cloud and edge platform.

🇺🇸 United States – Remote

💵 $121.4k - $218.6k / year

💰 Post-IPO Equity on 2001-07

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 6 days ago

PrizePicks

201 - 500

🎮 Gaming

⚽ Sports

Senior SRE ensuring reliable, scalable infrastructure for PrizePicks’ daily fantasy sports platform. Leading incident response, observability, Kubernetes operations, and reliability improvements.

🇺🇸 United States – Remote

💵 $120k - $175k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 6 days ago

Veeam Software

1001 - 5000

💼 Consulting

📦 Logistics

☁️ SaaS

Senior SRE building reliability engineering for Veeam Data Cloud’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, incident response, and compliance-ready delivery practices.

🇺🇸 United States – Remote

💵 $158.4k - $294.1k / year

💰 $500M Private Equity Round on 2019-01

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 6 days ago

Veeam Software

1001 - 5000

💼 Consulting

📦 Logistics

☁️ SaaS

Site Reliability Engineer building reliability practices for Veeam’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, automation, and incident-response systems in regulated environments.

🇺🇸 United States – Remote

💵 $138.9k - $231.4k / year

💰 $500M Private Equity Round on 2019-01

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info