Customer Reliability Engineer – Hypershield

Vaga não está no LinkedIn

🕒 6 dias atrás

🏄 California – Remoto

info

💵 $160.800 - $214.100 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Cisco

Cisco

10.000+ funcionários

Fundada em 1984

🔧 Hardware

🔐 Segurança

🏢 Corporativo

Hardware • Security • Enterprise

A Cisco é uma empresa multinacional de tecnologia que fornece hardware de rede, software e serviços para empresas, provedores de serviços e governos. A empresa desenvolve roteadores, switches, transceptores ópticos, silício programável e plataformas de computação de borda, além de oferecer segurança, colaboração (Webex), observabilidade e software habilitado por IA e serviços de suporte para ajudar organizações a projetar, operar e proteger redes e data centers em larga escala. A Cisco também oferece serviços profissionais, treinamentos e soluções gerenciadas em nuvem para apoiar a transformação digital e infraestrutura preparada para IA.

Descrição

• Own Hypershield cases escalated from Cisco TAC through resolution, engaging customers directly as needed • Diagnose complex production failures across the N9300 Smart Switch fabric and on-premises Kubernetes controller • Localize faults across switching and forwarding, security services and enforcement, and control-plane layers • Understand customer architectures and configurations to diagnose failures in unfamiliar production environments • Reproduce customer failures, partner with engineering on fixes, and own fixes back to customers • Turn individual cases into systemic improvements including runbooks, diagnostics, knowledge-base content, and engineering feedback • Develop proactive customer-health monitoring, tooling, and reliability practices as the installed base grows

🎯 Requisitos

• Bachelor's degree plus 8 years of experience, Master's degree plus 6 years, or equivalent industry experience • Experience supporting enterprise customers in an escalation capacity • Experience diagnosing and resolving complex production incidents under SLA pressure in unfamiliar environments • Production experience operating and troubleshooting Cisco Nexus / NX-OS, or equivalent depth on another major vendor • Experience localizing failures across layered data-center architectures spanning switching/forwarding, services/enforcement, and control-plane domains • Linux command-line operations and production troubleshooting experience • Working exposure to containers or Kubernetes • Network troubleshooting using packet capture and flow-telemetry analysis such as NetFlow/IPFIX • Enterprise virtualization knowledge sufficient to troubleshoot VM-based appliance deployments; vSphere experience • Operational Kubernetes and Helm proficiency, including TLS certificates, service-account authentication, API-server connectivity, service exposure, persistent storage, custom resources, and operators • Working knowledge of VXLAN EVPN fabrics • Network segmentation and firewall policy design, including zone-based or microsegmentation approaches • Familiarity with NetOps/NetSecOps operating models • Experience driving diagnosis and remediation through customer teams in environments without direct access • Experience with NX-OS automation and APIs such as NX-API, NETCONF/RESTCONF, gNMI, or Ansible • Ability to communicate incident status, root cause, and remediation to technical and executive audiences • Cisco Nexus Dashboard familiarity is a plus • CCNP Data Center, CCNP Enterprise, DevNet Professional, CCIE Data Center, CCIE Enterprise, or DevNet Expert is a plus

🏖️ Benefícios

• Medical, dental and vision insurance • 401(k) plan with Cisco matching contribution • Paid parental leave • Short- and long-term disability coverage • Basic life insurance • Potential Cisco restricted stock unit grants • 10 paid holidays per full calendar year • 1 floating holiday for non-exempt employees • Paid birthday day off • Paid year-end holiday shutdown • 4 paid personal wellness days • 16 days of paid vacation per full calendar year for non-exempt employees • Flexible vacation time off with no defined limit for eligible exempt employees • 80 hours of sick time off provided on hire date and each January 1st • Up to 80 hours of unused sick time carried forward • Additional paid time away for critical or emergency family issues • Optional 10 paid volunteer days per full calendar year • Annual bonuses for non-sales roles

Candidatar-se

Vagas Similares

🕒 6 dias atrás

Akamai Technologies

5001 - 10000

🔒 Cibersegurança

Senior Site Reliability Engineer improving Akamai's distributed network platform through performance analytics, reliability tuning, monitoring, automation, and troubleshooting.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $146.400 - $263.600 / ano

💰 Post-IPO Equity em 2001-07

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Akamai Technologies

5001 - 10000

🔒 Cibersegurança

Senior SRE ensuring reliability and uptime for Akamai’s dedicated AI hardware infrastructure. Automating provisioning, observability, and incident response across its distributed cloud and edge platform.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $121.400 - $218.600 / ano

💰 Post-IPO Equity em 2001-07

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

PrizePicks

201 - 500

🎮 Jogos

⚽ Esportes

Senior SRE ensuring reliable, scalable infrastructure for PrizePicks’ daily fantasy sports platform. Leading incident response, observability, Kubernetes operations, and reliability improvements.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $175.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 4

Veeam Software

1001 - 5000

💼 Consultoria

📦 Logística

☁️ SaaS

Senior SRE building reliability engineering for Veeam Data Cloud’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, incident response, and compliance-ready delivery practices.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $158.400 - $294.100 / ano

💰 $500.000.000 Private Equity Round em 2019-01

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 4

Veeam Software

1001 - 5000

💼 Consultoria

📦 Logística

☁️ SaaS

Site Reliability Engineer building reliability practices for Veeam’s Government and Sovereign Cloud SaaS platform. Designing Azure infrastructure, observability, automation, and incident-response systems in regulated environments.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $138.900 - $231.400 / ano

💰 $500.000.000 Private Equity Round em 2019-01

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório