Senior Platform Engineer, Network Infrastructure

🕒 July 16

🏄 California, Illinois, +1 more states – Remote

infoinfo

💵 $176k - $276k / year

⏰ Full Time

🟠 Senior

🛜 Network Engineer / Network Administrator

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 5%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Healthcare • Manufacturing • Artificial Intelligence

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• Design, build, and operate the Kubernetes platform that powers GNI network automation, telemetry, and operations across data center, colocation, and cloud environments. • Own the lifecycle management for GNI Kubernetes environments, including cluster onboarding, upgrades, capacity, availability, and recovery. • Develop production-quality software and automation for cluster provisioning, validation, upgrades, remediation, and safe multi-cluster delivery through GitOps. • Provide production support for network services hosted on the platform, working with Network Automation and service teams that retain ownership of application architecture, code, and features. • Diagnose complex Kubernetes platform and hosted-service failures involving control-plane health, cluster networking, storage, scheduling, workload placement, and multi-cluster dependencies. • Drive issues from initial signal through verified resolution. • Define production-readiness and observability standards for the platform and hosted network services, including health signals, capacity, alerts, runbooks, and recovery. • Participate in CFR’s production on-call rotation, including scheduled after-hours and weekend coverage. • Lead incident response and recovery, then drive corrective actions to completion.

🎯 Requirements

• Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience. • 8+ years of experience building or operating production Kubernetes platforms, network infrastructure, or distributed systems. • Deep experience with Kubernetes at scale, including cluster lifecycle, upgrades, networking, storage, and recovery. • Proficiency in at least one general-purpose programming language, such as Go or Python. • Experience with GitOps, infrastructure as code, CI/CD, and automated production delivery. • Experience deploying and supporting network automation or telemetry services on Kubernetes. • Experience with production on-call, incident response, root-cause analysis, and driving corrective actions to completion.

🏖️ Benefits

• Health insurance • Retirement plans • Paid time off • Flexible work arrangements • Professional development • Stock options • Equity

Apply Now

Similar Jobs

🕒 July 15

SysGroup

51 - 200

🤝 B2B

🔒 Cybersecurity

Network Operations Engineer responsible for design, implementation, and monitoring of network connectivity services. Ensuring secure and reliable infrastructure within SysGroup's client environments.

🕒 July 15

GARUD

11 - 50

🎖️ Defense

🏛️ Government

🚀 Aerospace

Network Engineer supporting a federal program management office in technology deployment and operations. Requires advanced networking skills and overall system knowledge with significant experience in security and compliance.

🕒 July 15

OneStream Software

1001 - 5000

💸 Finance

🏢 Enterprise

Senior Cloud DevOps Network Engineer at OneStream responsible for Azure cloud architecture and compliance with FedRAMP standards. Leading automation efforts and team collaboration for secure cloud environments.

🕒 July 10

Red River

501 - 1000

💼 Consulting

📦 Logistics

Sr. Network Engineer delivering complex multi-site networking solutions for enterprise customers. Leading SD-WAN deployments and providing guidance to junior engineers.

🕒 July 10

Data Concepts

201 - 500

💼 Consulting

📣 Marketing

📦 Logistics

Network Engineer responsible for design, installation, and security of enterprise Cisco networks. Must have CCNP and experience with routing, switching, and security protocols.