Senior Platform Engineer, Network Infrastructure

🕒 August 6

🇮🇳 India – Remote

⏰ Full Time

🟠 Senior

🛜 Network Engineer / Network Administrator

👻 Ghost score 11%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Healthcare • Manufacturing • Artificial Intelligence

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• Design, build, and operate the Kubernetes platform powering GNI network automation, telemetry, and operations across data center, colocation, and cloud environments • Own lifecycle management for GNI Kubernetes environments, including cluster onboarding, upgrades, capacity, availability, and recovery • Develop production-quality software and automation for cluster provisioning, validation, upgrades, remediation, and safe multi-cluster GitOps delivery • Provide production support for network services hosted on the platform, collaborating with Network Automation and service teams • Diagnose complex Kubernetes platform and hosted-service failures involving control-plane health, cluster networking, storage, scheduling, workload placement, and multi-cluster dependencies • Drive issues from initial signal through verified resolution • Define production-readiness and observability standards, including health signals, capacity, alerts, runbooks, and recovery • Participate in CFR’s production on-call rotation, including scheduled after-hours and weekend coverage • Lead incident response and recovery and drive corrective actions to completion

🎯 Requirements

• Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent experience • 8+ years of experience building or operating production Kubernetes platforms, network infrastructure, or distributed systems • Deep experience with Kubernetes at scale, including cluster lifecycle, upgrades, networking, storage, and recovery • Proficiency in at least one general-purpose programming language, such as Go or Python • Experience with GitOps, infrastructure as code, CI/CD, and automated production delivery • Experience deploying and supporting network automation or telemetry services on Kubernetes • Experience with production on-call, incident response, root-cause analysis, and driving corrective actions to completion • Strong knowledge of IP routing, data center fabrics, and cloud networking is a plus • Experience designing and operating large, multi-region Kubernetes fleets, including fleet-wide upgrades and recovery • Hands-on experience with Cluster API (CAPI) and Metal3 for bare-metal provisioning, cluster lifecycle, machine remediation, and upgrades • Experience building Kubernetes controllers or operators in Go using custom resources and reconciliation patterns • Experience designing or operating network automation and telemetry services on Kubernetes at global scale • Contributions to Cluster API, Metal3, or other open-source Kubernetes infrastructure projects

Apply Now

Similar Jobs

🕒 July 31

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Network Engineer for a global IT services company managing large-scale SaaS production networks. Collaborating with teams to optimize and troubleshoot network connectivity issues in mission-critical environments.

Firewalls

iOS

Switching

🕒 July 31

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Network Engineer maintaining and troubleshooting network components in a large-scale SaaS environment. Collaborating with teams to ensure network optimization and security.

Firewalls

iOS

Switching

🕒 July 31

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Network Engineer maintaining and troubleshooting Cisco-based networks for a global IT services company. Collaborating with teams to enhance network performance and reliability.

Firewalls

iOS

Switching

🕒 July 31

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Network Engineer at Miratech optimizing and troubleshooting enterprise networks in a global SaaS environment. Collaborate with teams and enhance network performance for large-scale applications.

Firewalls

iOS

Switching

🕒 July 31

Miratech

501 - 1000

🤝 B2B

💼 Consulting

☁️ SaaS

Senior Network Engineer managing large-scale SaaS production networks with expertise in IP networking. Collaborating with global teams to optimize data-center network environments.

Firewalls

iOS

Switching