Senior System Software Engineer, Software-Defined Networking

🕒 September 13

🏄 California, Colorado, +3 more states – Remote

infoinfo

💵 $224k - $356.5k / year

⏰ Full Time

🟠 Senior

🧑‍💻 Full-stack Engineer

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 1%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of NVIDIA

NVIDIA

10,000+ employees

Founded 1993

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Healthcare • Manufacturing • Artificial Intelligence

NVIDIA is a leading technology company specializing in accelerated computing and artificial intelligence. NVIDIA pioneers advancements in graphical processing units (GPUs), cloud computing, data centers, and virtual reality, with a focus on gaming, automotive, healthcare, and robotics industries. The company's innovations, such as NVIDIA Omniverse, transform traditional digital processes by enabling high-fidelity simulations and rendering tasks. Their applications span various industries, from autonomous vehicles using NVIDIA DRIVE to healthcare solutions with NVIDIA Clara, and AI-driven analytics and workflows.

📋 Description

• Advance software-defined networking for global GPU cloud infrastructure supporting AI workloads, cloud gaming, content delivery, and accelerated services • Collaborate with engineers and architects to define, review, and evolve multi-tenant control-plane and data-plane architecture using Open vSwitch, OVN, OpenFlow, and overlay networks • Provide technical leadership and mentorship through design and code reviews, implementation guidance, production-readiness decisions, and engineering-quality improvements • Develop and review production software for Kubernetes networking, network plugins, distributed control planes, Linux host networking, virtual-machine networking, and network automation • Translate product and infrastructure requirements into secure orchestration services, APIs, component boundaries, state models, compatibility strategies, and delivery plans • Drive Open vSwitch and OVN integration across flow behavior, configuration, lifecycle management, upgrades, interoperability, performance, failure recovery, and large-scale production operation • Build automated unit, integration, system, performance, scale, and upgrade tests connected to CI, deployment, and release-qualification workflows • Share the team’s on-call rotation and lead networking escalations with partner teams • Resolve incidents using packet captures, SDN state, telemetry, profiling, controlled experiments, and source debugging, and prevent recurrence • Define and improve reliability, performance, security, and resource-efficiency objectives • Engineer monitoring, telemetry, tracing, and service-level visibility for production networks

🎯 Requirements

• BS or MS in Computer Science, Computer Engineering, or a related field, or equivalent experience • 12+ years of experience designing, implementing, testing, and maintaining production software in both C and Go • Experience with Bash and Python for test, diagnostic, build, or operational automation • Experience developing, integrating, and troubleshooting Open vSwitch and OVN, including OpenFlow and control-plane-to-data-plane behavior • Production Kubernetes networking experience, including container network interfaces, network policy, node and pod traffic paths, network plugins, upgrades, and failure modes • Strong Linux networking fundamentals and practical understanding of IP, TCP, UDP, routing, switching, overlay networks, tunneling, network namespaces, and network policy • Experience architecting distributed systems with secure service APIs, state management, consistency, scalability, compatibility, and failure handling • Experience with automated testing, CI/CD, deployments, upgrades, observability, and performance analysis • Experience supporting production services through a shared on-call rotation and coordinating incident response • Collaborative architecture and technical leadership through written designs, cross-team work, critical reviews, mentoring, and production delivery • Contributions to Open vSwitch, OVN, OVN-Kubernetes, Kubernetes networking, or another relevant open-source project are a plus • Experience with large-scale cloud and accelerated virtualization systems, including SR-IOV, RDMA, SmartNICs, DPUs, NFV, KVM/QEMU, or container runtimes is a plus • Experience building secure, high-performance gRPC or REST services using transport security and strong authentication is a plus • Applied use of agentic AI and AI-assisted software-development tooling is a plus

🏖️ Benefits

• Competitive salaries • Generous benefits package • Equity • Benefits

Apply Now

Similar Jobs

🕒 September 12

PhillyTech (SaaS Talent)

11 - 50

💼 Consulting

📣 Marketing

🎯 Recruiter

Senior Full Stack Engineer building and scaling a custom coaching and personal-development community platform. Owning backend-focused features, integrations, testing, and production reliability.

🇺🇸 United States – Remote

💵 $90 - $120 / hour

⏰ Full Time

🟠 Senior

🧑‍💻 Full-stack Engineer

🕒 September 12

Orlando Informer

1 - 10

🏨 Hospitality

💼 Consulting

🍽️ Food & Beverage

Software Engineer building React Native apps, Node.js APIs, and MySQL systems for Orlando Informer’s theme-park vacation-planning and hospitality business. Integrating commerce platforms and improving guest experiences.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🧑‍💻 Full-stack Engineer

🕒 September 12

MetalBear

11 - 50

☁️ SaaS

🤝 B2B

Engineering Team Lead leading cloud integrations for MetalBear’s mirrord Kubernetes development platform. Building queue splitting, database branching, and related developer infrastructure features.

🇺🇸 United States – Remote

💰 $12.5M Seed Round - MetalBear on 2025-09

⏰ Full Time

🟠 Senior

🧑‍💻 Full-stack Engineer

🕒 September 12

FAR.AI

11 - 50

🤖 Artificial Intelligence

📚 Education

🤝 Non-profit

Tech Lead Manager leading GPU cluster infrastructure for FAR.AI, a nonprofit AI safety research institute. Owning platform architecture, security, operations, and team growth.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🧑‍💻 Full-stack Engineer

🕒 September 12

FAR.AI

11 - 50

🤖 Artificial Intelligence

📚 Education

🤝 Non-profit

Senior software engineer operating Kubernetes GPU clusters for FAR.AI, a nonprofit AI safety research institute. Owning scheduling, storage, reliability, security, and multi-provider infrastructure.

🇺🇸 United States – Remote

💵 $150k - $275k / year

⏰ Full Time

🟠 Senior

🧑‍💻 Full-stack Engineer