← Back to Jobs
NVIDIA logo
1 week agoworkday

Senior System Software Engineer, Software Defined Networking

NVIDIA / Computer Hardware Manufacturing
Bengaluru, Karnataka, IndiaOn-siteFull Time
Apply on company site
AI Summary

Design and develop next-generation multi-tenant cloud SDN control and data plane software for NVIDIA's AI clouds. Manage the full lifecycle of the SDN stack, including observability, reliability engineering, and CI/CD pipeline maintenance.

Requires a BS/MS in Computer Science or equivalent and at least 5 years of experience in large-scale distributed software development. Candidates must possess expert-level knowledge of OVN, OVS, and modern network protocols, along with strong programming skills in C and Go.

Job Details
Job Type: FULL TIME
Visa Sponsorship: No
Education: bachelor degree, postgraduate degree
Work Arrangement: On-site
Experience Level: 5 - 10 years
Hours: 40 hrs/week
Language: English
Key Skills
Software Defined NetworkingOVSOVNOpenFlowCGoKubernetesBashPythongRPCRESTCI/CDGitLabAnsibleTerraformArgoCD
Insider connections @NVIDIA
Members only

Find people at NVIDIA who may share hiring insights or referrals for this role.

Meet the people behind the opportunity

Create a free account to search for hiring managers and potential referral contacts.

No contact data is shown on this public page.

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response.
 

What you'll be doing:

  • Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow)

  • Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes

  • Drive upstream contributions to OVN-Kubernetes and related open-source projects; Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis

  • Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments

  • Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs

  • Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs

  • Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure; Drive reliability through incident management, resource monitoring, and performance tuning

  • Collaborate with SRE, DevOps, and network engineering teams on production readiness and operational tooling

What we need to see:

  • BS/MS in Computer Science or related technical field, or a comparable blend of education and relevant experience

  • 5+ years of proven experience in software development for large-scale distributed environments

  • Expert-level knowledge of OVN, OVS, OpenFlow, and modern network protocols

  • Strong programming skills in C and Go; advanced scripting in Bash and Python

  • Deep knowledge of Kubernetes, practical experience deploying and supporting CNIs (OVN-Kubernetes)

  • Hands-on experience with Infrastructure-as-Code and deployment tools (Ansible, Terraform, ArgoCD, Flux)

  • Experience designing and operating complex, multi-stage CI/CD pipelines

  • Hands-on experience developing secure, high-performance services using gRPC and REST with TLS and strong authentication

  • Strong knowledge of datacenter routing, switching, and Linux host/VM networking

Ways to stand out from the crowd:

  • Contributions to open-source projects (especially OVS, OVN, OVN-Kubernetes, or other Kubernetes networking projects)

  • Experience with hardware acceleration (GPU, DPU or equivalent experience) for networking

  • Practical experience with major cloud providers (AWS, Azure, GCP) and hybrid/multi-cloud deployments

  • SRE/DevOps top-level expertise — on-call, incident management, operations focused on service reliability targets, production ownership

  • Experience with observability platforms and tools (Prometheus, Grafana, Jaeger, OpenTelemetry, ELK)

About NVIDIA

NVIDIA, a pioneer in accelerated computing, revolutionized graphics with the GPU and is at the forefront of AI, gaming, and data-center innovations.

Computer Hardware Manufacturing
51,630 employees
Santa Clara, CA
Most staff in USA
$29B raised
Employee ratings
From Glassdoor · 7.1k reviews · View profile
Diversity & inclusion4.4
Compensation & benefits4.4
Culture & values4.4
Career opportunities4.3
Senior management4.2
Work–life balance4.0
90%
would recommend to a friend
90%
positive business outlook
Categories
Higher EducationGamingRoboticsVirtual RealityFoundational AISoftware
Specialties
GPU-accelerated computingartificial intelligencedeep learningvirtual realitygamingself-driving carssupercomputingroboticsvirtualizationparallel computingprofessional graphicsand automotive technology