NVIDIA
Posted 15h ago

Senior System Software Engineer, Software Defined Networking

NVIDIA
Taipei, Taipei City, Taiwan
OnsiteFull Time
Responsibilities
  • designing SDN software
  • building network orchestration
  • operating SDN solutions
Requirements
  • BS/MS in computer science or related field
  • 5+ years in large-scale distributed software development
  • Expert OVN/OVS/OpenFlow knowledge
  • Strong C/Go skills
  • Python/Bash scripting
  • Kubernetes and cloud networking expertise
Technical tools mentioned
OVSOVNOpenFlowgRPCRESTKubernetesGitLabLinuxAnsibleTerraformArgoCDFluxCGoBashPythonTLSAWSAzureGCPPrometheusGrafanaJaegerOpenTelemetryELK

Job description

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response.
 

What you'll be doing:

  • Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow)

  • Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes

  • Drive upstream contributions to OVN-Kubernetes and related open-source projects; Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis

  • Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments

  • Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs

  • Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs

  • Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure; Drive reliability through incident management, resource monitoring, and performance tuning

  • Collaborate with SRE, DevOps, and network engineering teams on production readiness and operational tooling

What we need to see:

  • BS/MS in Computer Science or related technical field, or a comparable blend of education and relevant experience

  • 5+ years of proven experience in software development for large-scale distributed environments

  • Expert-level knowledge of OVN, OVS, OpenFlow, and modern network protocols

  • Strong programming skills in C and Go; advanced scripting in Bash and Python

  • Deep knowledge of Kubernetes, practical experience deploying and supporting CNIs (OVN-Kubernetes)

  • Hands-on experience with Infrastructure-as-Code and deployment tools (Ansible, Terraform, ArgoCD, Flux)

  • Experience designing and operating complex, multi-stage CI/CD pipelines

  • Hands-on experience developing secure, high-performance services using gRPC and REST with TLS and strong authentication

  • Strong knowledge of datacenter routing, switching, and Linux host/VM networking

Ways to stand out from the crowd:

  • Contributions to open-source projects (especially OVS, OVN, OVN-Kubernetes, or other Kubernetes networking projects)

  • Experience with hardware acceleration (GPU, DPU or equivalent experience) for networking

  • Practical experience with major cloud providers (AWS, Azure, GCP) and hybrid/multi-cloud deployments

  • SRE/DevOps top-level expertise — on-call, incident management, operations focused on service reliability targets, production ownership

  • Experience with observability platforms and tools (Prometheus, Grafana, Jaeger, OpenTelemetry, ELK)

About NVIDIA

Designs GPU-accelerated computing and artificial intelligence hardware.

Similar jobs

System Software Engineer roles near Taipei, Taipei City
14h
Save
Mark Applied
Hide
System Software Engineer – GPU and SOC (2027 RDSS Intern)
Taipei, Taipei City, Taiwan
OnsiteFull Time, Internship
NVIDIA
NVIDIANASDAQ: NVDA: Designs graphics processing units and artificial intelligence hardware.
Current BS, MS, or PhD student in computer engineering, computer science, or related field; strong C/C++, low-level driver, x86/ARM/SoC, kernel, and system-level debugging experience.
C, C++, Linux, Android, Chrome, Windows, RTOS, x86, ARM, SoC
1d
Save
Mark Applied
Hide
System Software Engineer (Kernal)
Spring or Taipei or Houston
$131k-$205k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Produces personal computers, printers, and related digital imaging products.
7+ YOEBachelor's degree in a relevant engineering or computer science field and 7–10 years of experience, preferably developing Windows/Linux drivers. Requires C, firmware debugging, protocols, and source control expertise.
C, Python, GitHub, JTAG, SWD, UART, I2C, SPI, RTOS, Windows, Linux
1d
Save
Mark Applied
Hide
System Software Engineer (Kernal)
Spring or Taipei
$131k-$205k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufacturer of personal computers, printers, and imaging devices.
7+ YOEBachelor's degree in a relevant engineering or computer science field and 7–10 years of experience, preferably developing Windows/Linux drivers. Requires C, embedded debugging, hardware protocols, and source control experience.
Windows, Linux, C, Python, GitHub, JTAG, SWD, UART, I2C, SPI, RTOS, CI/CD
1d
Save
Mark Applied
Hide
System Software Engineer (Kernal)
Spring or Taipei or Houston
$131k-$205k/yr OnsiteFull Time
HP
HPNYSE: HPQ: Manufactures personal computers, printers, and 3D printing hardware.
7+ YOEBachelor's degree in a related engineering or computer science field and 7–10 years of experience, preferably developing Windows/Linux drivers. Requires C, firmware debugging, hardware protocols, and systems design expertise.
Windows, Linux, C, Python, JTAG, SWD, UART, I2C, SPI, RTOS, GitHub, CI/CD, oscilloscopes
2mo
Save
Mark Applied
Hide
Staff System Software Engineer
Hsinchu, Hsinchu, Taiwan
OnsiteFull Time
SiFive
SiFive: Designs and licenses high-performance RISC-V processor intellectual property.
5+ YOE5+ years developing architecture-level code or device drivers in C for multiprocessor systems; expert knowledge of PCIe, Ethernet, CXL; experience with Linux kernel/upstream, ACPI/UEFI/edk2 desired; debugging with GDB/JTAG/OpenOCD; git, Makefile, GNU toolchain, shell scripting.
Linux kernel, Linux, OpenSBI, u-boot, Yocto, OpenEmbedded, PCIe, Ethernet, CXL, ACPI, UEFI, edk2, GDB, JTAG, OpenOCD, git, Makefile, GNU toolchain, shell scripting, C, device drivers, virtualization, IOMMUs
3mo
Save
Mark Applied
Hide
System software engineer_Hsinchu
Hsinchu, Hsinchu, Taiwan
OnsiteFull Time
MediaTek
MediaTekTaiwan Stock Exchange: 2454: Designs and develops system-on-chip solutions for electronic devices.
0+ YOEExperience in C/C++ and ARM Linux embedded development; familiarity with embedded peripherals, multimedia, chip verification, and signal processing; passion for learning.
C, C++, ARM Linux
9mo
Save
Mark Applied
Hide
System Software Engineer (Taiwan)
Taipei, Taiwan
OnsiteFull Time
Etched
Etched: Designs specialized AI chips optimized for transformer architectures.
Develop, integrate, and debug firmware, boot processes, drivers, and system software for server platforms; focus on BIOS/BMC, security, performance, and data center orchestration.
C/C++, Linux, Git, Kubernetes, Docker, EFI, UEFI, NetBoot, BIOS, BMC, Kernel-mode driver development, Hardware logs
1y
Save
Mark Applied
Hide
Sr. System Software Engineer, Rack Solution_TC25273
Taoyuan City, Taoyuan, Taiwan
OnsiteFull Time
Supermicro
SupermicroNASDAQ: SMCI: Designs and manufactures high-performance server and storage systems.
3+ YOEBachelor's or master's in computer science or related field; 3+ years in AI/ML and Linux/networking debugging or testing; experience with frameworks, cloud, containers, clusters, and scripting.
Linux, MLPerf, LLM, RAG, PyTorch, TensorFlow, ONNX, DevOps, Docker, Containers, Kubernetes, Slurm, AWS, Azure, GCP, OpenStack, OpenShift, Windows, CUDA, oneAPI, ROCm, Intel, AMD, NVIDIA