Senior System Software Engineer, Software Defined Networking

NVIDIAIndia, BengaluruRemotefull timeSenior
ActiveVerified 3h ago

Job description

We are looking for a Senior System Software Engineer, Software Defined Networking to design, build, and operate highly performant and scalable SDN solutions for NVIDIA's AI Clouds hosting GPU-accelerated workloads — including hyperscale multi-node training, inference, cloud gaming, and cloud functions. This role spans the full lifecycle of our SDN stack — from designing and developing new control and data plane software to ensuring operational excellence in production through reliability engineering, CI/CD, observability, and incident response.

What you'll be doing:

  • Design and develop next-generation multi-tenant cloud SDN control and data plane software (OVS, OVN, OpenFlow)

  • Build Infrastructure-as-a-Service virtual network orchestration and services using gRPC and REST to support tenant workload security and performance SLAs for BMaaS, VMaaS, and Kubernetes

  • Drive upstream contributions to OVN-Kubernetes and related open-source projects; Develop software for network observability — monitoring, telemetry, intelligent metering, and performance analysis

  • Operate and support OVS-OVN based SDN solutions in large-scale NVIDIA AI Cloud environments

  • Own end-to-end observability for the SDN stack — build and maintain monitoring, alerting, distributed tracing, and dashboarding to ensure real-time insight into network health, performance, and tenant SLAs

  • Design, enhance, and maintain CI/CD pipelines (GitLab) across Linux host networking, OVS, OVN, and Kubernetes CNIs

  • Implement GitOps approaches or related experience for secure, seamless integration with cloud infrastructure; Drive reliability through incident management, resource monitoring, and performance tuning

  • Collaborate with SRE, DevOps, and network engineering teams on production readiness and operational tooling

What we need to see:

  • BS/MS in Computer Science or related technical field, or a comparable blend of education and relevant experience

  • 5+ years of proven experience in software development for large-scale distributed environments

  • Expert-level knowledge of OVN, OVS, OpenFlow, and modern network protocols

  • Strong programming skills in C and Go; advanced scripting in Bash and Python

  • Deep knowledge of Kubernetes, practical experience deploying and supporting CNIs (OVN-Kubernetes)

  • Hands-on experience with Infrastructure-as-Code and deployment tools (Ansible, Terraform, ArgoCD, Flux)

  • Experience designing and operating complex, multi-stage CI/CD pipelines

  • Hands-on experience developing secure, high-performance services using gRPC and REST with TLS and strong authentication

  • Strong knowledge of datacenter routing, switching, and Linux host/VM networking

Ways to stand out from the crowd:

  • Contributions to open-source projects (especially OVS, OVN, OVN-Kubernetes, or other Kubernetes networking projects)

  • Experience with hardware acceleration (GPU, DPU or equivalent experience) for networking

  • Practical experience with major cloud providers (AWS, Azure, GCP) and hybrid/multi-cloud deployments

  • SRE/DevOps top-level expertise — on-call, incident management, operations focused on service reliability targets, production ownership

  • Experience with observability platforms and tools (Prometheus, Grafana, Jaeger, OpenTelemetry, ELK)

Similar jobs

Job DescriptionThe Robotics Engineering Fellow serves as the technical authority and visionary for Vaux. This position defines the long-term strategic technology roadmap, leveraging foundational innovation modeling, mathematical modeling, and advanced automation principles to solve the logistics industry's most complex challenges. Rather than executing defined project tracks, the Robotics Engineering Fellow anticipates disruptive industry trends, conceptualizes next-generation autonomous systems, and establishes enterprise-wide engineering standards. This position works directly with senior executives, orchestrates multi-million dollar high-stakes R&D initiatives, and serves as an industry-wide thought leader to secure Vaux's competitive edge in logistics automation.ResponsibilitiesArchitect the long-term autonomy, control systems, and algorithmic frameworks for Vaux's entire fleet and logistics automation portfolio.Define, manage, and govern the technical interfaces and foundational infrastructure between diverse robotic platforms to ensure seamless current and future cross-compatibility.Serve as the compliance and risk management authority for robot design, system safety, architecture, and regulatory standards across the enterprise.Define and manage the technical interfaces between robotic platforms for current and future development.Cultivate and lead strategic relationships with primary global technology vendors, research universities, and enterprise-level suppliers to accelerate innovation.Meet, consult, and align regularly with Executives and key clients to pitch technological breakthroughs, navigate macroeconomic barriers, and guide company strategy.Actively mentor and grow Senior and Lead Engineers, cultivating a high-performing culture of innovation, deep technical competency, and psychological safety under tight timelines.Bridge the technical gap across business functions, embedding advanced automation solutions smoothly within Vaux's broader operations.Clearly define and manage project success metrics on an ongoing basis even after the completion of projects.Maintain a positive attitude in a highly intense environment.Other duties and projects, as assigned.RequirementsEducation:Master's degree in Engineering, Computer Science, or related field, required.PhD with focus on Autonomous Systems, AI, or advanced Control Theory, preferred.Experience:15+ years of experience as a high-impact Robotics Engineer or equivalent job function.Proven track record in commercial scale robotic systems or hardware-software systems from early-stage conceptual R&D to high-volume deployment.Computer Skills:Expert programming, systems analysis, and design skills, required.In-depth knowledge and working experience of all kinds of sensors used in robotics, required.Advanced Tooling Mastery such as SolidWorks skills and experience working with ROS, required.Additional Requirements:Understanding of organization, behavior, business practices, and inter-connectivity within ArcBest and its subsidiaries, required.Demonstrated leadership capabilities, required.Project Management and excellent organizational skills, preferred.Other DetailsWork Hours:Generally, 8:00 am - 5:00 pm with occasional irregular hours depending on workload.Travel Requirements:Occasional (25%-50%)Compensation:This is a salary position paid biweekly.About Us ArcBest is a multibillion-dollar integrated logistics company that's helped businesses build better supply chains for over a century. With 14,000+ employees across 250 campuses and service centers, we connect owned assets, a broad brokerage network and innovative technology to deliver end-to-end supply chain solutions for customers around the world. Our people make the difference. From customer service and operations to technology, sales and logistics professionals, every employee plays a role in supporting customers and solving complex logistics challenges. It's the kind of people-first culture that's earned recognition as a Best Company to Work For by U.S. News & World Report and one of America's Best Employers for Company Culture by Forbes. At ArcBest, you're part of a team grounded in our core values: Creativity, Integrity, Collaboration, Growth, Excellence and Wellness. Whether you're starting your career or bringing years of experience, you'll find opportunities to learn, grow and make an impact.

Verified Sep 23, 2026Posted 17 days ago