Moonlite AI
Senior Software Engineer, Network Platform
Chicago, IL or Remote · Senior
Sponsorship not specified$165k-$225kDetected 62 days ago
PythonGoFastAPIKubernetesTerraformLinuxPlatform EngineeringgRPCCybersecurityNetwork SecurityTCP/IPFirewallBGP/OSPFNetwork MonitoringResearchCommunicationProblem Solving
About the role
- You will be foundational to building our software-defined networking (SDN) platform that enables high-performance, isolated networking for distributed computing, model training, inference, and data-intensive workloads.
- Working closely with our network, infrastructure, and product teams, you'll design and implement the network orchestration and provisioning systems that manage DPU-accelerated networking, tenant isolation, and network lifecycle management - enabling researchers and engineers to access enterprise-grade networking with cloud-like simplicity.
- Job Responsibilities:
Responsibilities
- Software-Defined Networking Architecture: Collaborate with infrastructure to design and build scalable SDN orchestration systems leveraging NVIDIA Bluefield-3 DPUs to deliver programmable, high-performance networking for AI workloads with hardware-accelerated forwarding isolation.
- Research Cluster Networking: Design and implement networking systems for research computing environments including Kubernetes and SLURM clusters, enabling high-performance connectivity, optimized network topology for distributed workloads, and seamless integration with cluster orchestration systems.
- Network Provisioning & Lifecycle Management: Implement automated SDN provisioning systems that handle VPC creation, subnet allocation, routing configuration, and network resource lifecycle from deployment through decommissioning.
- DPU Platform Engineering: Develop platform capabilities for managing Bluefield-3 DPUs including SR-IOV virtual function management, OVS offload configuration, network function deployment, and integration with compute orchestration systems.
- Multi-Tenancy & Network Isolation: Build enterprise-grade network isolation using VPCs, VXLAN, and hardware-accelerated forwarding to ensure complete tenant separation while maintaining high-performance connectivity for GPU clusters and distributed workloads.
- High-Performance Networking: Collaborate with infrastructure to optimize network paths for RDMA, RoCE, and GPU-to-GPU communication, ensuring minimal latency and maximum throughput for distributed training and large-scale computational workloads.
- Network APIs & Integration: Develop robust APIs and SDKs for network resource management that integrate seamlessly with compute and storage platforms, enabling programmatic network provisioning and configuration.
- Collaborate with infrastructure to design and build scalable SDN orchestration systems leveraging NVIDIA Bluefield-3 DPUs to deliver programmable, high-performance networking for AI workloads with hardware-accelerated forwarding isolation.
- Design and implement networking systems for research computing environments including Kubernetes and SLURM clusters, enabling high-performance connectivity, optimized network topology for distributed workloads, and seamless integration with cluster orchestration systems.
- Implement automated SDN provisioning systems that handle VPC creation, subnet allocation, routing configuration, and network resource lifecycle from deployment through decommissioning.
Requirements
- Programming Skills: Experience with Go and Python for performance-critical networking components and services is highly valued.
- Linux Networking: Strong experience with Linux networking stack, including network namespaces, iptables/nftables, Open vSwitch, and kernel networking systems.
- DPU & SmartNIC Experience: Familiarity with DPU/SmartNIC architectures (Bluefield, or similar), SR-IOV, hardware offload capabilities, and programmable networking hardware - or strong ability to learn quickly.
- Problem-Solving & Architecture: Demonstrated ability to solve complex networking performance and scalability challenges while balancing pragmatic shipping with good long-term architecture.
- Strong experience with Linux networking stack, including network namespaces, iptables/nftables, Open vSwitch, and kernel networking systems.
- Familiarity with DPU/SmartNIC architectures (Bluefield, or similar), SR-IOV, hardware offload capabilities, and programmable networking hardware - or strong ability to learn quickly.
- Demonstrated ability to solve complex networking performance and scalability challenges while balancing pragmatic shipping with good long-term architecture.
Nice to have
- Hands-On Ownership: As an early engineer, you'll have end-to-end ownership of projects and the autonomy to influence our product and technology direction.
Skills
- Experience with Go and Python for performance-critical networking components and services is highly valued.
- Commitment to Growth: Growth mindset with continuous focus on learning and professional development.
- Background provisioning or managing networking for research computing environments (Kubernetes, SLURM, or HPC clusters)
- Experience with NVIDIA Bluefield DPU programming and DOCA framework
- Background with network function virtualization (NFV) and service function chaining
- Knowledge of Kubernetes networking (CNI plugins, network policies, service mesh)
- Experience building network control planes or SDN controllers
- Familiarity with network automation frameworks and infrastructure-as-code for networking
- Understanding of data center fabric architectures (spine-leaf, CLOS topologies)
- Experience with network security and compliance requirements in regulated industries
- Key Technologies
- Go, Python, NVIDIA Bluefield DPUs, Open vSwitch, VXLAN, SR-IOV, RDMA, RoCE, InfiniBand, BGP, Linux networking, Terraform, FastAPI, gRPC
Compensation
- We offer a competitive total compensation package combining a competitive base salary, startup equity, and industry-leading benefits.
Benefits
- We offer a competitive total compensation package combining a competitive base salary, startup equity, and industry-leading benefits.
- The total compensation range for this role is $165,000 - $225,000, which includes both base salary and equity.
- We provide generous benefits, including a 6% 401(k) match, fully covered health insurance premiums, and other comprehensive offerings to support your well-being and success as we grow together.
Apply directly at Moonlite AI →Create a free account for alerts like thisView Moonlite AI immigration profile
This listing is sourced directly from Moonlite AI's careers page and normalized into a canonical job model.