Unlock Full Resume Report

This position is no longer accepting applications

Positions open for more than 30 days are automatically closed and marked as expired

Don't let one closed door slow you down, here's your next move:

May 22, 2026

Senior Software Engineer, Network Platform

Senior • On-site

165,000 - 225,000 USD/yr

Chicago, IL

Moonlite delivers high-performance AI infrastructure for organizations running intensive computational research, large-scale model training, and demanding data processing workloads. We provide infrastructure deployed in our facilities or co-located in yours, delivering flexible on-demand or reserved compute that feels like an extension of your existing data center. Our team of AI infrastructure specialists combines bare-metal performance with cloud-native operational simplicity, enabling research teams and enterprises to deploy demanding AI workloads with enterprise-grade reliability and compliance.

Your Role:

You will be foundational to building our software-defined networking (SDN) platform that enables high-performance, isolated networking for distributed computing, model training, inference, and data-intensive workloads. Working closely with our network, infrastructure, and product teams, you'll design and implement the network orchestration and provisioning systems that manage DPU-accelerated networking, tenant isolation, and network lifecycle management – enabling researchers and engineers to access enterprise-grade networking with cloud-like simplicity.

Job Responsibilities:

  • Software-Defined Networking Architecture: Collaborate with infrastructure to design and build scalable SDN orchestration systems leveraging NVIDIA Bluefield-3 DPUs to deliver programmable, high-performance networking for AI workloads with hardware-accelerated forwarding isolation.
  • Research Cluster Networking: Design and implement networking systems for research computing environments including Kubernetes and SLURM clusters, enabling high-performance connectivity, optimized network topology for distributed workloads, and seamless integration with cluster orchestration systems.
  • Network Provisioning & Lifecycle Management: Implement automated SDN provisioning systems that handle VPC creation, subnet allocation, routing configuration, and network resource lifecycle from deployment through decommissioning.
  • DPU Platform Engineering: Develop platform capabilities for managing Bluefield-3 DPUs including SR-IOV virtual function management, OVS offload configuration, network function deployment, and integration with compute orchestration systems.
  • Multi-Tenancy & Network Isolation: Build enterprise-grade network isolation using VPCs, VXLAN, and hardware-accelerated forwarding to ensure complete tenant separation while maintaining high-performance connectivity for GPU clusters and distributed workloads.
  • High-Performance Networking: Collaborate with infrastructure to optimize network paths for RDMA, RoCE, and GPU-to-GPU communication, ensuring minimal latency and maximum throughput for distributed training and large-scale computational workloads.
  • Network APIs & Integration: Develop robust APIs and SDKs for network resource management that integrate seamlessly with compute and storage platforms, enabling programmatic network provisioning and configuration.
  • Network Observability: Implement comprehensive network monitoring, telemetry, and troubleshooting systems that provide visibility into network performance, utilization, and tenant traffic patterns.Security & Policy Management: Build platform network security features including security groups, firewall rules, and policy enforcement that protect tenant workloads while enabling flexible network configuration.

Requirements:

  • Experience: 5+ years in software engineering with proven experience building network platforms, SDN systems, or network automation for production environments.
  • Kubernetes Networking & Container Orchestration: Strong familiarity with Kubernetes networking architecture, CNI plugins, service networking, and network policies. Understanding of pod networking, services, ingress, and how Kubernetes manages network resources.
  • Networking Expertise: Deep understanding of networking fundamentals including TCP/IP, VLANs, VXLAN, BGP, OSPF, routing protocols, and data center network architectures.Software-Defined Networking: Background in SDN concepts, network virtualization, overlay networks, and programmable networking technologies.
  • Programming Skills: Experience with Go and Python for performance-critical networking components and services is highly valued.
  • Linux Networking: Strong experience with Linux networking stack, including network namespaces, iptables/nftables, Open vSwitch, and kernel networking systems.
  • DPU & SmartNIC Experience: Familiarity with DPU/SmartNIC architectures (Bluefield, or similar), SR-IOV, hardware offload capabilities, and programmable networking hardware – or strong ability to learn quickly.
  • High-Performance Networking: Understanding of RDMA, RoCE, Infiniband, and low-latency networking requirements for distributed computing and GPU workloads.
  • Problem-Solving & Architecture: Demonstrated ability to solve complex networking performance and scalability challenges while balancing pragmatic shipping with good long-term architecture.
  • Autonomy & Communication: Comfortable navigating ambiguity, defining requirements collaboratively, and communicating technical decisions through clear documentation.
  • Commitment to Growth: Growth mindset with continuous focus on learning and professional development.

Preferred Qualifications

  • Background provisioning or managing networking for research computing environments (Kubernetes, SLURM, or HPC clusters)
  • Experience with NVIDIA Bluefield DPU programming and DOCA framework
  • Background with network function virtualization (NFV) and service function chaining
  • Knowledge of Kubernetes networking (CNI plugins, network policies, service mesh)
  • Experience building network control planes or SDN controllers
  • Familiarity with network automation frameworks and infrastructure-as-code for networking
  • Understanding of data center fabric architectures (spine-leaf, CLOS topologies)
  • Experience with network security and compliance requirements in regulated industries
  • Background building networking for research institutions, HPC environments, or cloud providers

Key Technologies

  • Go, Python, NVIDIA Bluefield DPUs, Open vSwitch, VXLAN, SR-IOV, RDMA, RoCE, InfiniBand, BGP, Linux networking, Terraform, FastAPI, gRPC

Why Moonlite

  • Build Next-Generation Infrastructure: Your work will create the platform foundation that enables financial institutions to harness AI capabilities previously impossible with traditional infrastructure.
  • Hands-On Ownership: As an early engineer, you'll have end-to-end ownership of projects and the autonomy to influence our product and technology direction.
  • Shape Industry Standards: Contribute to defining how enterprise AI infrastructure should work for the most demanding regulated environments.
  • Collaborate with Experts: Work alongside seasoned engineers and industry professionals passionate about high-performance computing, innovation, and problem-solving.
  • Start-Up Agility with Industry Impact: Enjoy the dynamic, fast-paced environment of a startup while making an immediate impact in an evolving and critical technology space.

We offer a competitive total compensation package combining a competitive base salary, startup equity, and industry-leading benefits. The total compensation range for this role is $165,000 – $225,000, which includes both base salary and equity. Actual compensation will be determined based on experience, skills, and market alignment. We provide generous benefits, including a 6% 401(k) match, fully covered health insurance premiums, and other comprehensive offerings to support your well-being and success as we grow together.

#li-remote

Similar jobs you might like

Technology

Nebius

Forward Deployment Engineering Manager

Senior

Remote

225,800 - 281,000 USD/yr

🏢 Summary: Lead and grow a Forward Deployed Engineering team delivering production-quality AI integrations, reference architectures, and proofs of concept for an AI cloud platform. The role combines people leadership, technical architecture oversight, partner engagement, and hands-on prototyping across agentic AI, inference, infrastructure, and data systems. 🗂️ Requirements: 8+ years of hands-on AI application, ML systems, or AI infrastructure engineering experience, 2+ years of engineering team leadership or management experience, Deep practical knowledge of LLM APIs, inference runtimes, orchestration frameworks, vector databases, RAG architectures, and agentic pipelines, Hands-on experience with agentic frameworks, Strong Python programming and end-to-end AI prototyping skills, Experience defining reference architectures and technical patterns, Experience building API and developer-platform integrations, Ability to work with external partner engineering teams and internal product and engineering teams, Strong technical communication skills, Authorization to work in the country of application 📃 Skills: Python, LangChain, LangGraph, CrewAI, AutoGen, LLM, RAG, FastAPI, Flask, Docker, Kubernetes, Git, AWS, GCP, Azure, vLLM, SGLang, TensorRT-LLM, Transformers, Qdrant, Weaviate, Milvus, pgvector, CUDA, TensorRT, NeMo 🏢 Description: About Nebius: Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. About Nebius Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure. Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI. Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D. The role Nebius builds the infrastructure serious AI teams run on - GPU clusters, inference runtimes, agent development environments, data pipelines - all of it purpose-built for the most demanding AI workloads. What we are now building is the ecosystem function that ensures the best AI companies choose to build on us, integrate with us, and stay. As Manager of Forward Deployed Engineering, you will lead a team of FDEs who sit at the intersection of solution architecture and hands-on engineering. You will set the technical bar, develop your team's craft, and ensure the work your team ships - integrations, reference architectures, and proofs of concept - is production-quality and strategically sound. You will also stay technically engaged: reviewing architectures, unblocking hard scoping problems, and occasionally prototyping yourself when the situation calls for it. You're welcome to work remotely in the United States. Your Responsibilities Will Include People & Team Leadership Hire, develop, and retain a team of Forward Deployed Engineers across ecosystem focus areas: agentic, inference, infrastructure, and data. Set clear expectations for technical quality and partner engagement; coach engineers toward those standards. Run structured 1:1s, provide direct and actionable feedback, and own the growth of each person on your team. Build a team culture where moving fast and building well are not in tension. Partner with Recruiting to define what great looks like for FDE roles and actively source candidates from the AI builder community. Technical Oversight & Architecture Review and elevate the technical work your team produces - integration architectures, proofs of concept, reference patterns, and partner scoping assessments. Serve as a technical escalation point for complex partner engagements; step in hands-on when needed. Maintain a high bar for what goes into the reference architecture library - not just what works, but what should be emulated. Stay current with the AI tooling ecosystem so you can guide your team's technical judgment, not just their output. Ecosystem Presence Represent Nebius at hackathons, in open source communities, and at technical events. Build in public - demos, reference architectures, and integrations that establish Nebius as the platform serious AI builders choose. Stay current with the AI tooling ecosystem - you know what shipped last week and what it means for our stack. Platform focus areas Depending on your background and mutual fit, you will focus on one or more of the following: Agentic - agent frameworks, memory systems, tool integration, orchestration, MCP, and guardrails Managed Inference - inference runtimes, model serving, optimization tooling, speculative decoding, and KV-cache routing IaaS / Managed Infrastructure - cloud-native integrations, GPU orchestration, and enterprise platform connectors Data - vector databases, retrieval systems, RAG architectures, data pipeline integrations, and synthetic data tooling Partner & Internal Stakeholder Engagement Engage directly with senior partner engineering leaders and founding CTOs; your team handles the working level, and you handle the strategic level. Translate what your team is seeing in the field into actionable product requirements for Nebius platform teams. Represent the FDE function in platform planning discussions as the technical voice of ecosystem integration. Work with ISV, SI, and field teams to scale solution adoption and drive revenue once integrations are ready. We expect you to have 8+ years of hands-on engineering experience in AI application development, ML systems, or AI infrastructure. 2+ years managing or leading a team of engineers, with a track record of developing technical talent. Deep working knowledge of the AI developer stack - LLM APIs, inference runtimes, orchestration frameworks, vector databases, RAG architectures, and agentic pipelines - built through shipping, not reading. Hands-on experience with agentic frameworks such as LangChain, LangGraph, CrewAI, AutoGen, or equivalent. Strong Python programming skills and comfort prototyping end-to-end AI systems quickly. Experience defining reference architectures and technical patterns - not just implementing them. Proven ability to move from idea to working prototype fast - you have shipped meaningful things under time pressure and found it energizing. Experience building integrations across APIs and developer platforms - you understand where the complexity actually lives. Comfort working across external partner engineering teams and internal Nebius Product and Engineering teams simultaneously. Strong technical communication - you can explain architecture decisions and integration findings to a founding CTO and a non-technical partner lead in the same day. It will be an added bonus if you have Experience with inference frameworks and optimization: vLLM, SGLang, TensorRT-LLM, speculative decoding, quantization, batching, and KV-cache routing. Familiarity with NVIDIA's software stack: CUDA, TensorRT, NeMo, or equivalent. Experience with multimodal AI models - vision-language, speech, or structured data. Won or placed at major AI hackathons in the past 12 months. Worked as a developer advocate, solutions engineer, or technical partner manager at a leading AI platform or developer tooling company. Been an early engineer at a YC-backed AI startup - you built the product under real constraints. Open source projects or public demos with meaningful community adoption. Proficiency with DevOps tools: Docker, Kubernetes, and Git. Preferred Technical Stack Languages - Python ML frameworks - vLLM, SGLang, TensorRT-LLM, Transformers, and OpenAI / Anthropic SDKs Agentic frameworks - LangChain, LangGraph, CrewAI, AutoGen, smolagents, or equivalent Vector databases - Qdrant, Weaviate, Milvus, and pgvector API and web frameworks - FastAPI and Flask DevOps - Kubernetes, Docker, and Git Cloud platforms - AWS, GCP, and Azure Key Employee Benefits Health Insurance: 100% company-paid medical, dental, and vision coverage for employees and families. 401(k) Plan: Up to 4% company match with immediate vesting. Parental Leave: 20 weeks paid for primary caregivers, 12 weeks for secondary caregivers. Remote Work Reimbursement: Up to $85/month for mobile and internet. Disability & Life Insurance: Company-paid short-term, long-term, and life insurance coverage. Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Starting Base Compensation Range: $225,800 - $281,000 USD Benefits & Perks Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's it like to work at Nebius Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know Pay Transparency We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law. Base Compensation Range $225—$280 USD Benefits & Perks: Competitive compensation Career growth and learning opportunities Flexibility and ownership Collaborative and innovative culture Opportunity to work on impactful AI projects International environment and talented teams What's It Like To Work At Nebius: Fast moving - Bold thinking - Constant growth - Meaningful impact - Trust and real ownership - Opportunity to shape the future of AI Equal Opportunity Statement: Nebius is an equal opportunity employer. We are committed to fostering an inclusive and diverse workplace and to providing equal employment opportunities in all aspects of employment. We do not discriminate on the basis of race, color, religion, sex (including pregnancy), national origin, ancestry, age, disability, genetic information, marital status, veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by applicable law. Applicants must be authorized to work in the country in which they apply and will be required to provide proof of employment eligibility as a condition of hire. If you need accommodations during the application process, please let us know.

Technology

Databricks

Staff Partner Engineer, Azure

Senior

On-site

New York City, NY , +2

206,450 - 206,450 USD/yr

🏢 Summary: Partner Engineer role owning the technical relationship and joint product roadmap between Databricks and Microsoft Azure. The position drives cross-company engineering integrations, resolves dependencies, develops reference architectures, and enables field teams on Azure–Databricks solutions. 🗂️ Requirements: 10+ years in cloud solution architecture, partner engineering, or technical program management, Deep expertise in Microsoft Azure services and architecture patterns, Experience managing strategic technical partnerships with VP/C-level stakeholders, Hands-on ability to prototype integrations, debug complex issues, and validate technical feasibility, Experience with Python, SQL, infrastructure-as-code, and API integrations, Ability to manage complex multi-company engineering initiatives and drive consensus, Deep knowledge of the Databricks platform and Microsoft data/AI ecosystem, Strong technical communication and influence across product, engineering, and partner teams 📃 Skills: Azure, Databricks, Python, SQL, APIs, Infrastructure-as-code 🏢 Description: RDQ427R371 At Databricks , we are passionate about enabling data teams to solve the world’s toughest problems — from making the next mode of transportation a reality to accelerating the development of medical breakthroughs. We do this by building and running the world’s best data and AI infrastructure platform so our customers can use deep data insights to improve their business. Founded by engineers — and customer obsessed — we leap at every opportunity to tackle technical challenges, from designing next-gen UI/UX for interfacing with data to scaling our services and infrastructure across millions of virtual machines. And we're only getting started. About the Role We're seeking a Partner Engineer to own the technical relationship between Databricks Product/Engineering and Microsoft Azure. You will be the primary technical liaison managing our cross product roadmap, driving engineering integrations, and ensuring Databricks products work seamlessly within the Azure ecosystem. You will work directly with Azure Partner SAs, Product Managers, and engineering teams to align roadmaps, identify integration gaps, and accelerate feature delivery. This role reports to the Director of Partner Engineering and works closely with Databricks Product Management, Engineering, and Field teams. Impact you will have Own the Databricks-Azure product-to-cloud technical roadmap, serving as the primary interface between Databricks Product Management and Azure partner teams to ensure alignment on features, integrations, and co-innovation priorities. Track and drive resolution of cross-company dependencies and blockers, manage joint milestone timing with Azure teams, and own the coordination rhythms (recurring syncs, escalation paths) that keep both sides aligned through delivery. Drive engineering integration work between Databricks and Azure product teams coordinating cross-company technical initiatives from scoping through delivery. Create and maintain technical architecture artifacts showing "better together" integration patterns, holistic architecture diagrams, and reference implementations for Azure + Databricks solutions. Enable field teams with technical guidance on Azure-Databricks integration patterns, working with Solution Architects and Customer Success to accelerate Azure customer wins. What we’re looking for 10+ years of experience in cloud solution architecture, partner engineering, or technical program management roles, with deep expertise in Microsoft Azure services and architecture patterns. Proven experience managing strategic technical partnerships at VP/C-level, with ability to navigate complex multi-company engineering initiatives and drive consensus across organizations. Hands-on technical skills with ability to prototype integrations, debug complex issues, and validate technical feasibility of proposed solutions. Experience with Python, SQL, infrastructure-as-code, and API integration patterns. Demonstrated ability to manage high-stakes partner relationships where strategic interests diverge, including setting boundaries and managing expectations while preserving long-term partnership health. Deep knowledge of Databricks platform and Microsoft data/AI ecosystem Excellent communication and influence skills, with ability to translate technical requirements between Product Management, Engineering, and external partner teams. Preferred Prior experience in Partner Solutions Architect or Alliance Engineering roles at cloud providers (AWS, Azure, GCP) or major ISVs. Background working at or with Microsoft - understanding of Microsoft field organization, partner program structure, and decision-making processes. History of driving product-to-product integrations between major platforms. Familiarity with modern development workflows and AI-assisted development tools (Cursor, Claude Code, Codex). Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here . Zone 1 Pay Range $150,200 — $206,450 USD Pay Range Transparency Databricks is committed to fair and equitable compensation practices. The pay range(s) for this role is listed below and represents the expected salary range for non-commissionable roles or on-target earnings for commissionable roles. Actual compensation packages are based on several factors that are unique to each candidate, including but not limited to job-related skills, depth of experience, relevant certifications and training, and specific work location. Based on the factors above, Databricks anticipates utilizing the full width of the range. The total compensation package for this position may also include eligibility for annual performance bonus, equity, and the benefits listed above. For more information regarding which range your location is in visit our page here . Zone 2 Pay Range $150,200 — $206,450 USD About Databricks Databricks is the Data and AI company. More than 20,000 organizations worldwide — including adidas, AT&T, Bayer, Block, Mastercard, Rivian, Unilever, and 70% of the Fortune 500 — rely on the Databricks Data + AI Platform to build and scale data and AI apps, analytics and agents. Headquartered in San Francisco with 30+ offices around the globe, Databricks offers a unified platform that includes Genie, Lakebase, Agent Bricks, Lakeflow, Lakehouse, and Unity Catalog. To learn more, follow Databricks on LinkedIn , X , YouTube , and Instagram . Benefits At Databricks, we strive to provide comprehensive benefits and perks that meet the needs of all of our employees. For specific details on the benefits offered in your region click here . Our Commitment to Diversity and Inclusion At Databricks, we are committed to fostering a diverse and inclusive culture where everyone can excel. We take great care to ensure that our hiring practices are inclusive and meet equal employment opportunity standards. Individuals looking for employment at Databricks are considered without regard to age, color, disability, ethnicity, family or marital status, gender identity or expression, language, national origin, physical and mental ability, political affiliation, race, religion, sexual orientation, socio-economic status, veteran status, and other protected characteristics. Compliance If access to export-controlled technology or source code is required for performance of job duties, it is within Employer's discretion whether to apply for a U.S. government license for such positions, and Employer may decline to proceed with an applicant on this basis alone. Applicant Privacy Notice