Staff Software Engineer, Networking (Seattle or SF Bay Area)

Staff Engineer · Staff · Full Time · Remote

San Francisco Bay Area · RemoteUSD 170k – 276k1d ago
Apply for this role

Opens Docker's application page

Role

What you'll do.

Docker seeks a Staff Software Engineer to lead the networking infrastructure powering the Sandboxes platform, supporting AI and security initiatives. This role demands expertise in large-scale distributed systems, kernel-level networking technologies, and container orchestration, requiring 8+ years of infrastructure experience. The position provides technical leadership to a small engineering team while architecting high-performance, scalable networking solutions across local and cloud environments.

Responsibilities

  • Architect Core Networking Services: Design and implement foundational networking services that power Docker's Sandboxes platform across diverse environments including local laptops and public cloud infrastructure (AWS, Azure, GCP, OCI). Establish simple, reusable primitives that work seamlessly in both local and cloud deployments to reduce architectural complexity.
  • Build Scalable MicroVM Infrastructure: Develop highly scalable networking infrastructure supporting microVM orchestration, workload scheduling, and lifecycle management. Design systems capable of handling multi-tenant deployments with efficient resource allocation and predictable performance characteristics.
  • Develop High-Performance Networking Components: Engineer optimized data-plane and control-plane networking components specifically designed for multi-tenant agentic workloads. Focus on achieving sub-millisecond latency, high throughput, and minimal resource overhead in distributed environments.
  • Ensure System Reliability and Observability: Establish comprehensive monitoring, logging, and observability practices for network infrastructure. Design for high availability with redundancy patterns, implement robust alerting systems, and conduct performance analysis across Docker's Sandbox infrastructure.
  • Provide Technical Leadership: Lead a small team of engineers by setting technical direction, conducting design reviews, establishing coding standards, and mentoring team members. Drive consensus across stakeholders, communicate architectural decisions clearly, and foster a culture of operational excellence.
  • Cross-Functional Collaboration: Partner closely with product, platform, and security teams to translate requirements into robust networking capabilities. Participate in architectural discussions, design review processes, and ensure networking decisions align with broader Docker platform goals.
  • Production Operations and Incident Response: Participate in on-call rotation, respond to production incidents, debug issues across distributed systems in both local and cloud environments. Implement improvements to prevent incident recurrence and establish post-incident review processes.
  • DevOps and Infrastructure Excellence: Contribute to CI/CD pipeline enhancements, infrastructure-as-code improvements, and deployment automation. Establish best practices for infrastructure testing, validation, and progressive rollout strategies across the platform.

Qualifications

What we look for.

Technical

  • Large-Scale Distributed Systems

    8+ years designing, building, and operating distributed systems at scale with focus on networking architecture and infrastructure foundations. Proven ability to architect systems handling millions of concurrent connections and petabytes of data.

  • Kernel and User-Space Networking

    Deep expertise in Linux networking stack, including advanced kernel networking, Open vSwitch (OVS), Open Virtual Network (OVN), DPDK for high-performance packet processing, and eBPF for programmable network behavior. Experience optimizing network performance at kernel level.

  • Container Orchestration and Networking

    Extensive experience with container orchestration platforms (Kubernetes, Docker Swarm), container networking models, network policies, and service discovery mechanisms. Understanding of container runtime networking constraints and optimization techniques.

  • Service Mesh and Policy Enforcement

    Hands-on experience implementing or operating service mesh architectures (Istio, Linkerd, or similar). Knowledge of network policy enforcement, traffic management, circuit breaking, and observability within service mesh frameworks.

  • Cloud Infrastructure and Scalability

    Strong understanding of cloud platforms (AWS, Azure, GCP, OCI) including VPC design, security groups, load balancing, and scalability patterns. Experience architecting solutions that work across multiple cloud providers or hybrid environments.

  • Infrastructure-as-Code and Automation

    Proficiency with infrastructure-as-code tools (Terraform, CloudFormation, Ansible) and CI/CD pipeline construction. Experience implementing automated deployment, testing, and rollback strategies for infrastructure changes.

  • Distributed Systems Debugging

    Expert-level debugging and troubleshooting skills in distributed environments. Proficiency with packet analysis tools, system tracing, performance profiling, and root cause analysis of complex issues spanning multiple layers.

Education

  • Bachelor's Degree in Computer Science or Related Field

    Bachelor's degree in Computer Science, Software Engineering, Computer Engineering, or related technical field. Equivalent practical experience demonstrating mastery of core computer science fundamentals is acceptable.

Experience

  • Senior Infrastructure Engineering Leadership

    8+ years of combined experience building large-scale cloud or distributed systems with proven track record of technical leadership. Demonstrated success leading small teams through complex architectural decisions and delivering high-quality infrastructure software on tight timelines.

  • Technical Consensus Building

    Proven ability to drive technical consensus among engineering teams and stakeholders on architecture, design decisions, and technology choices. Experience navigating ambiguity, proposing pragmatic solutions, and rallying teams around technical direction.

  • High-Availability Systems Operation

    Experience designing, deploying, and operating highly available networking infrastructure serving millions of users or handling massive scale. Track record of achieving strong SLOs and reducing mean time to recovery (MTTR) for critical systems.

  • Cross-Platform Development

    Experience developing infrastructure solutions that work across multiple platforms, environments, or cloud providers. Understanding of platform-specific constraints and techniques for building abstraction layers that hide platform differences.

Skills

Required

  • Systems Architecture

    Ability to design systems from first principles, evaluate tradeoffs between reliability, performance, scalability, and complexity. Experience with distributed consensus, eventual consistency, and failure domain isolation.

  • Network Protocol Implementation

    Deep understanding of TCP/IP stack, DNS, HTTP/gRPC protocols, and ability to implement or extend networking protocols. Knowledge of performance implications of protocol choices and optimization techniques.

  • Performance Optimization

    Expertise in identifying performance bottlenecks through profiling and measurement, implementing optimizations at multiple levels (algorithm, architecture, kernel, hardware). Experience with latency and throughput tradeoff analysis.

  • Team Leadership and Mentoring

    Proven ability to guide engineering teams technically, conduct effective code reviews, establish development standards, and mentor engineers toward technical growth. Experience growing teams from individual contributors to independent problem-solvers.

  • Problem Solving and Critical Thinking

    Strong analytical skills combined with practical engineering judgment. Ability to break complex problems into manageable components, evaluate multiple approaches, and select solutions balancing technical elegance with pragmatic delivery.

  • Communication and Documentation

    Clear written and verbal communication skills for explaining complex technical concepts to diverse audiences. Proficiency creating architecture documents, design specifications, and runbooks that guide team decision-making.

Preferred

  • Golang Programming

    Nice to have

    Experience building high-performance systems in Go language. Familiarity with goroutines, channels, and Go's standard library for systems programming.

  • Rust Systems Programming

    Nice to have

    Experience implementing systems-level code in Rust, particularly for networking or performance-critical components. Understanding of Rust's ownership model and how it ensures memory safety in infrastructure code.

  • Container and Virtualization Technologies

    Nice to have

    Hands-on experience with microVM technologies, hypervisors, or container runtimes. Understanding of containerization tradeoffs and isolation mechanisms at kernel level.

  • AI/ML Infrastructure

    Nice to have

    Experience supporting AI workloads or ML inference platforms. Understanding of unique networking requirements for distributed AI training or real-time inference serving.

  • Open Source Contributions

    Nice to have

    Active contributions to open source networking projects, container runtimes, or infrastructure software. Demonstrated ability to collaborate in open source communities and navigate complex existing codebases.

  • Security and Network Isolation

    Nice to have

    Expertise in network security, encryption, access control, and tenant isolation. Experience implementing zero-trust networking models or fine-grained security policies.

Tech stack

Languages

GoCPython

Frameworks

KubernetesOpen vSwitch (OVS)Open Virtual Network (OVN)gRPC

Databases

etcd

Tools

TerraformDocker and Container RuntimesLinux Kernel TracingPrometheus and Grafanatcpdump and Wireshark

Other

eBPF (Extended Berkeley Packet Filter)DPDK (Data Plane Development Kit)Linux Networking Stack

Compensation

Pay and benefits.

Base·USD 170,350 – 275,550

Equity·Stock options

Benefits

  • Flexible Work Arrangement

    Remote-first work culture with offices in Seattle and San Francisco Bay Area. Design your work schedule to fit your lifestyle and maximize productivity.

  • Quarterly Whaleness Days

    Designated quarterly breaks plus end-of-year extended break to recharge, reflect, and return energized. Docker recognizes the importance of sustainable work practices.

  • Comprehensive Home Office Setup

    Home office setup allowance and equipment support to ensure you maintain an ergonomic, productive workspace. Technology stipend of $100 USD net/month for ongoing upgrades.

  • Generous Parental Leave

    16 weeks of paid parental leave available after 6 months of employment, supporting work-life balance during major life transitions.

  • Professional Development

    Annual training stipend for conferences, courses, certifications, and classes. Investment in your continuous learning and career growth in rapidly evolving infrastructure landscape.

  • Flexible Time Off

    Unlimited PTO policy encouraging you to take the time you need for personal wellness, travel, and rejuvenation outside of mandatory holidays.

  • Equity and Ownership

    Equity participation in Docker's growth as a rapidly expanding company. Align your interests with company success and benefit from long-term value creation.

  • Comprehensive Health and Retirement

    Medical benefits, dental, vision, and retirement plans tailored by country of employment. Full support for your health and financial security.

  • Community and Swag

    Docker branded merchandise and access to vibrant global community of developers and engineers building modern infrastructure.

Full posting

Original listing.

Docker has been one of the most loved brands in developer tooling, trusted by more than 20 million monthly users and over 20 billion container image pulls. From solo founders to the world's largest companies, developers rely on Docker to build, share, and run their applications across our suite of products including Docker Desktop, Docker Hub, and Docker Scout.

We are a globally distributed, remote-first team building the tools that define how software gets built and delivered. As AI agents redefine software development, Docker is at the center of that shift, providing the sandboxed environments, verified images, and secure infrastructure that make autonomous workflows trustworthy by default.

As a Staff Software Engineer in Networking, you will set the technical direction for the networking stack behind Docker's Sandboxes platform and make architectural decisions that ripple across Docker's product portfolio. Your work will directly enable Docker's expansion into AI and security products.

This role calls for someone who:

  1. Combines deep technical expertise in infrastructure at scale, virtual networking, service mesh, orchestration, and integration across the broader infrastructure stack;

  2. Communicates and collaborates well, thinks critically, and turns ideas into working systems

  3. Adapts to ambiguity and brings clarity

  4. Provides technical leadership to a small team of engineers, driving customer delight through the solutions they deliver

 

Responsibilities

  • Design, implement, and operate core networking services that power Docker's Sandboxes platform across local laptops and public cloud environments

  • Build scalable networking infrastructure for microVM orchestration, workload scheduling, and lifecycle management

  • Develop high-performance networking data- and control-plane components for multi-tenant agentic workloads

  • Ensure network reliability, observability, and performance across Docker's Sandbox infrastructure

  • Collaborate with product, platform, and security teams to deliver customer-focused capabilities

  • Participate in architectural discussions, code reviews, and design documents

  • Contribute to automation and CI/CD improvements across the deployment pipeline

  • Debug and resolve production issues across distributed systems in local and public-cloud environments

  • Take part in on-call rotation for your team; respond to incidents and drive continuous improvement of system reliability and availability

 

Qualifications

  • 8+ years of experience building and leading large-scale cloud or distributed systems, with a focus on networking foundations

  • Track record of driving technical consensus and delivering high-quality infrastructure software at scale, on tight timelines

  • Experience with microservices architecture, container networking, service mesh, policy enforcement, and container orchestration

  • Experience with kernel and user-space networking technologies, including Linux networking, OVS/OVN, DPDK, and eBPF

  • Experience designing and operating highly available, secure, and observable networking infrastructure

  • Strong understanding of cloud infrastructure (AWS, Azure, GCP, OCI) and related scalability patterns

  • Familiarity with CI/CD pipelines, monitoring, and infrastructure-as-code tooling

  • Excellent problem-solving and debugging skills in distributed environments

  • Strong communication skills and ability to collaborate across remote, cross-functional teams

  • Bachelor's degree in Computer Science, Engineering, or a related field, or equivalent practical experience

 

What to Expect

First 30 days

  • Onboard, integrate with our Sandbox infrastructure teams, and build context on Docker's Sandbox vision

  • Deep-dive into Sandbox architecture for local and public-cloud infrastructure, with a focus on driving simple, common foundational primitives usable across both environments

  • Identify the highest-leverage opportunities in the current architecture, along with strategic and tactical execution risks

First 90 days

  • Develop a technical roadmap to address the challenges identified; drive consensus across stakeholders

  • Provide technical leadership to a small team of engineers to deliver on the roadmap

  • Establish or improve technical standards for development, deployment, operation, and observability of networking services

One-year outlook

  • Drive simplification of foundational networking infrastructure across local and public-cloud environments

  • Establish repeatable engineering practices for quality, performance, and operational excellence

  • Nurture the team's confidence and capability by providing clarity, sustaining focus, and delivering pragmatically through successive iteration

  • Set a standard for what a small team of engineers can deliver and manage when given a charter, autonomy, and the tools they need

 

Docker considers visa sponsorship on a case-by-case basis based on business needs.

 

Compensation & Equity

United States: $170,350 – $275,550 + equity

Perks

  • Freedom & flexibility; fit your work around your life

  • Designated quarterly Whaleness Days plus end of year Whaleness break

  • Home office setup; we want you comfortable while you work

  • 16 weeks of paid Parental leave (after 6 months of employment)

  • Technology stipend equivalent to $100 USD net/month

  • PTO plan that encourages you to take time to do the things you enjoy

  • Training stipend for conferences, courses and classes

  • Equity; we are a growing start-up and want all employees to have a share in the success of the company

  • Docker Swag

  • Medical benefits, retirement and holidays vary by country

  • Remote-first culture, with offices in Seattle and Paris

Docker embraces diversity and equal opportunity. We are committed to building a team that represents a variety of backgrounds, perspectives, and skills. The more inclusive we are, the better our company will be.

#LI-REMOTE

Redirects to Docker's application page.

Other roles

More at Docker.

View all 24 roles