Zoox Logo

Zoox

Senior ML Storage Infrastructure Engineer

Reposted 2 Days Ago
Hybrid
2 Locations
192K-300K Annually
Senior level
Hybrid
2 Locations
192K-300K Annually
Senior level
Responsible for improving Zoox's HPC infrastructure, optimizing cloud/storage systems for ML workloads, designing APIs, and investigating distributed paradigms.
The summary above was generated by AI
Zoox is looking for a software engineer to work on our custom High-Performance Computing infrastructure and its supporting ecosystem of tools and services. This infrastructure is central to machine learning workflows across all Zoox software divisions, from data engineering to computer vision perception to simulation and more. You will take on a breadth of end-to-end responsibilities including distributed system design, algorithmic job scheduling, and adaptive cloud scaling in support of all of Zoox’s computational needs.

In this role, you will:

  • Design, build, and optimize a petabyte-scale, in-house HPC storage infrastructure, ensuring high performance and reliability for our machine learning workloads across both cloud and on-premise data centers.
  • Drive GPU efficiency by strategically collocating storage and compute, architecting a storage layer that keeps tens of thousands of GPUs fully utilized and prevents bottlenecks.
  • Drive key initiatives in training and storage optimization by partnering with ML practitioners, applying your deep understanding of frameworks such as PyTorch and TensorFlow to meet their evolving demands. 
  • Investigate and adopt new distributed system paradigms and cutting-edge technologies to ensure our infrastructure can scale to meet ever-growing computational and storage demands.
  • Create production-grade web service APIs, SDKs, and other essential tools to deliver a world-class developer experience for all software teams at Zoox.

Qualifications

  • Experience designing and building high-performance, distributed storage systems (object/file) for large-scale, GPU-bound workloads.
  • Proficiency in Python, Java, or similar languages for developing data-intensive, high-performance applications.
  • Hands-on experience with cloud platforms (AWS, GCP, Azure), using their storage, GPU, and observability services to provide usage showback for ML practitioners.
  • Bachelor's degree in Computer Science or a related field with a strong foundation in data structures and systems design.

Bonus Qualifications

  • Experience with parallel filesystems (e.g., Lustre, FSx) and their integration with container orchestrators via Kubernetes CSI drivers.
  • Deep knowledge of ML frameworks like PyTorch and TensorFlow, and workload schedulers such as SLURM or Kubernetes.
  • Familiarity with emerging AI paradigms, including agentic systems, and observability tools like OpenTelemetry.

About Zoox
Zoox is developing the first ground-up, fully autonomous vehicle fleet and the supporting ecosystem required to bring this technology to market. Sitting at the intersection of robotics, machine learning, and design, Zoox aims to provide the next generation of mobility-as-a-service in urban environments. We’re looking for top talent that shares our passion and wants to be part of a fast-moving and highly execution-oriented team.

Follow us on LinkedIn

Accommodations
If you need an accommodation to participate in the application or interview process please reach out to [email protected] or your assigned recruiter.

A Final Note:
You do not need to match every listed expectation to apply for this position. Here at Zoox, we know that diverse perspectives foster the innovation we need to be successful, and we are committed to building a team that encompasses a variety of backgrounds, experiences, and skills.

Top Skills

AWS
Azure
GCP
Java
Python
Slurm

Zoox Seattle, Washington, USA Office

1111 3rd Ave, Seattle, WA, United States, 98101

Similar Jobs

54 Minutes Ago
In-Office
Seattle, WA, USA
69K-93K Annually
Junior
69K-93K Annually
Junior
Aerospace • Information Technology • Cybersecurity • Defense • Manufacturing
The Provisioning Support Specialist assists engineers in producing data files for spare parts, reviews engineering documentation, and communicates with airline customers regarding provisioning products.
Top Skills: Adobe AcrobatMicrosoft AccessExcelMicrosoft PowerpointMicrosoft Word
59 Minutes Ago
Easy Apply
Hybrid
5 Locations
Easy Apply
213K-278K Annually
Senior level
213K-278K Annually
Senior level
Fintech • HR Tech
Lead and grow Gusto's engineering team for design systems, driving technical roadmaps, team scaling, and ensuring quality across products.
Top Skills: Accessibility ToolingCss-In-JsReactTypescript
2 Hours Ago
In-Office
2 Locations
77K-104K Annually
Entry level
77K-104K Annually
Entry level
Aerospace • Information Technology • Cybersecurity • Defense • Manufacturing
Assist in developing production methodologies, work within teams to improve efficiency, and gain hands-on experience in manufacturing engineering.
Top Skills: EngineeringLean PrinciplesManufacturing Technology

What you need to know about the Seattle Tech Scene

Home to tech titans like Microsoft and Amazon, Seattle punches far above its weight in innovation. But its surrounding mountains, sprinkled with world-famous hiking trails and climbing routes, make the city a destination for outdoorsy types as well. Established as a logging town before shifting to shipbuilding and logistics, the Emerald City is now known for its contributions to aerospace, software, biotech and cloud computing. And its status as a thriving tech ecosystem is attracting out-of-town companies looking to establish new tech and engineering hubs.

Key Facts About Seattle Tech

  • Number of Tech Workers: 287,000; 13% of overall workforce (2024 CompTIA survey)
  • Major Tech Employers: Amazon, Microsoft, Meta, Google
  • Key Industries: Artificial intelligence, cloud computing, software, biotechnology, game development
  • Funding Landscape: $3.1 billion in venture capital funding in 2024 (Pitchbook)
  • Notable Investors: Madrona, Fuse, Tola, Maveron
  • Research Centers and Universities: University of Washington, Seattle University, Seattle Pacific University, Allen Institute for Brain Science, Bill & Melinda Gates Foundation, Seattle Children’s Research Institute

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account