Skip to content
AI Engineer Jobs
DigitalOcean

Principal Engineer, Inference Memory and Storage Systems

DigitalOcean

Location
Hybrid (Seattle, Washington)
Compensation
$249k - $312k/yr
Employment
Full-time
Level
Senior Level
Posted 2 days ago

About the Role

DigitalOcean is building the Inference Cloud, a serving stack for running frontier open models on GPU fleets at production scale. This role owns the technical vision for the memory and storage layer, optimizing cost, latency, and throughput for large-scale LLM inference.

Skills

Distributed systems Memory hierarchies GPU HBM NVMe RDMA NVLink LLM inference C++ Go Rust Python Kubernetes System architecture Performance optimization KV cache management Networked I/O

Benefits

  • Employee assistance program

Perks

  • Equity compensation
  • Flexible time off
  • Hybrid Work
  • Conference reimbursement
  • Training and education reimbursement
  • Employee stock purchase program

Full job details