Skip to content
AI Engineer Jobs
StackYak

Senior AI Inference Engineer

StackYak

Location
Remote (United States)
Employment
Full-time
Level
Senior Level
Posted 1 week ago

About the Role

StackYak is an early-stage startup building the infrastructure layer for AI, unifying compute, GPU infrastructure, networking, and inference into a single product. This founding role involves owning the inference infrastructure, optimizing model serving, and automating deployment decisions for production workloads.

Skills

Large language models GPU infrastructure Inference optimization Python Distributed systems Quantization Tensor parallelism Pipeline parallelism Serving runtimes vLLM SGLang TensorRT-LLM Benchmarking Linux Capacity planning Multi-node serving

Perks

  • Remote Work
  • Meaningful equity

Full job details